AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,479 results
Hardware

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache

DGX agent

arXiv:2604.06370v1 Announce Type: cross Abstract: The serving paradigm of large language models (LLMs) is rapidly shifting towards complex multi-agent workflows where specialized agents collaborate ov

hardwarearxiv-cs-lg
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start

DGX agent

arXiv:2604.06664v1 Announce Type: cross Abstract: Modern LLM service providers increasingly rely on autoscaling and parallelism reconfiguration to respond to rapidly changing workloads, but cold-start

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

Full State-Space Visualisation of the 8-Puzzle: Feasibility, Design, and Educational Use

DGX agent

arXiv:2604.06186v1 Announce Type: cross Abstract: Search algorithms are a foundational topic in artificial intelligence education, yet even simple domains can generate large state spaces that challeng

hardwarearxiv-cs-ai
10 Apr 2026
Hardware

Graph Neural ODE Digital Twins for Control-Oriented Reactor Thermal-Hydraulic Forecasting Under Partial Observability

DGX agent

arXiv:2604.07292v1 Announce Type: new Abstract: Real-time supervisory control of advanced reactors requires accurate forecasting of plant-wide thermal-hydraulic states, including locations where physi

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

DGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

hardwarearxiv-cs-cv
10 Apr 2026
Hardware

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of …

DGX agent

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of Open Source, will show you how to build with it. Live worksh

hardwarejerry-liu--x
10 Apr 2026
Hardware

Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning

DGX agent

arXiv:2604.07345v1 Announce Type: cross Abstract: The rapid growth of generative artificial intelligence (AI) has introduced unprecedented computational demands, driving significant increases in the e

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

National Robotics Week — Latest Physical AI Research, Breakthroughs and Resources

DGX agent

This National Robotics Week, NVIDIA is highlighting the breakthroughs that are bringing AI into the physical world — as well as the growing wave of robots transforming industries, from agricultural an

hardwarenvidia-blog
10 Apr 2026
Hardware

NestPipe: Large-Scale Recommendation Training on 1,500+ Accelerators via Nested Pipelining

DGX agent

arXiv:2604.06956v1 Announce Type: cross Abstract: Modern recommendation models have increased to trillions of parameters. As cluster scales expand to O(1k), distributed training bottlenecks shift from

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

Object-Centric Stereo Ranging for Autonomous Driving: From Dense Disparity to Census-Based Template Matching

DGX agent

arXiv:2604.07980v1 Announce Type: new Abstract: Accurate depth estimation is critical for autonomous driving perception systems, particularly for long range vehicle detection on highways. Traditional

hardwarearxiv-cs-cv
10 Apr 2026
Hardware

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and …

DGX agent

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and removing unnecessary constraints so developers can truly exp

hardwarezhipu-ai--x
10 Apr 2026
Hardware

OpenPRC: A Unified Open-Source Framework for Physics-to-Task Evaluation in Physical Reservoir Computing

DGX agent

arXiv:2604.07423v1 Announce Type: new Abstract: Physical Reservoir Computing (PRC) leverages the intrinsic nonlinear dynamics of physical substrates, mechanical, optical, spintronic, and beyond, as fi

hardwarearxiv-cs-ro
10 Apr 2026
Hardware

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing

DGX agent

arXiv:2512.06774v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a leading representation for high-fidelity 3D assets, yet protecting these assets via digital watermarking r

hardwarearxiv-cs-cv
10 Apr 2026
Hardware

RISC-V chip design startup SiFive nabs $400M investment

DGX agent

SiFive Inc., a startup that sells chip designs based on the open-source RISC-V architecture, has raised $400 million in funding at a $3.65 billion valuation. Atreides Management led the Series G round

hardwaresiliconangle
10 Apr 2026
Hardware

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

DGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

Splats under Pressure: Exploring Performance-Energy Trade-offs in Real-Time 3D Gaussian Splatting under Constrained GPU Budgets

DGX agent

arXiv:2604.07177v1 Announce Type: cross Abstract: We investigate the feasibility of real-time 3D Gaussian Splatting (3DGS) rasterisation on edge clients with varying Gaussian splat counts and GPU comp

hardwarearxiv-cs-lg
10 Apr 2026
Hardware

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us curren…

DGX agent

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us currently. SemiAnalysis employees consume 1.86B tokens per month.

hardwaredylan-patel--x
10 Apr 2026
Hardware

The power of JAX

DGX agent

The power of JAX Introducing gyaradax 🐉: A JAX solver for local flux-tube gyrokinetics with custom CUDA kernels for acceleration. This entire code was vibecoded by @ggalletti_ and me in a month. Valid

hardwarefrancois-chollet--x
10 Apr 2026
Hardware

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training

DGX agent

arXiv:2603.04385v3 Announce Type: replace Abstract: Feed-forward transformer models have driven rapid progress in 3D vision, but state-of-the-art methods such as VGGT and pi^3 have a computational

hardwarearxiv-cs-cv
10 Apr 2026
Hardware

How Estée Lauder Companies uses Cloud Run worker pools for its pull-based agentic workloads

DGX agent

Cloud Run has long provided developers with a straightforward, opinionated platform for running code. You can easily deploy request-driven web applications using Cloud Run services, or execute run-to-

hardwaregoogle-cloud-ai
9 Apr 2026
Hardware

New course: Efficient Inference with SGLang: Text and Image Generation, built in partnership with LMSys @lmsysorg and RadixArk @radixark, an…

DGX agent

New course: Efficient Inference with SGLang: Text and Image Generation, built in partnership with LMSys @lmsysorg and RadixArk @radixark, and taught by Richard Chen @richardczl, a Member of Technical

hardwareandrew-ng--x
9 Apr 2026
Hardware

Nvidia published DWDP (Distributed Weight-Data Parallelism), a new inference parallelism strategy focused on prefill. It sounds slightly ins…

DGX agent

Nvidia published DWDP (Distributed Weight-Data Parallelism), a new inference parallelism strategy focused on prefill. It sounds slightly insane until you remember the target machine is GB200 NVL72. Th

hardwaredylan-patel--x
9 Apr 2026
Hardware

Strength and Destiny Collide: ‘Samson: A Tyndalston Story’ Arrives in the Cloud

DGX agent

A timeless story of grit, faith and rebellion takes center stage as Samson: A Tyndalston Story joins the GeForce NOW library today. The highly anticipated release from Liquid Swords can now be streame

hardwarenvidia-blog
9 Apr 2026
Hardware

The sad thing is that Ro is the guy preaching for socialism while he is the most active insider trader in Congress. He is a terrible represe…

DGX agent

The sad thing is that Ro is the guy preaching for socialism while he is the most active insider trader in Congress. He is a terrible representative of Silicon Valley. Nancy Pelosi take a seat. There i

hardwareelon-musk--x
9 Apr 2026
Hardware

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at…

DGX agent

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at least some strategy, since the robot first has to spread th

hardwareclem-delangue--x
9 Apr 2026
Hardware

This is literally why I wrote Taming Silicon Valley.

DGX agent

This is literally why I wrote Taming Silicon Valley. Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government funded

hardwaregary-marcus--x
8 Apr 2026
Hardware

We're seeing even more autonomous AI coworkers. The new MLE agent on the market is Disarray. In Kaggle competitions, Disarray: - won 28 meda…

DGX agent

We're seeing even more autonomous AI coworkers. The new MLE agent on the market is Disarray. In Kaggle competitions, Disarray: - won 28 medals across diverse domains (vision, NLP, tabular data) - plac

hardwareallie-k--miller--x
8 Apr 2026
Model Releases

ggml-cpu/ops: vectorize flash-attention V-cache F16 to F32 conversion by jinzihao · Pull Request #26947 · ggml-org/llama.cpp

DGX agent

Overview ggml_cpu_fp16_to_fp32 leverages hardware F16C intrinsics (AVX-512, AVX2, etc.), faster than the software-only ggml_fp16_to_fp32_row, bringing 17-31% gain in prompt processing rate for a small

model-releasesr-localllama
13 Aug 2026
Applications

Hybrid-LUT: Channel-Aware Hybrid Lookup Table and Filtering for Efficient Image Denoising

DGX agent

arXiv:2608.11646v1 Announce Type: new Abstract: Lookup table (LUT)-based image denoising methods have attracted increasing attention due to their high efficiency and hardware-friendly properties. Howe

applicationsarxiv-cs-cv
13 Aug 2026
Safety

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

DGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

safetyarxiv-cs-cl
13 Aug 2026
Model Releases

Google launches five new Pixel devices, array of Gemini Intelligence features

DGX agent

Google LLC today expanded its Pixel consumer hardware line with a foldable handset, three smartphones and a smartwatch. The devices will ship with an upgraded version of the company’s Gemini Intellige

model-releasessiliconangle
12 Aug 2026
Tutorials

Epically Powerful: An open-source software and mechatronics infrastructure for wearable robotic systems

DGX agent

arXiv:2511.05033v2 Announce Type: replace Abstract: Epically Powerful is an open-source robotics infrastructure that streamlines the underlying framework of wearable robotic systems - managing communi

tutorialsarxiv-cs-ro
11 Aug 2026
Research

Janus: An Algorithm-Evaluator Co-Evolution Framework for LLM-Driven Discovery under Expensive Evaluation Budgets

DGX agent

arXiv:2608.08189v1 Announce Type: new Abstract: LLM-driven program discovery relies on rapid evaluator feedback, but many scientific and engineering tasks require high-fidelity simulations, hardware e

researcharxiv-cs-ai
11 Aug 2026
Model Releases

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

DGX agent

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Reconfigurable Structural Robotic Assembly: Interlocking 3D Aggregations with Self-Aligning Compound Nested Lattice Modules

DGX agent

arXiv:2608.07576v1 Announce Type: new Abstract: Robotic construction systems often treat the material system and the robot as separate design problems, locating intelligence primarily in hardware, sen

safetyarxiv-cs-ro
11 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 is the ‘killer app’ that is going to sell A LOT of DGX Sparks

DGX agent

Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getti

model-releasesr-localllama
10 Aug 2026
Model Releases

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

DGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Learning Fault-Tolerant Locomotion with Adaptive Gait Timing

DGX agent

arXiv:2608.07328v1 Announce Type: cross Abstract: Hardware failures require legged robots to rapidly reorganize coordination and gait timing to maintain stability and mobility. This is particularly ch

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

Extremely slow DSpark draft model performance (1-2 t/s) with DeepSeek-V4-Flash on llama-server compared to MTP?

DGX agent

Hey everyone, I could use some advice on setting up speculative decoding correctly with llama-server. My Hardware: GPUs: RTX 4090 + RTX 6000 Pro (120GB total VRAM) RAM: 32GB I am currently testing the

model-releasesr-localllama
8 Aug 2026
Safety

CheckOne: Lightweight Fault Detection and Mitigation for Vision Transformers

DGX agent

arXiv:2608.04035v1 Announce Type: cross Abstract: The wide adoption of Vision Transformers (ViTs) in safety-critical applications raises reliability concerns related to hardware faults. Algorithm-Base

safetyarxiv-cs-ai
6 Aug 2026
Research

LLM-based Vulnerability Discovery in Business Process Documentation

DGX agent

arXiv:2608.04271v1 Announce Type: cross Abstract: Just like software and hardware, business processes are susceptible to vulnerabilities that can lead to product quality issues, delays, and increased

researcharxiv-cs-cl
6 Aug 2026
Research

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

DGX agent

The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling their use across diverse hardware and software plat

researchapple-ml-research
6 Aug 2026
Industry

Report: OpenAI’s upcoming AI speaker will be shaped like a donut and cost around $300

DGX agent

More details have emerged from Bloomberg about OpenAI Group PBC’s long-awaited hardware device, which was previously described as an artificial intelligence-enabled speaker that will be a “physical ma

industrysiliconangle
6 Aug 2026
Local Ai

Representational separation between unitary and channel quantum generative models via shared classical randomness at shallow depth

DGX agent

arXiv:2608.05110v1 Announce Type: cross Abstract: Near-term quantum hardware limits circuit depth and often imposes geometrically local connectivity for quantum generative models, restricting the outp

local-aiarxiv-cs-ai
6 Aug 2026
Research

A Survey on Design Methodologies for Accelerating Deep Learning on Heterogeneous Architectures

DGX agent

arXiv:2311.17815v3 Announce Type: replace-cross Abstract: Given their increasing size and complexity, the need for efficient execution of deep neural networks has become increasingly pressing in the d

researcharxiv-cs-ai
5 Aug 2026
Model Releases

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

DGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

model-releasesr-localllama
5 Aug 2026
Model Releases

Morphology-Aware Implicit Super-Resolution Network for Pathological Images

DGX agent

arXiv:2608.03664v1 Announce Type: new Abstract: Accurate diagnosis in Digital Pathology (DP) relies on high-resolution whole-slide images, yet clinical deployment is often limited by hardware costs. S

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Scaling agentic AI: How UiPath built its high-performance GPU platform on AI Hypercomputer

DGX agent

As a market leader in enterprise agentic automation and business orchestration, UiPath is helping to pioneer an industry shift toward agentic AI. With it, the company is deploying autonomous agents to

model-releasesgoogle-cloud-ai
5 Aug 2026
← Previous
1…4344454647…94
Next →