AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
23 Apr 2026

Stream-CQSA: Avoiding Out-of-Memory in Attention Computation via Flexible Workload Scheduling

HardwareDGX agent

arXiv:2604.20819v1 Announce Type: new Abstract: The scalability of long-context large language models is fundamentally limited by the quadratic memory cost of exact self-attention, which often leads t

22 Apr 2026

ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants

HardwareDGX agent

arXiv:2604.18616v1 Announce Type: cross Abstract: LLM-based coding agents can generate functionally correct GPU kernels, yet their performance remains far below hand-optimized libraries on critical co

From Rainforests to Recycling Plants: 5 Ways NVIDIA AI Is Protecting the Planet

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware
DGX agent

NVIDIA AI and accelerated computing are advancing sustainability, climate science and energy efficiency through five key applications. These include RecycleOS, an AI and robotics solution that helps r

21 Apr 2026

Condense, Don't Just Prune: Enhancing Efficiency and Performance in MoE Layer Pruning

HardwareDGX agent

arXiv:2412.00069v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has garnered significant attention for its ability to scale up neural networks while utilizing the same or even fewer

Enabling AI ASICs for Zero Knowledge Proof

HardwareDGX agent

arXiv:2604.17808v1 Announce Type: cross Abstract: Zero-knowledge proof (ZKP) provers remain costly because multi-scalar multiplication (MSM) and number-theoretic transforms (NTTs) dominate runtime as

MerLin: A Discovery Engine for Photonic and Hybrid Quantum Machine Learning

Model ReleasesDGX agent

arXiv:2602.11092v2 Announce Type: replace Abstract: Identifying where quantum models may offer practical benefits in near term quantum machine learning (QML) requires moving beyond isolated algorithmi

RACE Attention: A Strictly Linear-Time Attention Layer for Training on Outrageously Large Contexts

HardwareDGX agent

arXiv:2510.04008v5 Announce Type: replace Abstract: Softmax Attention has a quadratic time complexity in sequence length, which becomes prohibitive to run at long contexts, even with highly optimized

Scalable and Adaptive Parallel Training of Graph Transformer on Large Graphs

HardwareDGX agent

arXiv:2604.16715v1 Announce Type: cross Abstract: Graph foundation models have demonstrated remarkable adaptability across diverse downstream tasks through large-scale pretraining on graphs. However,

Towards Real-Time ECG and EMG Modeling on mu NPUs

ResearchDGX agent

arXiv:2604.18067v1 Announce Type: new Abstract: The miniaturisation of neural processing units (NPUs) and other low-power accelerators has enabled their integration into microcontroller-scale wearable

Web-Gewu: A Browser-Based Interactive Playground for Robot Reinforcement Learning

HardwareDGX agent

arXiv:2604.17050v1 Announce Type: new Abstract: With the rapid development of embodied intelligence, robotics education faces a dual challenge: high computational barriers and cumbersome environment c

20 Apr 2026

Closing the Theory-Practice Gap in Spiking Transformers via Effective Dimension

ResearchDGX agent

arXiv:2604.15769v1 Announce Type: cross Abstract: Spiking transformers achieve competitive accuracy with conventional transformers while offering 38-57imes energy efficiency on neuromorphic hardware,

Scalable Posterior Uncertainty for Flexible Density-Based Clustering

HardwareDGX agent

arXiv:2603.03188v2 Announce Type: replace-cross Abstract: We introduce a novel framework for uncertainty quantification in clustering that combines martingale posterior distributions with density-base

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was th…

HardwareDGX agent

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was the most popular model on our training platform. K2.6 on Firew

17 Apr 2026

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability

HardwareDGX agent

arXiv:2604.13048v1 Announce Type: cross Abstract: Modern cloud-native platforms expose thousands of time series metrics through systems like Prometheus, yet formulating correct queries in domain-speci

Reference-Free Sampling-Based Model Predictive Control

HardwareDGX agent

arXiv:2511.19204v3 Announce Type: replace Abstract: We present a sampling-based model predictive control (MPC) framework that enables emergent locomotion without relying on handcrafted gait patterns o

16 Apr 2026

The Multi-GPU Era: Why Heterogeneous Compute Is Becoming Enterprise Standard

HardwareDGX agent

Heterogeneous computing, which combines multiple types of GPUs and processors, is becoming standard in enterprise environments to handle diverse workloads more efficiently than traditional homogeneous

15 Apr 2026

Cline with Ollama on a RTX4090 (24GRAM) and i9 with 64 GRAM

Local AiDGX agent

This Reddit post from r/ollama discusses a user's experience running Cline (an AI coding agent) with Ollama on a high-end local hardware setup consisting of an NVIDIA RTX 4090 with 24GB VRAM and an In

Roop Unleashed 4.3.1 not fully utilizing RTX 5070 Ti / 5080X (Low GPU/RAM usage)

HardwareDGX agent

This Reddit thread from r/StableDiffusion discusses a performance issue where Roop Unleashed version 4.3.1 — a face-swapping tool commonly used alongside Stable Diffusion — fails to fully utilize the

Running a 31B model locally made me realize how insane LLM infra actually is

Local AiDGX agent

A Reddit post from r/ollama in which a user shares their experience running a 31B parameter model locally using Ollama, reflecting on the surprisingly demanding hardware and infrastructure requirement

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token…

HardwareDGX agent

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token Town' - the crack team of AI Engineers and Model Behavior En

14 Apr 2026

CUTEv2: Unified and Configurable Matrix Extension for Diverse CPU Architectures with Minimal Design Overhead

ResearchDGX agent

arXiv:2604.11615v1 Announce Type: cross Abstract: Matrix extensions have emerged as an essential feature in modern CPUs to address the surging demands of AI workloads. However, existing designs often

Device-Conditioned Neural Architecture Search for Efficient Robotic Manipulation

SafetyDGX agent

arXiv:2604.10170v1 Announce Type: cross Abstract: The growing complexity of visuomotor policies poses significant challenges for deployment with heterogeneous robotic hardware constraints. However, mo

IceCache: Memory-efficient KV-cache Management for Long-Sequence LLMs

HardwareDGX agent

arXiv:2604.10539v1 Announce Type: cross Abstract: Key-Value (KV) cache plays a crucial role in accelerating inference in large language models (LLMs) by storing intermediate attention states and avoid

NVIDIA Ising Introduces AI-Powered Workflows to Build Fault-Tolerant Quantum Systems

HardwareDGX agent

NVIDIA Ising is the world's first family of open-source quantum AI models, designed to help researchers and enterprises build quantum processors capable of running useful applications. The family span

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

HardwareDGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

Tessera: Unlocking Heterogeneous GPUs through Kernel-Granularity Disaggregation

HardwareDGX agent

arXiv:2604.10180v1 Announce Type: cross Abstract: Disaggregation maps parts of an AI workload to different types of GPUs, offering a path to utilize modern heterogeneous GPU clusters. However, existin

Vestibular reservoir computing

ResearchDGX agent

arXiv:2604.09943v1 Announce Type: new Abstract: Reservoir computing (RC) is a computational framework known for its training efficiency, making it ideal for physical hardware implementations. However,

what's the best place to buy GPU server?

HardwareDGX agent

This Reddit thread from r/ollama discusses community recommendations for purchasing or renting GPU servers to run Ollama and local LLMs. It likely covers options ranging from dedicated GPU server prov

13 Apr 2026

$200/month is enough to buy an H100 GPU for 6 hours every workday

HardwareDGX agent

Soumith Chintala shared a post highlighting that $200 per month is sufficient to rent access to an NVIDIA H100 GPU for approximately 6 hours every workday, making high-end AI compute more accessible t

10 Apr 2026

🫡 @aiDotEngineer in London was extremely good. Absolutely incredible speaker line up, great curation, escaped the Silicon Valley bubble mas…

HardwareDGX agent

"**AI Engineer Europe (London) — Conference Recap**

Behind the Analysis with Google Cloud and Team USA: Architecting AI infrastructure for U.S. Winter Olympians

HardwareDGX agent

In freeskiing and snowboarding, traditional video replay shows you what happened during a complex aerial maneuver, but it fails to explain the physics of how it was possible. At the speed of the sport

That’s a wrap on HumanX. Custom comics, hats, a happy hour with @getmetronome & @nvidia, and two sessions on what actually matters for AI-na…

HardwareDGX agent

Together Computer (Together AI) participated in HumanX 2026, a major AI conference held April 6–9 in San Francisco, where they hosted activations including custom comics, branded hats, and a happy ...

9 Apr 2026

Cut Checkpoint Costs with About 30 Lines of Python and NVIDIA nvCOMP

HardwareDGX agent

Training LLMs requires periodic checkpoints — full snapshots of model weights, optimizer states, and gradients — whose storage costs can reach $200,000/month for a 405B model on 128 NVIDIA DGX B200...

How to Accelerate Protein Structure Prediction at Proteome-Scale

HardwareDGX agent

NVIDIA, Google DeepMind, EMBL-EBI, and Seoul National University collaborated to extend the AlphaFold Protein Structure Database (AFDB) beyond monomeric structures to proteome-scale quaternary stru...

Running Large-Scale GPU Workloads on Kubernetes with Slurm

HardwareDGX agent

NVIDIA's open-source project **Slinky** (developed by SchedMD, now part of NVIDIA) enables organizations to run full Slurm clusters directly on Kubernetes infrastructure by managing the complete li...

8 Apr 2026

Integrate Physical AI Capabilities into Existing Apps with NVIDIA Omniverse Libraries

HardwareDGX agent

NVIDIA has introduced a modular, library-based architecture for Omniverse, exposing core components—RTX rendering, PhysX-based simulation, and data storage pipelines—as standalone, headless-first C...

New Vultr Identity and Access Management Upgrades

HardwareDGX agent

Vultr has upgraded its Identity and Access Management (IAM) framework by replacing legacy ACLs with fine-grained, policy-driven access controls, enabling organizations to assign roles, manage permi...

7 Apr 2026

NVIDIA’s New AI Just Changed Everything

HardwareDGX agent

I was unable to directly access the YouTube video at the provided URL, and my search did not return results specifically identifying that video's content. YouTube videos cannot be fetched or read a...

12 Aug 2026

Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence

HardwareDGX agent

arXiv:2608.10720v1 Announce Type: new Abstract: Omni-modal dialogue models can understand multimodal inputs and synthesize spoken replies, yet their responses remain visually disembodied. We introduce

From Detection to Understanding: TAR and TAR-Bench for Multi-Task Traffic Anomaly Reasoning

HardwareDGX agent

arXiv:2608.10317v1 Announce Type: new Abstract: We present TAR (Traffic Anomaly Reasoning) and TAR-Bench datasets, resources for training and evaluating video-language models beyond anomaly detection.

Hand-Written PTX Tensor-Core GEMM Kernels: A Multi-Precision Study on NVIDIA L4

HardwareDGX agent

arXiv:2608.10103v1 Announce Type: cross Abstract: High-performance Tensor Core kernels rely on a low-level PTX pipeline built from asynchronous data movement with cp.async, warp-level matrix loads wit

How to Choose Full-Stack Observability for NVIDIA AI Factories

HardwareDGX agent

A full‑stack observability framework for NVIDIA AI factories links telemetry from compute, networking, storage, orchestration and application layers using specialized tools (DCGM, NVSM, UFM, NetQ, NMX

HyWA: Architecture-Preserving Personalized Voice Activity Detection for Full-Duplex Voice Assistants

HardwareDGX agent

arXiv:2510.12947v3 Announce Type: replace-cross Abstract: Voice activity detection (VAD) serves as an early gate in voice-assistant pipelines for smart devices. Because conventional VADs respond to sp

IBM inks $240M infrastructure deal with AI-optimized cloud operator Together AI

HardwareDGX agent

IBM Corp. will provide infrastructure to artificial intelligence startup Together AI Inc. as part of a 240 million deal announced today. The partnership comes a few weeks after the latter company rais

Lambda prices $926 million senior secured term loan B facility, the first investment-grade-rated term loan B financing by a private neocloud

HardwareDGX agent

First private neocloud to execute an investment-grade-rated financing in the term loan B market, further broadening the investor base for AI infrastructure financing. Facility supports the purchase an

Neuroevolution Arena: Nested Ecological Evaluation of Update-and-Inheritance Regimes across Neural Architectures

HardwareDGX agent

arXiv:2608.10323v1 Announce Type: new Abstract: Competitive artificial-life systems can rank trained controllers differently under training and ecological evaluation. We present Neuroevolution Arena,

NVIDIA AI Factory Compute Is Becoming an Investable Asset Class

HardwareDGX agent

We announced partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish independent financing platforms designed to mobilize over $500 billion of third-party capit

TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling

HardwareDGX agent

arXiv:2608.10402v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models is moving toward multi-turn agentic workloads, where rollout tasks repeatedly pause for external e

Tiered KV cache for large LLMs on Amazon SageMaker HyperPod with Curvine

HardwareDGX agent

Running large language model inference at scale forces a KV cache trade-off: oversized GPU instances or slow time-to-first-token. This post builds a tiered KV cache on Amazon SageMaker HyperPod that e

TransitReID: Transit OD Data Collection with Occlusion-Resistant Dynamic Passenger Re-Identification

HardwareDGX agent

arXiv:2504.11500v3 Announce Type: replace-cross Abstract: Transit Origin-Destination (OD) data are fundamental for optimizing public transit services, yet current collection methods, such as manual su

11 Aug 2026

Beyond Isotropic Assumptions: Continuity-Constrained Segmentation and GPU Morphometry for Nanoscale GBM Analysis

HardwareDGX agent

arXiv:2608.07575v1 Announce Type: new Abstract: Confocal microscopy of optically cleared and swelled tissue resolves complex biological structures in 3D, but such acquisitions are highly anisotropic:

Big Tech's AI boom echoes the 1870s railroad buildout, and Nvidia shifting risk to institutional capital may expose investors if AI revenues fail to materialize (Ben Thompson/Stratechery)

HardwareDGX agent

Ben Thompson / Stratechery: Big Tech's AI boom echoes the 1870s railroad buildout, and Nvidia shifting risk to institutional capital may expose investors if AI revenues fail to materialize — On Januar

Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching

HardwareDGX agent

arXiv:2608.09444v1 Announce Type: cross Abstract: A main promise of looped language models (LMs) is depth-adaptive inference. By iterating a block of shared layers a variable number of times, the mode

Design Space of Self--Consistent Electrostatic Machine Learning Interatomic Potentials

HardwareDGX agent

arXiv:2603.14700v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have become widely used tools in atomistic simulations. For much of the history of this field,

Differentiate the Solver, Not the Equation: Reverse-Sweep Adjoints for Block Implicit Simulation

HardwareDGX agent

arXiv:2608.08559v1 Announce Type: cross Abstract: Differentiable simulation is a key component in learning, control, and inverse problems, where gradients through nonlinear implicit solvers are requir

EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference

HardwareDGX agent

arXiv:2608.07964v1 Announce Type: cross Abstract: Load Balancing has emerged as a critical problem in expert-parallel distributed inference of Mixture-of-Experts (MoE) models. As routing distributions

ERF-GS: Reconstructing Fast Motion from Disjoint Event-RGB Viewpoints

HardwareDGX agent

arXiv:2608.08531v1 Announce Type: new Abstract: Deep learning-driven representations such as neural radiance fields (NeRFs) and 3D Gaussian splatting (3DGS) have revolutionized the field of dynamic 3D

EsaacSim: A Multimodal Event Camera Add-on for NVIDIA Isaac Sim

HardwareDGX agent

arXiv:2608.08522v1 Announce Type: new Abstract: Event-based vision is becoming an increasingly important sensing paradigm for robotics, yet its adoption remains limited by sensor availability and the

From Approachability Residuals to Anytime-Valid Evidence: The Online Convex Geometry of Testing by Betting

HardwareDGX agent

arXiv:2608.09450v1 Announce Type: new Abstract: Betting-based sequential tests and Blackwell approachability are linked by a rate-explicit reduction through support-function residuals. For a compact c

Holder Signed Distance: A Differentiable, Signed, Parallelizable Metric for Robotics

HardwareDGX agent

arXiv:2608.07707v1 Announce Type: new Abstract: Computing distances between sets is essential in robotic motion planning and control, where differentiable gradients enable real-time optimization. The

← Previous
1…89101112…75
Next →