AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
All
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
Hardware

OlmoEarth v1.1: A more efficient family of OlmoEarth models

DGX agent

arXiv:2605.20804v1 Announce Type: new Abstract: We present a set of improvements to the OlmoEarth family. These improvements allow us to cut compute costs during training (1.7 imes reduction in GPU ho

hardwarearxiv-cs-cv
21 May 2026
Hardware
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR

DGX agent

arXiv:2605.20863v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has recently unlocked strong reasoning capabilities in large language models (LLMs), triggering

hardwarearxiv-cs-lg
21 May 2026
Hardware

PulseCol: Periodically Refreshed Column-Sparse Attention for Accelerating Diffusion Language Models

DGX agent

arXiv:2605.20813v1 Announce Type: new Abstract: Inference in diffusion large language models (dLLMs) is computationally expensive, as full self-attention must be repeatedly executed at each step of th

hardwarearxiv-cs-cl
21 May 2026
Hardware

So much for Nvidia’s cash flow. Wild.

DGX agent

So much for Nvidia’s cash flow. Wild. When does circular financing stop? Similar to a Ponzi Scheme, when cash outflows become greater than inflows. In this last quarter, NVDA cash grew by only ~600, a

hardwaregary-marcus--x
21 May 2026
Hardware

Taiwan is seeking to detain three people for forging documents to export Nvidia-powered Super Micro servers to China, Hong Kong, and Macau, breaking US rules (Bloomberg)

DGX agent

Bloomberg: Taiwan is seeking to detain three people for forging documents to export Nvidia-powered Super Micro servers to China, Hong Kong, and Macau, breaking US rules — Taiwanese officials are seeki

hardwaretechmeme
21 May 2026
Hardware

Understanding Deterioration Random Effects for Causal Discovery in Infrastructure Management

DGX agent

arXiv:2605.20400v1 Announce Type: cross Abstract: Infrastructure deterioration poses significant challenges for asset management, yet existing approaches rely on population-averaged models that overlo

hardwarearxiv-cs-lg
21 May 2026
Hardware

Accelerating Sparse Transformer Inference on GPU

DGX agent

arXiv:2506.06095v4 Announce Type: replace Abstract: Large language models (LLMs) are popular around the world due to their powerful understanding capabilities. As the core component of LLMs, accelerat

hardwarearxiv-cs-lg
20 May 2026
Hardware

Banger take

DGX agent

Banger take Gavin Baker: 'I've been optimistic that the fundamental shortage of wafers, which is really controlled by Taiwan Semi, will prevent a bubble.' 'If Taiwan Semi did what Jensen wanted, Nvidi

hardwaredylan-patel--x
20 May 2026
Hardware

COBALT: Crowdsourcing Robot Learning via Cloud-Based Teleoperation with Smartphones

DGX agent

arXiv:2605.19138v1 Announce Type: cross Abstract: The scarcity of large-scale, high-quality demonstration data remains a bottleneck in scaling imitation learning for robotic manipulation. We present C

hardwarearxiv-cs-ai
20 May 2026
Hardware

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs

DGX agent

arXiv:2605.19269v1 Announce Type: new Abstract: Transformer training systems are built around dense linear algebra, yet a nontrivial fraction of end-to-end time is spent on surrounding memory-bound op

hardwarearxiv-cs-lg
20 May 2026
Hardware

Decentralized Direct Volume Rendering: A Browser-Native GPU Architecture for MRI Digital Twins in Resource-Constrained Settings

DGX agent

arXiv:2605.19737v1 Announce Type: cross Abstract: Digital Twin (DT) technology holds immense potential for surgical planning and personalized medicine. However, generating interactive, patient-specifi

hardwarearxiv-cs-cv
20 May 2026
Hardware

Document and sources: China banned Nvidia's RTX 5090D V2 chip, aimed at gamers and animators, while Jensen Huang was visiting China with Donald Trump last week (Financial Times)

DGX agent

Financial Times: Document and sources: China banned Nvidia's RTX 5090D V2 chip, aimed at gamers and animators, while Jensen Huang was visiting China with Donald Trump last week — Beijing aims to suppo

hardwaretechmeme
20 May 2026
Hardware

Exa Labs raises 250M at 2.2B valuation for its AI search tools

DGX agent

Search startup Exa Labs Inc. today announced that it has raised 250 million in funding to purchase more infrastructure. The round was led by Andersen Horowitz. It comes less than a year after Exa’s pr

hardwaresiliconangle
20 May 2026
Hardware

FiLark: a streaming-first software framework for end-to-end exploration, annotation, and algorithm integration in distributed acoustic sensing

DGX agent

arXiv:2605.20132v1 Announce Type: cross Abstract: Distributed acoustic sensing (DAS) systems generate continuous, ultra-high-channel-count data streams at rates that exceed the capabilities of convent

hardwarearxiv-cs-lg
20 May 2026
Hardware

GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems

DGX agent

arXiv:2605.19945v1 Announce Type: cross Abstract: Mixture-of-Expert (MoE) models enable efficient inference by employing smaller experts and activating only a subset of them per token. MoE serving eng

hardwarearxiv-cs-ai
20 May 2026
Hardware

HiLiftAeroML: High-Fidelity Computational Fluid Dynamics Dataset for High-Lift Aircraft Aerodynamics

DGX agent

arXiv:2605.19565v1 Announce Type: cross Abstract: This paper describes the first-ever open-source high-fidelity CFD dataset of a high-lift aircraft for the purpose of AI surrogate model development. T

hardwarearxiv-cs-lg
20 May 2026
Hardware

Hyrax: An Extensible Framework for Rapid ML Experimentation and Unsupervised Discovery in the Era of Rubin, Roman, and Euclid

DGX agent

arXiv:2605.18959v1 Announce Type: cross Abstract: The NSF-DOE Vera C. Rubin Observatory, Roman Space Telescope, Euclid, and other next-generation surveys will deliver imaging, spectroscopic, and time-

hardwarearxiv-cs-lg
20 May 2026
Hardware

Lambda partners with Hudson River Trading to power quantitative research and development

DGX agent

Lambda Labs announced a partnership with Hudson River Trading (HRT), a quantitative trading firm, to provide computational infrastructure and GPU resources for HRT's quantitative research and developm

hardwarelambda-labs
20 May 2026
Hardware

Nvidia almost doubles its data center revenue as it powers to another solid earnings beat

DGX agent

Chipmaker Nvidia Corp., the world’s most valuable company, crushed earnings expectations once again today, benefiting from massive demand for high-end artificial intelligence chips. The company report

hardwaresiliconangle
20 May 2026
Hardware

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond

DGX agent

arXiv:2605.19660v1 Announce Type: cross Abstract: The rapid advancement toward long-context reasoning and multi-modal intelligence has made the memory footprint of the Key-Value (KV) cache a dominant

hardwarearxiv-cs-cl
20 May 2026
Hardware

PiKV: KV Cache Management System for Mixture of Experts

DGX agent

arXiv:2508.06526v3 Announce Type: replace-cross Abstract: As large-scale language models continue to scale up in both size and context length, the memory and communication cost of key-value (KV) cache

hardwarearxiv-cs-ai
20 May 2026
Hardware

Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization

DGX agent

arXiv:2605.19782v1 Announce Type: new Abstract: LLM discovery and optimization systems are increasingly applied across domains, implementing a common propose-evaluate-revise loop. Such optimization or

hardwarearxiv-cs-ai
20 May 2026
Hardware

'Ryzen 5800X3D 10th Anniversary Edition' may help you avoid paying for a new PC

DGX agent

AMD is preparing an AM4 10th Anniversary Edition Ryzen 7 5800X3D for Q2 2026 launch, reviving its popular Zen 3 X3D chip with the same core specifications but updated packaging. The move is framed as

hardwarears-technica
20 May 2026
Hardware

Soft Learning

DGX agent

arXiv:2605.18889v1 Announce Type: cross Abstract: Modern machine learning forces practitioners to choose between powerful but expensive deep networks and fast but limited classical algorithms. Here we

hardwarearxiv-cs-ai
20 May 2026
Hardware

Sources: Nvidia's business development group, not its VC arm NVentures, has led much of its ~$90B dealmaking push across 145+ companies over the past 16 months (Financial Times)

DGX agent

Financial Times: Sources: Nvidia's business development group, not its VC arm NVentures, has led much of its ~90B dealmaking push across 145+ companies over the past 16 months — Nvidia's Huang bankrol

hardwaretechmeme
20 May 2026
Hardware

SuperInfer: SLO-Aware Rotary Scheduling and Memory Management for LLM Inference on Superchips

DGX agent

arXiv:2601.20309v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) serving faces a fundamental tension between stringent latency Service Level Objectives (SLOs) and limited GPU memor

hardwarearxiv-cs-ai
20 May 2026
Hardware

TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload

DGX agent

arXiv:2605.20179v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization an

hardwarearxiv-cs-cl
20 May 2026
Hardware

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of …

DGX agent

🌍Today we release Mosaic, a probabilistic weather model that shifts the Pareto frontier of ML weather forecasting. It matches the skill of state-of-the-art models while generating a 24-member, 10-day

hardwareemad-mostaque--x
20 May 2026
Hardware

Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption

DGX agent

arXiv:2605.19593v1 Announce Type: new Abstract: Modern deployments of Large Language Models (LLMs) increasingly require serving multiple models with diverse architectures, sizes, and specialization on

hardwarearxiv-cs-ai
20 May 2026
Hardware

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing

DGX agent

arXiv:2605.18796v1 Announce Type: cross Abstract: LLM cascades and model routing promise lower inference cost by sending easy queries to a small model and escalating hard ones to a large model, but mo

hardwarearxiv-cs-cl
20 May 2026
Hardware

3D Skew Gaussian Splatting with Any Camera Trajectory Visualization Engine

DGX agent

arXiv:2605.18334v1 Announce Type: new Abstract: While 3D Gaussian Splatting (3DGS) has revolutionized real-time photorealistic view synthesis, its fundamental reliance on symmetric Gaussian distributi

hardwarearxiv-cs-cv
19 May 2026
Hardware

Breaking: Andrej Karpathy just filed his Q1 2026 13F. Here's everything you need to know about his recent 13F Top 10 positions: 1. VanEck Se…

DGX agent

Breaking: Andrej Karpathy just filed his Q1 2026 13F. Here's everything you need to know about his recent 13F Top 10 positions: 1. VanEck Semiconductor ETF - SMH - [Put] — 2.04B 2. Nvidia - NVDA - [Pu

hardwaredylan-patel--x
19 May 2026
Hardware

Bundle Adjustment in the Eager Mode

DGX agent

arXiv:2409.12190v4 Announce Type: replace-cross Abstract: Bundle adjustment (BA) is a critical technique in various robotic applications such as simultaneous localization and mapping (SLAM), augmented

hardwarearxiv-cs-cv
19 May 2026
Hardware

Charon: A Unified and Fine-Grained Simulator for Large-Scale LLM Training and Inference

DGX agent

arXiv:2605.17164v1 Announce Type: cross Abstract: Deploying large-scale LLM training and inference with optimal performance is exceptionally challenging due to a complex design space of parallelism st

hardwarearxiv-cs-ai
19 May 2026
Hardware

cuNRTO: GPU-Accelerated Nonlinear Robust Trajectory Optimization

DGX agent

arXiv:2603.02642v2 Announce Type: replace Abstract: Robust trajectory optimization enables autonomous systems to operate safely under uncertainty by computing control policies that satisfy the constra

hardwarearxiv-cs-ro
19 May 2026
Hardware

DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention

DGX agent

arXiv:2605.18753v1 Announce Type: cross Abstract: Current hierarchical attention methods, such as NSA and InfLLMv2, select the top-k relevant key-value (KV) blocks based on coarse attention scores and

hardwarearxiv-cs-ai
19 May 2026
Hardware

Decart raises $300M for its AI optimization software, world models

DGX agent

Artificial intelligence developer Decart.ai Inc. today announced that it has raised 300 million in funding at a nearly 4 billion valuation. Radical Ventures led the round with participation from Nvidi

hardwaresiliconangle
19 May 2026
Hardware

Dell’s CFO sees the enterprise AI buildout as a generational opportunity still in its opening act

DGX agent

AI factory momentum has reached a critical inflection point as capital joins silicon and energy as key constraints in the race to build enterprise AI infrastructure. That dynamic is playing out in rea

hardwaresiliconangle
19 May 2026
Hardware

Efficient 3D Content Reconstruction and Generation

DGX agent

arXiv:2605.18052v1 Announce Type: new Abstract: Automatic 3D content creation seeks to replace labor-intensive modeling and scanning pipelines with systems that can synthesize or recover 3D assets dir

hardwarearxiv-cs-cv
19 May 2026
Hardware

Elon Musk met with NVIDIA’s Ian Buck, who gave him the new Vera CPU for SpaceXAI.

DGX agent

I cannot verify this claim as accurate. The URL provided doesn't follow standard X post formats, and there is no publicly documented 'Vera CPU' from NVIDIA. This appears to be either fabricated or spe

hardwareelon-musk--x
19 May 2026
Hardware

Google's SynthID AI watermarking tech is being adopted by OpenAI, Nvidia, and more

DGX agent

Google's SynthID watermarking technology is being adopted by major AI companies including OpenAI, Kakao, and ElevenLabs to watermark AI-generated content. Google has also partnered with NVIDIA to wate

hardwarears-technica
19 May 2026
Hardware

GPU-Accelerated Deep Learning for Heatwave Prediction and Urban Heat Risk Assessment

DGX agent

arXiv:2605.16435v1 Announce Type: cross Abstract: Heatwaves are an important problem in cities, and climate change makes this problem more difficult. In this paper, we present a GPU-based deep learnin

hardwarearxiv-cs-ai
19 May 2026
Hardware

Guard: Scalable Straggler Detection and Node Health Management for Large-Scale Training

DGX agent

arXiv:2605.17879v1 Announce Type: cross Abstract: Training frontier-scale foundation models involves coordinating tens of thousands of GPUs over multi-month runs, where even minor performance degradat

hardwarearxiv-cs-ai
19 May 2026
Hardware

Hawkeye: Reproducing GPU-Level Non-Determinism

DGX agent

arXiv:2603.20421v2 Announce Type: replace-cross Abstract: We present Hawkeye, a system for analyzing and reproducing GPU-level arithmetic operations. Using our framework, anyone can re-execute on a CP

hardwarearxiv-cs-lg
19 May 2026
Hardware

IVF-TQ: Streaming-Robust Approximate Nearest Neighbor Search via a Codebook-Free Residual Layer

DGX agent

arXiv:2605.17415v1 Announce Type: cross Abstract: We propose IVF-TQ, an IVF index with a codebook-free residual layer: a fixed random rotation followed by precomputed Lloyd-Max scalar quantization dep

hardwarearxiv-cs-ai
19 May 2026
Hardware

KVDrive: A Holistic Multi-Tier KV Cache Management System for Long-Context LLM Inference

DGX agent

arXiv:2605.18071v1 Announce Type: new Abstract: Supporting long-context LLMs is challenging due to the substantial memory demands of the key-value (KV) cache. Existing offloading systems store the ful

hardwarearxiv-cs-cl
19 May 2026
Hardware

LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models

DGX agent

arXiv:2605.17289v1 Announce Type: cross Abstract: Unstructured sparsity is now natively accelerated by recent GPU kernels and dataflow hardware, shifting the bottleneck from inference execution to the

hardwarearxiv-cs-ai
19 May 2026
Hardware

Lightweight Gaussian Process Inference in C++ on Metal and CUDA

DGX agent

arXiv:2605.17898v1 Announce Type: new Abstract: Gaussian process (GP) inference in Python is dominated by libraries such as GPyTorch and GPflow, which are built on deep-learning frameworks and inherit

hardwarearxiv-cs-lg
19 May 2026
← Previous
1…2122232425…37
Next →