AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,734 results
Hardware

LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models

DGX agent

arXiv:2605.17289v1 Announce Type: cross Abstract: Unstructured sparsity is now natively accelerated by recent GPU kernels and dataflow hardware, shifting the bottleneck from inference execution to the

hardwarearxiv-cs-ai
19 May 2026
Hardware
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Lightweight Gaussian Process Inference in C++ on Metal and CUDA

DGX agent

arXiv:2605.17898v1 Announce Type: new Abstract: Gaussian process (GP) inference in Python is dominated by libraries such as GPyTorch and GPflow, which are built on deep-learning frameworks and inherit

hardwarearxiv-cs-lg
19 May 2026
Hardware

LogRouter: Adaptive Two-Level LLM Routing for Log Question Answering in Big Data Systems

DGX agent

arXiv:2605.18015v1 Announce Type: new Abstract: Production log analytics in self-hosted, resource-constrained environments requires natural-language access to massive log streams without the cost of r

hardwarearxiv-cs-lg
19 May 2026
Hardware

LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation

DGX agent

arXiv:2605.18739v1 Announce Type: new Abstract: We present LongLive-2.0, an NVFP4-based parallel infrastructure throughout the full training and inference workflow of long video generation, addressing

hardwarearxiv-cs-cv
19 May 2026
Hardware

NanoQuant: Efficient Sub-1-Bit Quantization of Large Language Models

DGX agent

arXiv:2602.06694v2 Announce Type: replace Abstract: Weight-only quantization has become a standard approach for efficiently serving large language models (LLMs). However, existing methods fail to effi

hardwarearxiv-cs-lg
19 May 2026
Hardware

Neural-network methods for two-dimensional finite-source reflector design

DGX agent

arXiv:2604.02184v2 Announce Type: replace Abstract: We address the inverse problem of designing two-dimensional reflectors that transform light from a finite, extended source into a prescribed far-fie

hardwarearxiv-cs-lg
19 May 2026
Hardware

NVIDIA and Google Cloud Empower the Next Wave of AI Builders

DGX agent

At this year’s Google I/O conference, NVIDIA and Google Cloud are accelerating the work of more than 100,000 developers in the companies’ joint developer community, which provides curated learning pat

hardwarenvidia-blog
19 May 2026
Hardware

NVIDIA-Verified Agent Skills Provide Capability Governance for AI Agents

DGX agent

NVIDIA agent skills enable specialized agents to deliver targeted capabilities across enterprise workflows , with governance provided through verified, scanned, and signed skills. JFrog AI Catalog aut

hardwarenvidia-developer
19 May 2026
Hardware

Pocket Foundation Models: Distilling TFMs into CPU-Ready Gradient-Boosted Trees

DGX agent

arXiv:2605.18654v1 Announce Type: cross Abstract: A fraud scorer needs to answer in under 2 ms. The best tabular foundation models (TFMs) take 151-1,275 ms on GPU. We close this gap by distilling the

hardwarearxiv-cs-ai
19 May 2026
Hardware

Progressive Generalization Augmentation with Deeply Coupled RND-PPO and Domain-Prioritized Noise Injection for Robust Crop Management Reinforcement Learning

DGX agent

arXiv:2605.17428v1 Announce Type: cross Abstract: Our preliminary experiments on gym-DSSAT maize irrigation tasks revealed that +/-2 degrees C temperature noise causes an 11.9% reduction in economic r

hardwarearxiv-cs-ai
19 May 2026
Hardware

SignMuon: Communication-Efficient Distributed Muon Optimization

DGX agent

arXiv:2605.16311v1 Announce Type: new Abstract: Distributed training of large neural networks is bottlenecked by full-precision gradient communication and by coordinatewise optimizers that ignore the

hardwarearxiv-cs-lg
19 May 2026
Hardware

Sources: Zyphra, which trains and runs inference for its open-weight models on AMD hardware, is raising a 500M Series B at a valuation of at least 5B (Anna Tong/Forbes)

DGX agent

Anna Tong / Forbes: Sources: Zyphra, which trains and runs inference for its open-weight models on AMD hardware, is raising a 500M Series B at a valuation of at least 5B — The Series B round, which ch

hardwaretechmeme
19 May 2026
Hardware

Sparse Mamba Decoder for Quantum Error Correction: Efficient Defect-Centric Processing of Surface Code Syndromes

DGX agent

arXiv:2605.17156v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for building fault-tolerant quantum computers, requiring decoders that are simultaneously accurate, fast,

hardwarearxiv-cs-lg
19 May 2026
Hardware

Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons

DGX agent

arXiv:2605.18389v1 Announce Type: new Abstract: Optimal transport provides a powerful framework for comparing measures while respecting the geometry of their support, but comes with an expensive compu

hardwarearxiv-cs-lg
19 May 2026
Hardware

Stable Audio 3

DGX agent

arXiv:2605.17991v1 Announce Type: cross Abstract: Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models c

hardwarearxiv-cs-ai
19 May 2026
Hardware

StreamingEffect: Real-Time Human-Centric Video Effect Generation

DGX agent

arXiv:2605.17019v1 Announce Type: new Abstract: Streaming video effect generation is highly desirable for live human-centric applications such as e-commerce streaming, entertainment, and vlogging, yet

hardwarearxiv-cs-cv
19 May 2026
Hardware

Systematic Optimization of Real-Time Diffusion Model Inference on Apple M3 Ultra

DGX agent

arXiv:2605.16259v1 Announce Type: cross Abstract: While real-time image generation using diffusion models has advanced rapidly on NVIDIA GPUs, systematic optimization research on non-CUDA platforms su

hardwarearxiv-cs-ai
19 May 2026
Hardware

TeleRAG: Efficient Retrieval-Augmented Generation Inference with Lookahead Retrieval

DGX agent

arXiv:2502.20969v4 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external data sources to enhance factual correctness and domain

hardwarearxiv-cs-lg
19 May 2026
Hardware

TierCheck: Tiered Checkpointing for Fault Tolerance in Large Language Model Training

DGX agent

arXiv:2605.17821v1 Announce Type: cross Abstract: Large Language Model (LLM) training is frequently interrupted by a heterogeneous spectrum of failures, from common GPU crashes to catastrophic cluster

hardwarearxiv-cs-ai
19 May 2026
Hardware

Token-Space Mask Prediction for Efficient Vision Transformer Segmentation

DGX agent

arXiv:2605.18177v1 Announce Type: new Abstract: Query-based Vision Transformer segmentation models typically reconstruct dense spatial feature maps to predict masks, inheriting design patterns from co

hardwarearxiv-cs-cv
19 May 2026
Hardware

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks

DGX agent

arXiv:2605.17170v1 Announce Type: new Abstract: Agentic workloads have emerged as a major workload for LLM inference. They differ significantly from chat-only workloads, requiring long-context process

hardwarearxiv-cs-lg
19 May 2026
Hardware

Truthful Calibration Errors for Multi-Class Prediction

DGX agent

arXiv:2510.06388v2 Announce Type: replace Abstract: Calibrated predictions are useful because their numerical values can be interpreted as probabilities. Calibration errors are therefore widely used t

hardwarearxiv-cs-lg
19 May 2026
Hardware

VeriCache: Turning Lossy KV Cache into Lossless LLM Inference

DGX agent

arXiv:2605.17613v1 Announce Type: cross Abstract: The large size of the KV cache has become a major bottleneck for serving LLMs with increasing context lengths. In response, many KV cache compression

hardwarearxiv-cs-lg
19 May 2026
Hardware

Vultr Announces Milan, Italy, as 33rd Cloud Data Center Region

DGX agent

Vultr expanded its global cloud infrastructure by opening a new data center region in Milan, Italy, marking the company's 33rd cloud data center location worldwide. This expansion provides European cu

hardwarevultr
19 May 2026
Hardware

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human geno…

DGX agent

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human genome on a single GPU in <2 days. Here are the tricks we used:

hardwareclem-delangue--x
19 May 2026
Hardware

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

DGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

hardwarearxiv-cs-ai
18 May 2026
Hardware

A Unified Non-Parametric and Interpretable Point Cloud Analysis via t-FCW Graph Representation

DGX agent

arXiv:2605.15475v1 Announce Type: new Abstract: We introduce an empowered transposed Fully Connected Weighted (t-FCW) graph representation to embed point clouds into a metric space. While original t-F

hardwarearxiv-cs-cv
18 May 2026
Hardware

AgriMind: An Ensemble Deep Learning Framework for Multi-Class Plant Disease Classification

DGX agent

arXiv:2605.16076v1 Announce Type: cross Abstract: Plant disease detection is still largely manual in Bangladesh, where extension workers eyeball leaf samples across millions of smallholdings. We built

hardwarearxiv-cs-ai
18 May 2026
Hardware

Bridging Silicon and the Hippocampus: Algebro-Deterministic Memory 'VaCoAl' as a Substrate for Vector-HaSH and TEM

DGX agent

arXiv:2605.15652v1 Announce Type: cross Abstract: Vector-HaSH and the Tolman-Eichenbaum Machine (TEM) propose that the hippocampal-entorhinal circuit factorizes content from a prestructured grid-cell

hardwarearxiv-cs-ai
18 May 2026
Hardware

Decart, which offers real-time generative video and GPU optimization tech, raised 300M at a ~4B valuation, up from 3.1B after raising 153M in August 2025 (Robbie Whelan/Wall Street Journal)

DGX agent

Robbie Whelan / Wall Street Journal: Decart, which offers real-time generative video and GPU optimization tech, raised 300M at a ~4B valuation, up from 3.1B after raising 153M in August 2025 — Decart'

hardwaretechmeme
18 May 2026
Hardware

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization

DGX agent

arXiv:2605.15824v1 Announce Type: new Abstract: Human-centric video customization, particularly at the garment level, has shown significant commercial value. However, existing approaches cannot suppor

hardwarearxiv-cs-cv
18 May 2026
Hardware

Fine-Tuning NVIDIA Cosmos Predict 2.5 with LoRA/DoRA for Robot Video Generation

DGX agent

This guide demonstrates how to fine-tune NVIDIA's Cosmos Predict 2.5 video generation model using parameter-efficient techniques like LoRA (Low-Rank Adaptation) and DoRA (Mixture of Experts-based adap

hardwarehugging-face
18 May 2026
Hardware

frax: Fast Robot Kinematics and Dynamics in JAX

DGX agent

arXiv:2604.04310v2 Announce Type: replace Abstract: In robot control, planning, and learning, there is a need for rigid-body dynamics libraries that are highly performant, easy to use, and compatible

hardwarearxiv-cs-ro
18 May 2026
Hardware

How the AI industry's intense pressure to keep up and its financial rewards are creating 'walls of resentment' between Silicon Valley workers and their spouses (Alessandra Ram/Wired)

DGX agent

Alessandra Ram / Wired: How the AI industry's intense pressure to keep up and its financial rewards are creating “walls of resentment” between Silicon Valley workers and their spouses — Are you marrie

hardwaretechmeme
18 May 2026
Hardware

NVIDIA CEO Jensen Huang at Dell Technologies World: “Demand Is Going Parabolic, Utterly Parabolic”

DGX agent

Agentic AI inference at one-tenth the cost per token with NVIDIA Vera Rubin NVL72. Agent sandboxes run 50% faster on NVIDIA Vera than traditional CPUs — while enterprise data queries are up to 3x fast

hardwarenvidia-blog
18 May 2026
Hardware

Pwn2Own Berlin 2026: participants earned a total of $1,298,250 for 47 vulnerabilities, with successful exploits of AI products like Codex, Cursor, and LM Studio (Eduard Kovacs/SecurityWeek)

DGX agent

Eduard Kovacs / SecurityWeek: Pwn2Own Berlin 2026: participants earned a total of $1,298,250 for 47 vulnerabilities, with successful exploits of AI products like Codex, Cursor, and LM Studio — Partici

hardwaretechmeme
18 May 2026
Hardware

Together with SpaceXAI, we’re training a significantly larger model from scratch, using 10x more total compute. With Colossus 2’s million H1…

DGX agent

Together with SpaceXAI, we’re training a significantly larger model from scratch, using 10x more total compute. With Colossus 2’s million H100-equivalents and our combined data and training techniques

hardwarecursor--x
18 May 2026
Hardware

TrainMover: An Interruption-Resilient Runtime for ML Training

DGX agent

arXiv:2412.12636v3 Announce Type: replace-cross Abstract: Large-scale ML training jobs are frequently interrupted by hardware and software anomalies, failures, and management events. Existing solution

hardwarearxiv-cs-ai
18 May 2026
Hardware

Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs

DGX agent

The first NVIDIA Vera CPUs arrived at three of the world's leading AI labs on Friday — Anthropic in San Francisco, OpenAI in Mission Bay, SpaceXAI in Palo Alto — followed by a delivery to Oracle Cloud

hardwarenvidia-blog
18 May 2026
Hardware

Reduce your GPU power limit

DGX agent

Setting a GPU power limit using nvidia-smi reduces heat output by approximately 20% with only a 5-8% inference speed loss. Undervolting the GPU can reduce power consumption by 5-15% with zero performa

hardwarer-ollama
16 May 2026
Hardware

SF vibes are frenetic over the huge divide in outcomes and career uncertainty for software engineers; over 5 years ~10K people in AI attained retirement wealth (Deedy/@deedydas)

DGX agent

Deedy / @deedydas: SF vibes are frenetic over the huge divide in outcomes and career uncertainty for software engineers; over 5 years ~10K people in AI attained retirement wealth — The vibes in SF fee

hardwaretechmeme
16 May 2026
Hardware

Shenzhen-listed RoboTechnik, which claims to be the largest silicon photonics tool maker and whose stock is up 340% over the past year, files for a HK listing (Zinnia Lee/Forbes)

DGX agent

Zinnia Lee / Forbes: Shenzhen-listed RoboTechnik, which claims to be the largest silicon photonics tool maker and whose stock is up 340% over the past year, files for a HK listing — RoboTechnik Intell

hardwaretechmeme
16 May 2026
Hardware

A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 (Rebecca Bellan/TechCrunch)

DGX agent

Rebecca Bellan / TechCrunch: A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 — Every major A

hardwaretechmeme
15 May 2026
Hardware

An Amortized Efficiency Threshold for Comparing Neural and Heuristic Solvers in Combinatorial Optimization

DGX agent

arXiv:2605.14624v1 Announce Type: cross Abstract: A common critique of neural combinatorial-optimization solvers is that they are less energy-efficient than CPU metaheuristics, given the operational e

hardwarearxiv-cs-ai
15 May 2026
Hardware

BEAM: Binary Expert Activation Masking for Dynamic Routing in MoE

DGX agent

arXiv:2605.14438v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enhance the efficiency of large language models by activating only a subset of experts per token. However, standa

hardwarearxiv-cs-ai
15 May 2026
Hardware

DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration

DGX agent

arXiv:2605.14526v1 Announce Type: cross Abstract: Differentiable simulation of soft bodies is a foundation for system identification, trajectory optimization, and Real2Sim transfer. Yet, existing meth

hardwarearxiv-cs-ro
15 May 2026
Hardware

EMA: Efficient Model Adaptation for Learning-based Systems

DGX agent

arXiv:2605.13942v1 Announce Type: new Abstract: Machine learning (ML) is increasingly applied to optimize system performance in tasks such as resource management and network simulation. Unlike traditi

hardwarearxiv-cs-lg
15 May 2026
Hardware

EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization

DGX agent

arXiv:2605.14249v1 Announce Type: new Abstract: We present EnergyLens, an end-to-end framework for energy-aware large language model (LLM) inference optimization. As LLMs scale, predicting and reducin

hardwarearxiv-cs-lg
15 May 2026
← Previous
1…2223242526…37
Next →