AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
Hardware

CubicQuant: Parametric Non-Uniform Codebooks for High-Throughput LLM Inference with 1-8-Bit Weights

DGX agent

arXiv:2608.06763v1 Announce Type: new Abstract: Weight quantization for large-language-model inference must balance adaptive reconstruction levels with representations regular enough for efficient GPU

hardwarearxiv-cs-lg
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Fast LapSum: Exact Differentiable Top-k at Million Scale

DGX agent

arXiv:2608.06912v1 Announce Type: new Abstract: The top-k operation is a fundamental building block of modern sparse computation, enabling token routing, expert activation, memory selection, and atten

hardwarearxiv-cs-ai
10 Aug 2026
Hardware

If Anthropic begin shipping chips, will NVIDIA begin shipping frontier models? Who will win?

DGX agent

On August 10 2026 at 1:39 AM UTC, user Itamar Friedman (@itamar_mar) posted a short tweet asking whether Anthropic’s potential launch of its own chips would prompt NVIDIA to release frontier AI models

hardwareitamar-friedman--x
10 Aug 2026
Hardware

Jensen Huang just told the incredible story of how Elon Musk became NVIDIA’s first customer for its AI supercomputer - when literally nobody…

DGX agent

Jensen Huang just told the incredible story of how Elon Musk became NVIDIA’s first customer for its AI supercomputer - when literally nobody else wanted it “When I announced this thing, nobody wanted

hardwareelon-musk--x
10 Aug 2026
Hardware

Meta says Muse Glimmer has 30B parameters and is 'small enough' to need only one GPU; Meta plans a $1B fund to invest in US communities near its data centers (Vlad Savov/Bloomberg)

DGX agent

Vlad Savov / Bloomberg: Meta says Muse Glimmer has 30B parameters and is “small enough” to need only one GPU; Meta plans a $1B fund to invest in US communities near its data centers — Meta Platforms I

hardwaretechmeme
10 Aug 2026
Hardware

Nvidia partners with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR on a $500B funding package for AI infrastructure development (Financial Times)

DGX agent

Financial Times: Nvidia partners with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR on a $500B funding package for AI infrastructure development — Apollo, Blackstone and Goldman Sa

hardwaretechmeme
10 Aug 2026
Hardware

Scalable High-Fidelity Macromolecular Docking for GPU-Accelerated Supercomputers

DGX agent

arXiv:2608.07078v1 Announce Type: cross Abstract: Flexible macromolecular docking offers high-fidelity predictions of biomolecular interactions, but remains prohibitively expensive at scale. Among exi

hardwarearxiv-cs-ai
10 Aug 2026
Hardware

Sources: AI cloud computing provider Lambda is selling a $917M leveraged loan to finance the purchase of GPUs as part of a contract with Nvidia (Bloomberg)

DGX agent

Bloomberg: Sources: AI cloud computing provider Lambda is selling a $917M leveraged loan to finance the purchase of GPUs as part of a contract with Nvidia — Lambda Inc., an AI cloud-computing provider

hardwaretechmeme
10 Aug 2026
Hardware

Summarize First, Download Later: Onboard VLMs for Bandwidth-Efficient Earth Observation

DGX agent

arXiv:2608.06959v1 Announce Type: new Abstract: Modern Earth observation (EO) satellites carry increasingly advanced sensors that produce vast volumes of high-resolution, multispectral data, yet downl

hardwarearxiv-cs-cv
10 Aug 2026
Hardware

Ultra-High Interactivity on NVIDIA GPUs? - TileRT InferenceX

DGX agent

The TileRT InferenceX article (Aug 10 2026) examines whether the TileRT software stack on NVIDIA GPUs can compete with dedicated inference systems such as Cerebras, Groq LPUs and SambaNova for ultra‑h

hardwaresemianalysis
10 Aug 2026
Hardware

Ultra-High Interactivity on NVIDIA GPUs? TileRT InferenceX Can TileRT software on NVIDIA GPU compete with Cerebras, Groq LPU, SambaNova? Bat…

DGX agent

Ultra-High Interactivity on NVIDIA GPUs? TileRT InferenceX Can TileRT software on NVIDIA GPU compete with Cerebras, Groq LPU, SambaNova? Batch Size 1, Disaggregated engine, High throughput prefill eng

hardwaredylan-patel--x
10 Aug 2026
Hardware

300b on 32gb MoE-streaming findings + optimisations

DGX agent

The past week I've been running DSv4 inference on my laptop by keeping everything RAM-resident except the MXFP4-experts (since expert pool is ~147GB and won't fit) TL;DR - read speed is the limiter mo

hardwarer-localllama
9 Aug 2026
Hardware

Open Model: Google Weather Next 2

DGX agent

I am not a meteorologist, but I just read a very interesting article: https://arstechnica.com/science/2026/08/deepminds-hurricane-model-bought-forecasters-an-extra-day/ In a paper published on Thursda

hardwarer-localllama
9 Aug 2026
Hardware

CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning

DGX agent

arXiv:2512.02551v4 Announce Type: replace-cross Abstract: In this paper, we propose CUDA-L2, a system that combines large language models (LLMs) and reinforcement learning (RL) to automatically optimi

hardwarearxiv-cs-ai
7 Aug 2026
Hardware

IcFuzz: Fuzzing Isaac Sim with Semantic Stage Guidance and Multi-level Mutation

DGX agent

arXiv:2608.06088v1 Announce Type: new Abstract: Robotics simulators serve as a foundational infrastructure for embodied AI, facilitating safe and scalable robotic system development. NVIDIA Isaac Sim

hardwarearxiv-cs-ro
7 Aug 2026
Hardware

Learning to Rank Tensor Network Contraction Plans for GPU-Accelerated Quantum Circuit Simulation

DGX agent

arXiv:2608.05819v1 Announce Type: new Abstract: Classical simulation remains essential for developing and validating quantum algorithms, but its cost grows rapidly with circuit size. Tensor-network co

hardwarearxiv-cs-lg
7 Aug 2026
Hardware

PaDoc: Layout-Grounded Parallel Decoding for Document Parsing

DGX agent

arXiv:2608.06146v1 Announce Type: new Abstract: End-to-end document parsers provide a unified interface, but serialize page layouts and regional contents into one autoregressive sequence. This formula

hardwarearxiv-cs-ai
7 Aug 2026
Hardware

Parallel-in-Time Nonlinear Optimal Control via GPU-native Sequential Convex Programming

DGX agent

arXiv:2603.10711v3 Announce Type: replace Abstract: Real-time solution of nonlinear optimal control problems remains challenging on embedded robotic hardware, where conventional solvers often rely on

hardwarearxiv-cs-ro
7 Aug 2026
Hardware

Sources: Nvidia agrees to invest 2B in Lancium, the power infrastructure developer of the Stargate campus in Texas, plus 1B more if it hits certain thresholds (The Information)

DGX agent

The Information: Sources: Nvidia agrees to invest 2B in Lancium, the power infrastructure developer of the Stargate campus in Texas, plus 1B more if it hits certain thresholds — Nvidia has agreed to i

hardwaretechmeme
7 Aug 2026
Hardware

Sources: the US Commerce Department's BIS is reviewing how Chinese AI companies access Nvidia chips overseas, including by legally renting foreign data centers (Mackenzie Hawkins/Bloomberg)

DGX agent

Mackenzie Hawkins / Bloomberg: Sources: the US Commerce Department's BIS is reviewing how Chinese AI companies access Nvidia chips overseas, including by legally renting foreign data centers — A key U

hardwaretechmeme
7 Aug 2026
Hardware

AI-based single-shot structured-light depth reconstruction for real-time laparoscopic surgical guidance

DGX agent

arXiv:2608.05109v1 Announce Type: cross Abstract: Significance. Accurate intraoperative depth perception is important for autonomous and semi-autonomous robotic laparoscopic surgery. Conventional frin

hardwarearxiv-cs-ro
6 Aug 2026
Hardware

AMD acquires Taalas to hardwire AI models into silicon

DGX agent

Advanced Micro Devices Inc. said today it has agreed to buy Taalas Inc., a Toronto startup that hardwires artificial intelligence models directly into silicon, in a deal that pushes the chipmaker deep

hardwaresiliconangle
6 Aug 2026
Hardware

AMD acquires Toronto-based Taalas, which integrates model weights directly into silicon to boost inference performance, for an undisclosed sum (Tobias Mann/The Register)

DGX agent

Tobias Mann / The Register: AMD acquires Toronto-based Taalas, which integrates model weights directly into silicon to boost inference performance, for an undisclosed sum — Early tech demos show model

hardwaretechmeme
6 Aug 2026
Hardware

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

DGX agent

arXiv:2608.04041v1 Announce Type: new Abstract: Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classific

hardwarearxiv-cs-lg
6 Aug 2026
Hardware

ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing

DGX agent

arXiv:2608.04956v1 Announce Type: new Abstract: Recent video models increasingly support generation, reference conditioning, and editing within a single model, yet typically expose them as separate op

hardwarearxiv-cs-cv
6 Aug 2026
Hardware

Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs (Anna Tong/Forbes)

DGX agent

Anna Tong / Forbes: Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs — The same Silicon Va

hardwaretechmeme
6 Aug 2026
Hardware

GASP: GPU-Accelerated Safe Planner for Real-Time Collision-Aware Motion Generation with Latent Trajectory Sampling

DGX agent

arXiv:2608.04612v1 Announce Type: new Abstract: We present GASP, a GPU-Accelerated Safe Planner for real-time, collision-aware joint-space motion generation in known environments. GASP combines a clam

hardwarearxiv-cs-ro
6 Aug 2026
Hardware

Gaussian-LIC2: LiDAR-Inertial-Camera Gaussian Splatting SLAM

DGX agent

arXiv:2507.04004v3 Announce Type: replace Abstract: This paper presents the first photo-realistic LiDAR-Inertial-Camera Gaussian Splatting SLAM system that simultaneously addresses visual quality, geo

hardwarearxiv-cs-ro
6 Aug 2026
Hardware

Hugging Face Storage Buckets are now on http://Vast.ai Connect your HF Storage Bucket as a Cloud Connection in your Vast settings, and every…

DGX agent

Hugging Face Storage Buckets are now on http://Vast.ai Connect your HF Storage Bucket as a Cloud Connection in your Vast settings, and every GPU instance you rent can pull datasets and checkpoints str

hardwareclem-delangue--x
6 Aug 2026
Hardware

Into the Omniverse: How Open World Models Push the Frontier of Physical AI

DGX agent

In July, NVIDIA joined more than 200 companies and organizations in signing “Open Weights and American AI Leadership,” an open letter arguing that AI leadership will be measured not by any single fron

hardwarenvidia-blog
6 Aug 2026
Hardware

Ooredoo, Nvidia, Nokia, and Indosat launch Zankore, Indonesia's first dedicated AI compute and neocloud platform; Ooredoo, which owns a 49% stake, commits $800M (Melissa Hancock/Fortune)

DGX agent

Melissa Hancock / Fortune: Ooredoo, Nvidia, Nokia, and Indosat launch Zankore, Indonesia's first dedicated AI compute and neocloud platform; Ooredoo, which owns a 49% stake, commits $800M — Ooredoo Gr

hardwaretechmeme
6 Aug 2026
Hardware

RORA: Realistic Object Reconstruction with Articulation

DGX agent

arXiv:2608.04842v1 Announce Type: new Abstract: Replicating real-world environments into simulation by realistic visual representation like NeRF and 3D Gaussian Splatting (3DGS) has emerged as an effe

hardwarearxiv-cs-ro
6 Aug 2026
Hardware

SparseDitto: Customizing GPU Kernels for Different Sparsity Patterns with LLM-Based Agentic System

DGX agent

arXiv:2608.05033v1 Announce Type: cross Abstract: Sparse matrix kernels are fundamental to scientific computing, graph analytics, and machine learning. Their GPU performance depends strongly on the in

hardwarearxiv-cs-lg
6 Aug 2026
Hardware

StaticSegFormer: An Efficient High-Performance Semantic Segmentation Based on Static Structured Pruning

DGX agent

arXiv:2608.04811v1 Announce Type: new Abstract: Structured pruning enhances the efficiency of deep neural networks (DNNs) by eliminating groups of parameters during inference. Previous methods mostly

hardwarearxiv-cs-cv
6 Aug 2026
Hardware

Sydney-based AI data center company Firmus raised 2B in funding from Coatue, Nvidia, and others at a 10.5B post-money valuation, up from $5.5B in April (Nichiket Sunil/Reuters)

DGX agent

Nichiket Sunil / Reuters: Sydney-based AI data center company Firmus raised 2B in funding from Coatue, Nvidia, and others at a 10.5B post-money valuation, up from 5.5B in April — Firmus said on Friday

hardwaretechmeme
6 Aug 2026
Hardware

A Deployment-Friendly Foundational Framework for Efficient Computational Pathology

DGX agent

arXiv:2602.14010v2 Announce Type: replace-cross Abstract: Pathology foundation models (PFMs) generalize well across computational pathology tasks but remain costly for gigapixel whole-slide image anal

hardwarearxiv-cs-ai
5 Aug 2026
Hardware

Accelerating Dynamic Graph Clustering on GPU Architectures with cuGraph

DGX agent

arXiv:2608.03695v1 Announce Type: cross Abstract: This work addresses community detection in temporal networks through GPU-accelerated extensions of spectral clustering and modularity-based algorithms

hardwarearxiv-cs-lg
5 Aug 2026
Hardware

AcceptMoE: Commitment-Weighted Self-Sizing Verifier Expert Sets for Efficient MoE Speculative Decoding

DGX agent

arXiv:2608.02989v1 Announce Type: cross Abstract: Speculative decoding verifies a tree of draft tokens in one target-model forward pass. For a mixture-of-experts (MoE) target, however, parallel verifi

hardwarearxiv-cs-cl
5 Aug 2026
Hardware

Compound and Parallel Modes of Tropical Convolutional Neural Networks

DGX agent

arXiv:2504.06881v2 Announce Type: replace-cross Abstract: Convolutional neural networks (CNNs) are foundational to many state-of-the-art computer vision systems, yet their reliance on multiplication-i

hardwarearxiv-cs-ai
5 Aug 2026
Hardware

Learning and Clustering on Temporal Graphs: Principles, Primitives, and Pooling

DGX agent

arXiv:2608.03696v1 Announce Type: new Abstract: This work focuses on the problem of learning on temporal graphs, with particular emphasis on the task of clustering: obtaining coarse-grained representa

hardwarearxiv-cs-lg
5 Aug 2026
Hardware

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

DGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

hardwarearxiv-cs-ai
5 Aug 2026
Hardware

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

DGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

hardwarearxiv-cs-ai
5 Aug 2026
Hardware

RTX 3090 MiniMax H3 Speed Comparison: FP8 Scaled vs INT8 ConvRot (W8A8)

DGX agent

Setup: GPU: RTX 3090 24GB RAM: 32GB ComfyUI 0.30.0 PyTorch 2.13.0+cu130 CUDA 13.0 SageAttention enabled (sageattention-2.2.0+cu130torch2.10.0andhigher.post6-cp310-abi3-win_amd64) Spectrum node + Euler

hardwarer-stablediffusion
5 Aug 2026
Hardware

A Fortran General-Purpose Transpiler: Proof of Concept

DGX agent

arXiv:2608.00130v1 Announce Type: cross Abstract: Fortran has been the cornerstone of high-performance computing for decades and remains unmatched in many domains. Yet the language faces an expertise

hardwarearxiv-cs-cl
4 Aug 2026
Hardware

AI cloud startup Volta Infra raised 300M led by a16z and Altimeter at a 2.4B valuation and says it landed a $10B contract with an unnamed leading AI developer (Dina Bass/Bloomberg)

DGX agent

Dina Bass / Bloomberg: AI cloud startup Volta Infra raised 300M led by a16z and Altimeter at a 2.4B valuation and says it landed a 10B contract with an unnamed leading AI developer — Volta Infra Holdi

hardwaretechmeme
4 Aug 2026
Hardware

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory…

DGX agent

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory compute closer to the stars. The next chapter of AI infrastr

hardwareelon-musk--x
4 Aug 2026
Hardware

As AI Increases Demands on Memory, Storage Steps Up

DGX agent

Surging AI demands are driving the need for massive datasets and context windows that burst past the confines of system memory. But rising needs aren’t met by simply adding more storage capacity. What

hardwarenvidia-blog
4 Aug 2026
Hardware

Bole: Efficient Tree Speculation for Hybrid-Attention Language Models

DGX agent

arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autore

hardwarearxiv-cs-cl
4 Aug 2026
← Previous
1234…37
Next →