AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
Hardware

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. T…

DGX agent

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. This is what happens when world-class AI infrastructure meets

hardwareelon-musk--x
8 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

DBNN: Neural Spike Classification Using a Deep Binarized Neural Network

DGX agent

arXiv:2607.05590v1 Announce Type: cross Abstract: Implantable brain-computer interfaces require on-node spike sorting to reduce telemetry bandwidth and power while maintaining reliable neural decoding

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression

DGX agent

arXiv:2607.06523v1 Announce Type: new Abstract: Long-context language model inference is increasingly limited by the memory bandwidth and capacity required to store key-value caches, yet existing comp

hardwarearxiv-cs-ai
8 Jul 2026
Hardware

Design-CP: Context Parallelism for Design of Protein Nanoparticles

DGX agent

arXiv:2607.05439v1 Announce Type: new Abstract: Many all-atom generative protein models can in principle design large multimeric complexes by jointly modelling all chains, but their quadratic token- a

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

HJCD-IK: GPU-Accelerated Inverse Kinematics through Batched Hybrid Jacobian Coordinate Descent

DGX agent

arXiv:2510.07514v2 Announce Type: replace Abstract: Inverse Kinematics (IK) is a core problem in robotics, in which joint configurations are found to achieve a (collision-free) desired end-effector po

hardwarearxiv-cs-ro
8 Jul 2026
Hardware

Inference chip startup SambaNova valued at 11B in 1B funding round

DGX agent

Chip startup SambaNova Inc. today announced that it has raised 1 billion in funding at a 11 billion valuation. General Atlantic led the Series F round with contributions from more than a dozen others.

hardwaresiliconangle
8 Jul 2026
Hardware

InfluMatch: Frontier-Quality KOL Search at 4B-Model Cost

DGX agent

arXiv:2607.05968v1 Announce Type: cross Abstract: Matching influencers (KOLs) to free-form, multi-part Thai marketing criteria is today served either by keyword search over structured profiles, which

hardwarearxiv-cs-ai
8 Jul 2026
Hardware

Life Cycle Assessment of Pre-training the Lucie 7B Open-Source Large Language Model on the Jean Zay Supercomputer

DGX agent

arXiv:2607.05408v1 Announce Type: cross Abstract: The environmental impact of training large language models (LLMs) is increasingly scrutinised, yet most published estimates focus on operational energ

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory

DGX agent

arXiv:2607.05511v1 Announce Type: new Abstract: Agentic video understanding equips models with long-term memory to autonomously process and respond to continuous, long-horizon multimodal streams. Howe

hardwarearxiv-cs-cv
8 Jul 2026
Hardware

Nvidia and AI chip startup d-Matrix are combining their hardware in a new system to power AI models, the latest example of Nvidia partnering with rivals (Phoebe Liu/The Information)

DGX agent

Phoebe Liu / The Information: Nvidia and AI chip startup d-Matrix are combining their hardware in a new system to power AI models, the latest example of Nvidia partnering with rivals — Nvidia has a pl

hardwaretechmeme
8 Jul 2026
Hardware

Nvidia lost ~$1T in market value in less than two months, dropping 16% from its May all-time-high, trading at 18x forward earnings, its lowest level since 2019 (Bloomberg)

DGX agent

Bloomberg: Nvidia lost ~1T in market value in less than two months, dropping 16% from its May all-time-high, trading at 18x forward earnings, its lowest level since 2019 — After losing roughly 1 trill

hardwaretechmeme
8 Jul 2026
Hardware

Open, convenient and predictable: Introducing Provisioned Throughput

DGX agent

Provisioned Throughput gives you reserved inference capacity for frontier open models like MiniMax M3 and GLM-5.2. Token-based pricing, a 99% uptime SLA, and up to 90% lower cost than proprietary APIs

hardwaretogether-ai-blog
8 Jul 2026
Hardware

RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation

DGX agent

arXiv:2607.06558v1 Announce Type: new Abstract: Scaling robot learning requires massive, diverse trajectory data, yet collection is currently bottlenecked by physical teleoperation, where every demons

hardwarearxiv-cs-ro
8 Jul 2026
Hardware

Sources: China plans to let some of its biggest AI companies buy a small number of Nvidia's H200 chips to offset a domestic computing shortage in recent months (Qianer Liu/The Information)

DGX agent

Qianer Liu / The Information: Sources: China plans to let some of its biggest AI companies buy a small number of Nvidia's H200 chips to offset a domestic computing shortage in recent months — China pl

hardwaretechmeme
8 Jul 2026
Hardware

Sources: Reno-based AI chip startup Positron is in talks to raise ~750M in two phases, at valuations of 3.5B in the first tranche and ~$5B in the second (Bloomberg)

DGX agent

Bloomberg: Sources: Reno-based AI chip startup Positron is in talks to raise ~750M in two phases, at valuations of 3.5B in the first tranche and ~5B in the second — AI chip startup Positron is in talk

hardwaretechmeme
8 Jul 2026
Hardware

SUSE, NVIDIA, and Vultr Simplify the Path from AI Pilots to Production

DGX agent

SUSE, NVIDIA, and Vultr have partnered to create a streamlined solution for deploying AI applications from pilot phase to production environments. The collaboration leverages SUSE's AI Factory framewo

hardwarevultr
8 Jul 2026
Hardware

Tensordyne targets AI inference market with logarithmic math and Juniper-derived rack architecture

DGX agent

The race to serve AI inference faster and cheaper is exposing the hard limits of conventional chip architecture. As demand for real-time AI responses accelerates, the industry’s standard response — st

hardwaresiliconangle
8 Jul 2026
Hardware

Tuning-Free Latent Diffusion Models for Ultrahigh-Resolution Image Editing

DGX agent

arXiv:2607.06136v1 Announce Type: new Abstract: Recent diffusion-based generative models have shown impressive performance in image generation and editing. However, due to memory limitations and the h

hardwarearxiv-cs-cv
8 Jul 2026
Hardware

UBEP: Re-architecting Expert Parallelism Communication Library for Production Superpods

DGX agent

arXiv:2607.06202v1 Announce Type: cross Abstract: The deployment of Mixture-of-Experts (MoE) models on production high-bandwidth superpods, such as NVIDIA's NVL72/576 and Huawei's CloudMatrix384, intr

hardwarearxiv-cs-ai
8 Jul 2026
Hardware

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a han…

DGX agent

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a handful of labs. RL changes who can build frontier AI and just wo

hardwareswyx--x
8 Jul 2026
Hardware

We're releasing ZML/LLMD, our homegrown LLM server built on top of our homegrown high performance heterogeneous inference stack. It ships wi…

DGX agent

We're releasing ZML/LLMD, our homegrown LLM server built on top of our homegrown high performance heterogeneous inference stack. It ships with 5 architectures out of the box: NVIDIA, AMD, Metal, Intel

hardwareyann-lecun--x
8 Jul 2026
Hardware

A Deep Learning-based surrogate model for Severe Accidents in nuclear reactors using ASTEC

DGX agent

arXiv:2607.04450v1 Announce Type: cross Abstract: Integral codes like the Accident Source Term Evaluation Code (ASTEC) are powerful tools to study the physics of Severe Accidents (SAs) in nuclear reac

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

A Multi-Task Deep Learning Framework for Real-Time Intelligent Video Surveillance with Temporal Event Validation

DGX agent

arXiv:2607.03131v1 Announce Type: cross Abstract: Modern video surveillance systems generate far more video streams than human operators can effectively monitor, making automated analysis essential fo

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

AI debt financing is on an insane trajectory, per @SemiAnalysis_:

DGX agent

AI debt financing is on an insane trajectory, per @SemiAnalysis_: Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters Over 7T AI debt by 2029, There can be no N

hardwaregary-marcus--x
7 Jul 2026
Hardware

AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters

DGX agent

Max single-threaded CPUs at scale are a new category of CPUs built for the agentic AI era. Across the creation and deployment of an agentic system, the CPU is on the critical path for reasoning, respo

hardwarenvidia-blog
7 Jul 2026
Hardware

AquaGen: Scaling generative models to molecular dynamics precision on thousands of atoms

DGX agent

arXiv:2607.03513v1 Announce Type: cross Abstract: We present AquaGen, the first all-atom, explicit solvent, periodic-boundary-condition-aware generative model that produces molecular configurations fr

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

CPR: Chained Perceptual Refinement for Coarse-to-Fine Medical Image Classification

DGX agent

arXiv:2607.02591v1 Announce Type: new Abstract: High resolution medical images contain fine grained, spatially sparse cues that are critical for diagnosis, yet preserving full resolution incurs substa

hardwarearxiv-cs-cv
7 Jul 2026
Hardware

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

DGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Fast Asymptotically Optimal Kinodynamic Planning via Vectorization

DGX agent

arXiv:2607.03987v1 Announce Type: new Abstract: Sampling-based motion planners have been shown to be effective for systems with complex kinodynamic constraints and high dimensionality. However, these

hardwarearxiv-cs-ro
7 Jul 2026
Hardware

Generative wave propagator

DGX agent

arXiv:2607.04440v1 Announce Type: cross Abstract: Seismic wavefield simulation is fundamental to seismology, but conventional finite-difference (FD) methods remain limited by numerical dispersion and

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Graph Sparse Sampling: Breaking the Curse of the Horizon in Continuous MDP Planning

DGX agent

arXiv:2607.05359v1 Announce Type: new Abstract: Planning under uncertainty in continuous domains is essential for autonomous systems, yet computationally demanding. Tree-based search methods such as M

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion

DGX agent

arXiv:2607.03012v1 Announce Type: cross Abstract: Video Diffusion Transformers (VDiTs) have demonstrated significant capabilities in high-fidelity video generation. However, their ability to produce l

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Lights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy Consumption

DGX agent

arXiv:2607.04553v1 Announce Type: cross Abstract: We present a bidirectional framework for estimating the energy consumption of text-to-video (T2V) and text-to-video-audio (T2VA) models from architect

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

MLSYSIM: First-Principles Infrastructure Modeling for Machine Learning Systems

DGX agent

arXiv:2607.02558v1 Announce Type: cross Abstract: As machine learning shifts from laboratory curiosity to critical infrastructure, the systems that sustain it span an extraordinary range, from sub-mil

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community

DGX agent

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development can still be gated b

hardwarenvidia-blog
7 Jul 2026
Hardware

NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads

DGX agent

NVIDIA Vera is a CPU designed to help AI factories scale agentic AI and reinforcement learning by shortening CPU execution time, increasing task throughput, and enabling smarter, longer-thinking agent

hardwarenvidia-developer
7 Jul 2026
Hardware

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

DGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization

DGX agent

arXiv:2607.02834v1 Announce Type: new Abstract: Molecular optimization often starts from a pretrained generative model that captures a broad prior over valid molecular structures. At test time, howeve

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the lates…

DGX agent

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the latest Isaac models and frameworks to the open robotics community.

hardwareclem-delangue--x
7 Jul 2026
Hardware

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

DGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

RoMa v2: Harder Better Faster Denser Feature Matching

DGX agent

arXiv:2511.15706v3 Announce Type: replace Abstract: Dense feature matching aims to estimate all correspondences between two images of a 3D scene and has recently been established as the gold standard

hardwarearxiv-cs-cv
7 Jul 2026
Hardware

Scaling Weisfeiler-Leman Expressiveness Analysis to Massive Graphs with GPUs

DGX agent

arXiv:2607.02603v1 Announce Type: cross Abstract: The stable coloring of the Weisfeiler-Leman (1-WL) test is a cornerstone of Graph Neural Networks because it provides an upper bound to the expressive

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

SILO: Simulation-in-the-Loop Sim-to-Real Transfer for Multi-Stage Cable Routing

DGX agent

arXiv:2607.04616v1 Announce Type: cross Abstract: Linear-deformable manipulation remains challenging due to the complex deformations of objects such as cables and ropes. Prior data-driven approaches,

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

DGX agent

arXiv:2607.03333v1 Announce Type: cross Abstract: LLM agents are becoming a common interface for research, coding, and question answering, yet their Thought-Action-Observation loop is often serial: th

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

TabPack: Efficient Hyperparameter Ensembles for Tabular Deep Learning

DGX agent

arXiv:2607.05380v1 Announce Type: new Abstract: In deep learning for tabular data, efficient ensembles of multilayer perceptrons (MLPs) have recently emerged as effective and practical architectures.

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a w…

DGX agent

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a welcoming audience for my workshops and presentation. Thanks

hardwareswyx--x
7 Jul 2026
Hardware

The S-ICDF Dataset: Sionna-Simulated Dynamic Interference Characterization and Direction Finding

DGX agent

arXiv:2607.03411v1 Announce Type: cross Abstract: Jamming and spoofing threaten wireless and satellite navigation by disrupting or manipulating radio frequency (RF) signals, undermining availability,

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kerne…

DGX agent

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kernels all working together to make inference faster for users. P

hardwaretogether-ai--x
7 Jul 2026
← Previous
1…7891011…37
Next →