AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
2 Jun 2026

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

HardwareDGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

Procurement records show at least seven Chinese universities that support the country's military and defense industry are seeking access to Nvidia's H200 chips (Bloomberg)

HardwareDGX agent

Bloomberg: Procurement records show at least seven Chinese universities that support the country's military and defense industry are seeking access to Nvidia's H200 chips — At least seven Chinese univ

Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation

Hardware

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2601.22813v2 Announce Type: replace Abstract: The NVFP4 lower-precision format, supported in hardware by NVIDIA Blackwell GPUs, promises to allow, for the first time, end-to-end fully-quantized

Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills

HardwareDGX agent

arXiv:2503.05641v4 Announce Type: replace-cross Abstract: Combining existing pre-trained LLMs is a promising approach for diverse reasoning tasks. However, task-level expert selection is often too coa

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

HardwareDGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

SUPREME: A Multi-GPU Framework for Reproducible Image Unlearning Method Evaluation

HardwareDGX agent

arXiv:2606.00380v1 Announce Type: cross Abstract: Machine unlearning removes the influence of specific training data from a trained model without retraining it from scratch. Evaluating an unlearning m

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVID…

HardwareDGX agent

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVIDIA NemoClaw and OpenShell, connect it to Slack, Outlook, Git

The future of design doesn’t interrupt your process. @NousResearch Hermes keeps designers in creative flow - Rhino → ComfyUI → @Blender in o…

HardwareDGX agent

The future of design doesn’t interrupt your process. @NousResearch Hermes keeps designers in creative flow - Rhino → ComfyUI → @Blender in one seamless agent workflow. No context switching. Just visio

Threshold-Based Exclusive Batching for LLM Inference

HardwareDGX agent

arXiv:2606.00516v1 Announce Type: new Abstract: Mixed batching (MB)--interleaving prefill and decode in a single batch--has become the standard scheduling strategy for large language model (LLM) infer

ViBE: Co-Optimizing Workload Skew and Hardware Variability for MoE Serving

HardwareDGX agent

arXiv:2606.00735v1 Announce Type: cross Abstract: In distributed Mixture-of-Experts (MoE) inference, input-dependent token routing interacts with GPU performance variability to create persistent strag

1 Jun 2026

$200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀

HardwareDGX agent

200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀 ANOTHER big AI Lab is massively investing in London. This time it is @runwayml

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition

HardwareDGX agent

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition Introducing the Cosmos Coalition A new global initiative with NVIDIA and leading AI labs to build and open-source

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent.

HardwareDGX agent

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent. This is the NVIDIA RTX Spark Superchip. A new beginning for personal computers. Designed for creators,

Accelerate LLM model loading and increase context windows with GPUDirect on Amazon FSx for Lustre and TurboQuant

HardwareDGX agent

If you’re iterating on deploying large language models (LLMs) on AWS GPU instances, you’ve probably noticed the larger the model to be loaded into GPU High Bandwidth Memory (HBM), the longer the painf

AI companies wouldn't enslave engineers, or steal chips from Nvidia, or illegally tap a power line. So why, asks @AGSNYT, do they feel entit…

HardwareDGX agent

AI companies wouldn't enslave engineers, or steal chips from Nvidia, or illegally tap a power line. So why, asks @AGSNYT, do they feel entitled to 'a brazen theft of intellectual property' from news o

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning

HardwareDGX agent

arXiv:2605.31481v1 Announce Type: new Abstract: As robot control shifts toward large-scale reinforcement learning with in-loop dynamics computation, the community's reliance on CPU-bound libraries suc

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning

HardwareDGX agent

arXiv:2603.09221v2 Announce Type: replace Abstract: Associative memory has long underpinned the design of sequential models. Beyond recall, humans reason by projecting future states and selecting goal

Box says it created 13 new AI-focused roles, like AI architect and AI solutions manager, and plans to grow its staff to 3,000 by early 2027, up from 2,900 (Kalley Huang/New York Times)

HardwareDGX agent

Kalley Huang / New York Times: Box says it created 13 new AI-focused roles, like AI architect and AI solutions manager, and plans to grow its staff to 3,000 by early 2027, up from 2,900 — Box, a Silic

CoMo3R-SLAM: Collaborative Monocular Dense SLAM with Learned 3D Reconstruction Priors for Outdoor Multi-Agent Systems

HardwareDGX agent

arXiv:2605.30488v1 Announce Type: new Abstract: Collaborative dense SLAM is essential for multi-robot teams to achieve scalable and consistent 3D perception across large-scale outdoor environments. Ex

Cross-Entropy Optimization of Physically Grounded Task and Motion Plans

HardwareDGX agent

arXiv:2512.11571v2 Announce Type: replace Abstract: Autonomously performing tasks often requires robots to plan high-level discrete actions and continuous low-level motions to realize them. Previous T

Deep-learning-based low-energy trigger algorithms for the Hyper-Kamiokande experiment

HardwareDGX agent

arXiv:2605.31391v1 Announce Type: cross Abstract: Modern machine learning techniques have become increasingly important in particle physics because of their powerful pattern-recognition capabilities,

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

HardwareDGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

Develop Physical AI Reasoning, World, and Action Models with NVIDIA Cosmos 3

HardwareDGX agent

NVIDIA Cosmos 3 is a frontier foundation model for physical AI that combines physical reasoning, world generation, and action generation within a single open model. The model uses a Mixture-of-Transfo

DSD-GS: Dynamic-Static Decomposition of Gaussian Splatting for Efficient and High-Fidelity Dynamic Scene Reconstruction

HardwareDGX agent

arXiv:2605.30863v1 Announce Type: new Abstract: Dynamic scene reconstruction and novel view synthesis are fundamental to next-generation visual intelligence applications such as virtual reality, robot

Elastic ViTs from Pretrained Models without Retraining

HardwareDGX agent

arXiv:2510.17700v2 Announce Type: replace Abstract: Vision foundation models achieve remarkable performance but are only available in a limited set of pre-determined sizes, forcing sub-optimal deploym

Five thoughts from Nvidia CEO Jensen Huang’s GTC Taipei 2026 keynote

HardwareDGX agent

Useful artificial intelligence has arrived, and if Nvidia Chief Executive Jensen Huang is right, it is about to reshape not only data centers but also the structure of the global economy and the tech

🆕Grok Imagine’s Video Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://www.latent.space/p/video…

HardwareDGX agent

🆕Grok Imagine’s Video Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://www.latent.space/p/video-agents @EthanHe_42, former @xai world model lead and @nvidia

How Cosmos 3 Helps Physical AI Think Before It Acts

HardwareDGX agent

NVIDIA Cosmos 3 uses a mixture-of-transformers architecture that pairs a reasoning transformer with a generation transformer, enabling it to understand object interactions and spatial-temporal relatio

https://x.com/huggingface/status/2061495792246915131

HardwareDGX agent

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost 60,000 followers

Intel: Our upcoming AI chip will be cheaper, run cooler than Nvidia, AMD options

HardwareDGX agent

Intel plans to ship its Crescent Island AI GPU in 2026, targeting Nvidia and AMD on cost and power efficiency. The chip uses cheaper LPDDR5X memory and is designed to run in air-cooled server racks ra

It's cool to see that the http://hf.co/blog has become progressively a major source of learning and news about AI for the community. Just lo…

HardwareDGX agent

It's cool to see that the http://hf.co/blog has become progressively a major source of learning and news about AI for the community. Just look at the recent content created & major announcements & tut

Jensen Huang says Anthropic, OpenAI, and SpaceX are among the first big users for Nvidia's new Vera CPUs, which are 1.8x faster for AI workloads than x86 chips (Ian King/Bloomberg)

HardwareDGX agent

Ian King / Bloomberg: Jensen Huang says Anthropic, OpenAI, and SpaceX are among the first big users for Nvidia's new Vera CPUs, which are 1.8x faster for AI workloads than x86 chips — Nvidia Corp. sai

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts

HardwareDGX agent

arXiv:2605.30359v1 Announce Type: cross Abstract: Generating high-performance GPU kernels remains challenging due to the need for both correctness and hardware-aware optimization. While large language

Neurosim: A Fast Simulator for Neuromorphic Robot Perception

HardwareDGX agent

arXiv:2602.15018v2 Announce Type: replace-cross Abstract: Neurosim is a fast, real-time, high-performance library for simulating sensors such as dynamic vision sensors, RGB cameras, depth sensors, and

NVIDIA AI Cloud Ecosystem Expands Worldwide to Meet Global AI Compute Demand

HardwareDGX agent

The NVIDIA AI Cloud ecosystem is accelerating the global buildout of AI factory infrastructure. Partners are expanding capacity to meet growing demand from enterprises, startups, nations, AI labs and

Nvidia announcing a 550B model wasn't on my bingo card They are now the strongest american open-source lab

HardwareDGX agent

Nvidia announced development of a 550 billion parameter language model, positioning itself as a leading open-source AI research organization competing with traditional academic and independent labs. T

Nvidia debuts RTX Spark processor for Windows laptops, compact desktops

HardwareDGX agent

Nvidia Corp. today introduced a system-on-chip designed to power Windows laptops and compact desktops. RTX Spark, as the processor is called, will roll out alongside Windows upgrades that will make th

NVIDIA DSX OS Delivers Open, Modular Software for Operating AI Factories at Scale

HardwareDGX agent

NVIDIA DSX OS is an open source, modular software platform designed for operating and scaling multi-tenant AI factories. It provides foundational technologies for building and operating AI factories,

NVIDIA Factory Operations Blueprint Gives Factories a New AI Brain

HardwareDGX agent

As factories move from isolated automation to plant-wide intelligence, manufacturers need AI systems that can connect live machine signals, quality systems, work instructions and operational alerts in

Nvidia gives developers the tool to build secure, autonomous AI workers that scale

HardwareDGX agent

Not content with just providing the infrastructure for the next generation of artificial intelligence agents, Nvidia Corp. is also providing the tools for developers to build them. At Nvidia GTC Taipe

Nvidia ramps up production of Vera Rubin, the foundation of the next generation of AI factories

HardwareDGX agent

Nvidia Corp. said early Monday at the Computex conference in Taipei that it’s gearing up the production of its forthcoming Vera Rubin platform, which is set to become the foundation of a new generatio

Nvidia releases Cosmos3-Super-Image2Video . 64B parametres

HardwareDGX agent

Cosmos3-Super-Image2Video is a 64B model for temporally coherent image-to-video generation . NVIDIA's Cosmos platform is designed to accelerate Physical AI development by enabling machines to understa

Nvidia releasesCosmos3-Super-Text2Image model . 64 billion paramteres

HardwareDGX agent

Cosmos3-Super-Text2Image is a 64 billion parameter omnimodal world model capable of generating high-quality images from text inputs as part of NVIDIA's Cosmos 3 foundation model platform. The model is

NVIDIA Vera CPU Sets a New Standard for Agentic Workloads in AI Factories

HardwareDGX agent

NVIDIA Vera is a purpose-built CPU for agentic AI and reinforcement learning, delivering twice the efficiency and 50% faster performance than traditional rack-scale CPUs. The processor helps AI factor

On Efficient Scaling of GNNs via IO-Aware Layers Implementations

HardwareDGX agent

arXiv:2605.31500v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) are bottlenecked by sparse, irregular memory access. Popular frameworks such as DGL and PyTorch Geometric support general

ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs

HardwareDGX agent

arXiv:2602.07721v3 Announce Type: replace-cross Abstract: KV-cache retrieval is essential for long-context LLM inference, yet existing methods struggle with distribution drift and high latency at scal

PithTrain: A Compact and Agent-Native MoE Training System

HardwareDGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

Reducing the GPU Memory Bottleneck with Lossless Compression for ML -- Extended

HardwareDGX agent

arXiv:2605.30728v1 Announce Type: new Abstract: Machine learning (ML) training and inference often process data sets far exceeding GPU memory capacity, forcing them to rely on PCIe for on-demand tenso

SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer

HardwareDGX agent

arXiv:2605.30409v1 Announce Type: cross Abstract: Real-time streaming video-to-video editing (V2V) is critical for interactive applications such as live broadcasting and gaming, yet it remains a formi

Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Generation

HardwareDGX agent

arXiv:2605.30716v1 Announce Type: cross Abstract: Generating clinically useful pathology reports for pathology cases from whole-slide images (WSIs) is challenging due to gigapixel resolution, long vis

SubsurfaceGen: Procedural Generation of Field-Scale Earth Models and Seismic Data

HardwareDGX agent

arXiv:2605.30541v1 Announce Type: new Abstract: Full waveform inversion (FWI) is the gold standard for subsurface imaging, with applications from carbon sequestration to energy and mineral exploration

Taiwan’s Industry Titans Turbocharge World’s AI Infrastructure Buildout With NVIDIA

HardwareDGX agent

Taiwan is home to more than 500 NVIDIA ecosystem partners. More than 1 million NVIDIA MGX rack components for NVIDIA Vera Rubin infrastructure come together in Taiwan, from across 25 factory sites. As

the decade of agents! @theemozilla

HardwareDGX agent

the decade of agents! @theemozilla We have been working closely with @nvidia to ensure Hermes Agent works smoothly on their new @NVIDIARTXSpark superchip and integrates with the new OpenShell runtime,

This could be Windows’ M1 moment — but expect it to cost a ton

HardwareDGX agent

Nvidia's announcement that it's getting into the consumer laptop chip space with RTX Spark is huge. Apple has proved for years that Arm-based chips can perform incredibly well while also delivering gr

This pod was an incredible gift to the community: not only our first pod about @xAI, but Ethan really indulged on all our questions on how t…

HardwareDGX agent

This pod was an incredible gift to the community: not only our first pod about @xAI, but Ethan really indulged on all our questions on how to train a SOTA Videogen world model, including specific area

Today we’re announcing London as Runway’s European HQ, and doubling down on our investment in the UK AI ecosystem. In under two years, our L…

HardwareDGX agent

Today we’re announcing London as Runway’s European HQ, and doubling down on our investment in the UK AI ecosystem. In under two years, our London team has become central to our research, including fou

Unbox one of NVIDIA's first co-packaged optics samples with us. See why we bet on CPO early.

HardwareDGX agent

When we design large GPU clusters, the network is no longer a background system. It's part of the compute envelope. At the 800G and NVIDIA GB300 NVL72 scale, the back-end fabric accounts for 86% of ne

We have been working closely with @nvidia to ensure Hermes Agent works smoothly on their new @NVIDIARTXSpark superchip and integrates with t…

HardwareDGX agent

We have been working closely with @nvidia to ensure Hermes Agent works smoothly on their new @NVIDIARTXSpark superchip and integrates with the new OpenShell runtime, which connects Hermes to @Microsof

We have worked with @nvidia to integrate their official Agent Skills catalog into the Hermes Skills Hub. These skills teach your agent how t…

HardwareDGX agent

We have worked with @nvidia to integrate their official Agent Skills catalog into the Hermes Skills Hub. These skills teach your agent how to use CUDA-X libraries, Omniverse and Physical AI workflows,

We 💚 open source

HardwareDGX agent

We 💚 open source So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost

← Previous
1…1314151617…29
Next →