AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,734 results
Hardware

BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding

DGX agent

arXiv:2606.00144v1 Announce Type: cross Abstract: Speculative decoding speeds up autoregressive decoding by using a drafter to propose multiple tokens that a verifier validates in parallel. In resourc

hardwarearxiv-cs-ai
2 Jun 2026
Hardware
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cohort-Scale Neural Atlases of Ultrasound Video

DGX agent

arXiv:2606.00890v1 Announce Type: new Abstract: Ultrasound is the most widely used real-time imaging modality in clinical practice, yet per-frame video annotation remains a major bottleneck: expert la

hardwarearxiv-cs-cv
2 Jun 2026
Hardware

CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving

DGX agent

arXiv:2603.28768v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has recently emerged as the mainstream architecture for efficiently scaling large language models while maintaining n

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Deploy Agentic-Ready AI at the Edge with Memory Efficiency in NVIDIA JetPack 7.2

DGX agent

NVIDIA JetPack 7.2 enables deployment of AI agents to edge devices with optimized memory and performance for real-world applications. The release directly supports one-command deployment of NVIDIA Nem

hardwarenvidia-developer
2 Jun 2026
Hardware

Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap

DGX agent

arXiv:2512.10236v2 Announce Type: replace-cross Abstract: Modern ML workloads demand distributing training and inference across multiple GPUs. However, these parallelization techniques often suffer fr

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Don't Let a Few Network Failures Slow the Entire AllReduce

DGX agent

arXiv:2606.01680v1 Announce Type: cross Abstract: Network failures are among the most frequent hardware faults in large-scale GPU clusters and a leading cause of training-job interruptions. Modern col

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing

DGX agent

arXiv:2511.04791v2 Announce Type: replace Abstract: Modern LLM serving systems must sustain high throughput while meeting strict latency SLOs across two distinct inference phases: compute-intensive pr

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

DGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

hardwarearxiv-cs-cl
2 Jun 2026
Hardware

First technical Deepdive on M3 on the internet😎

DGX agent

First technical Deepdive on M3 on the internet😎 MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major sparse atte

hardwaretogether-ai--x
2 Jun 2026
Hardware

FLARE: Diffusion for Hybrid Language Model

DGX agent

arXiv:2606.01774v1 Announce Type: cross Abstract: Autoregressive (AR) large language models (LLMs) have achieved broad practical success, but sequential decoding remains a key bottleneck for low-laten

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

FreqLite: A Lightweight Frequency-Decomposed Linear Model with Adaptive Reversible Normalization for Robust Long-Term Time-Series Forecasting

DGX agent

arXiv:2606.01339v1 Announce Type: cross Abstract: Long-term time-series forecasting needs models that are accurate yet efficient enough for commodity hardware. Lightweight linear forecasters are remar

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Industrial Software Leaders Build Secure, Autonomous AI Engineers With NVIDIA NemoClaw

DGX agent

Accelerated computing has revolutionized industrial engineering, compressing simulation times from weeks to hours. Today’s remaining challenges sit in the end-to-end workflow surrounding the simulatio

hardwarenvidia-blog
2 Jun 2026
Hardware

Introducing workspaces for Lambda Cloud

DGX agent

Lambda workspaces help teams organize cloud resources, control access, and separate dev, staging, and production in shared GPU environments. A junior researcher kills a production training run. A cont

hardwarelambda-labs
2 Jun 2026
Hardware

Join the livestream to hear from our team members @karan4d and @yoniebans live from @nvidia! https://www.youtube.com/watch?v=pgQDbRMa2Eg

DGX agent

Nous Research is hosting a livestream featuring team members Karan and Yoni speaking from NVIDIA, likely discussing AI research, model development, or collaboration between Nous Research and NVIDIA. T

hardwarenous-research--x
2 Jun 2026
Hardware

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

DGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

LithoGRPO: Fast Inverse Lithography via GRPO Reinforced Flow Matching

DGX agent

arXiv:2606.00228v1 Announce Type: new Abstract: In semiconductor manufacturing, lithography projects circuit layouts onto silicon wafers through an optical mask. As circuit features shrink below the w

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Lodestar: An Online-Learning LLM Inference Router

DGX agent

arXiv:2606.00946v1 Announce Type: cross Abstract: Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for G

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Marvell shares closed up 32.52% on Tuesday after Jensen Huang hailed the chipmaker as the 'next trillion-dollar company' at Computex (Sawdah Bhaimiya/CNBC)

DGX agent

Sawdah Bhaimiya / CNBC: Marvell shares closed up 32.52% on Tuesday after Jensen Huang hailed the chipmaker as the “next trillion-dollar company” at Computex — Nvidia's CEO Jensen Huang hailed Marvell

hardwaretechmeme
2 Jun 2026
Hardware

Microsoft announces Surface RTX Spark AI supercomputer development box

DGX agent

Microsoft Corp. today announced it’s bringing agentic development capabilities into the hands of developers with a new desktop form factor supercomputer called the Surface RTX Spark Dev Box. Developed

hardwaresiliconangle
2 Jun 2026
Hardware

MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major…

DGX agent

MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major sparse attention, paged MSA decode, optimized index scoring

hardwaretogether-ai--x
2 Jun 2026
Hardware

NVIDIA Jetson Brings Agentic AI to the Physical World

DGX agent

Agentic AI is getting physical. At COMPUTEX on Tuesday, NVIDIA announced NVIDIA JetPack 7.2 and NVIDIA NemoClaw support on NVIDIA Jetson. JetPack 7.2 brings agentic AI skills, Yocto project support, N

hardwarenvidia-blog
2 Jun 2026
Hardware

Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving

DGX agent

arXiv:2606.01839v1 Announce Type: cross Abstract: LLM-based agents resolve a user task through many turns of dependent inference and tool calls, producing a workload whose total cost is unknown when t

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Open by Design: How NVIDIA and DigitalOcean Are Building the Stack for the Always-On Agentic Era

DGX agent

NVIDIA and DigitalOcean are collaborating to develop an open infrastructure stack designed to support autonomous AI agents that operate continuously. The initiative emphasizes open-source principles a

hardwaredigitalocean
2 Jun 2026
Hardware

Opus 4.8

DGX agent

Opus 4.8 likely refers to a software release, update, or feature announcement covered by Ben's Bites, a technology newsletter focused on AI and software developments. Without access to the specific ar

hardwareben-s-bites
2 Jun 2026
Hardware

PortBERT: Navigating the Depths of Portuguese Language Models

DGX agent

arXiv:2606.02100v1 Announce Type: new Abstract: Transformer models dominate modern NLP, but efficient, language-specific models remain scarce. In Portuguese, most focus on scale or accuracy, often neg

hardwarearxiv-cs-cl
2 Jun 2026
Hardware

Practical Aspects on Solving Differential Equations Using Deep Learning: A Primer

DGX agent

arXiv:2408.11266v5 Announce Type: replace Abstract: Deep learning is now common across many scientific fields, including the study of partial differential equations. This article provides a brief, acc

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

DGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Procurement records show at least seven Chinese universities that support the country's military and defense industry are seeking access to Nvidia's H200 chips (Bloomberg)

DGX agent

Bloomberg: Procurement records show at least seven Chinese universities that support the country's military and defense industry are seeking access to Nvidia's H200 chips — At least seven Chinese univ

hardwaretechmeme
2 Jun 2026
Hardware

Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation

DGX agent

arXiv:2601.22813v2 Announce Type: replace Abstract: The NVFP4 lower-precision format, supported in hardware by NVIDIA Blackwell GPUs, promises to allow, for the first time, end-to-end fully-quantized

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills

DGX agent

arXiv:2503.05641v4 Announce Type: replace-cross Abstract: Combining existing pre-trained LLMs is a promising approach for diverse reasoning tasks. However, task-level expert selection is often too coa

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

DGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

SUPREME: A Multi-GPU Framework for Reproducible Image Unlearning Method Evaluation

DGX agent

arXiv:2606.00380v1 Announce Type: cross Abstract: Machine unlearning removes the influence of specific training data from a trained model without retraining it from scratch. Evaluating an unlearning m

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVID…

DGX agent

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVIDIA NemoClaw and OpenShell, connect it to Slack, Outlook, Git

hardwarenous-research--x
2 Jun 2026
Hardware

The future of design doesn’t interrupt your process. @NousResearch Hermes keeps designers in creative flow - Rhino → ComfyUI → @Blender in o…

DGX agent

The future of design doesn’t interrupt your process. @NousResearch Hermes keeps designers in creative flow - Rhino → ComfyUI → @Blender in one seamless agent workflow. No context switching. Just visio

hardwarecomfyui--x
2 Jun 2026
Hardware

Threshold-Based Exclusive Batching for LLM Inference

DGX agent

arXiv:2606.00516v1 Announce Type: new Abstract: Mixed batching (MB)--interleaving prefill and decode in a single batch--has become the standard scheduling strategy for large language model (LLM) infer

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

ViBE: Co-Optimizing Workload Skew and Hardware Variability for MoE Serving

DGX agent

arXiv:2606.00735v1 Announce Type: cross Abstract: In distributed Mixture-of-Experts (MoE) inference, input-dependent token routing interacts with GPU performance variability to create persistent strag

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

$200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀

DGX agent

200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀 ANOTHER big AI Lab is massively investing in London. This time it is @runwayml

hardwarecristobal-valenzuela--x
1 Jun 2026
Hardware

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition

DGX agent

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition Introducing the Cosmos Coalition A new global initiative with NVIDIA and leading AI labs to build and open-source

hardwarecristobal-valenzuela--x
1 Jun 2026
Hardware

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent.

DGX agent

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent. This is the NVIDIA RTX Spark Superchip. A new beginning for personal computers. Designed for creators,

hardwarenous-research--x
1 Jun 2026
Hardware

Accelerate LLM model loading and increase context windows with GPUDirect on Amazon FSx for Lustre and TurboQuant

DGX agent

If you’re iterating on deploying large language models (LLMs) on AWS GPU instances, you’ve probably noticed the larger the model to be loaded into GPU High Bandwidth Memory (HBM), the longer the painf

hardwareaws-ml-blog
1 Jun 2026
Hardware

AI companies wouldn't enslave engineers, or steal chips from Nvidia, or illegally tap a power line. So why, asks @AGSNYT, do they feel entit…

DGX agent

AI companies wouldn't enslave engineers, or steal chips from Nvidia, or illegally tap a power line. So why, asks @AGSNYT, do they feel entitled to 'a brazen theft of intellectual property' from news o

hardwaregary-marcus--x
1 Jun 2026
Hardware

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning

DGX agent

arXiv:2605.31481v1 Announce Type: new Abstract: As robot control shifts toward large-scale reinforcement learning with in-loop dynamics computation, the community's reliance on CPU-bound libraries suc

hardwarearxiv-cs-ro
1 Jun 2026
Hardware

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning

DGX agent

arXiv:2603.09221v2 Announce Type: replace Abstract: Associative memory has long underpinned the design of sequential models. Beyond recall, humans reason by projecting future states and selecting goal

hardwarearxiv-cs-lg
1 Jun 2026
Hardware

Box says it created 13 new AI-focused roles, like AI architect and AI solutions manager, and plans to grow its staff to 3,000 by early 2027, up from 2,900 (Kalley Huang/New York Times)

DGX agent

Kalley Huang / New York Times: Box says it created 13 new AI-focused roles, like AI architect and AI solutions manager, and plans to grow its staff to 3,000 by early 2027, up from 2,900 — Box, a Silic

hardwaretechmeme
1 Jun 2026
Hardware

CoMo3R-SLAM: Collaborative Monocular Dense SLAM with Learned 3D Reconstruction Priors for Outdoor Multi-Agent Systems

DGX agent

arXiv:2605.30488v1 Announce Type: new Abstract: Collaborative dense SLAM is essential for multi-robot teams to achieve scalable and consistent 3D perception across large-scale outdoor environments. Ex

hardwarearxiv-cs-ro
1 Jun 2026
Hardware

Cross-Entropy Optimization of Physically Grounded Task and Motion Plans

DGX agent

arXiv:2512.11571v2 Announce Type: replace Abstract: Autonomously performing tasks often requires robots to plan high-level discrete actions and continuous low-level motions to realize them. Previous T

hardwarearxiv-cs-ro
1 Jun 2026
Hardware

Deep-learning-based low-energy trigger algorithms for the Hyper-Kamiokande experiment

DGX agent

arXiv:2605.31391v1 Announce Type: cross Abstract: Modern machine learning techniques have become increasingly important in particle physics because of their powerful pattern-recognition capabilities,

hardwarearxiv-cs-lg
1 Jun 2026
Hardware

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

DGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

hardwarearxiv-cs-cl
1 Jun 2026
← Previous
1…1617181920…37
Next →