AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
3 Jun 2026

Quobly raises $150M for its silicon spin qubit technology

HardwareDGX agent

French quantum computing startup Quobly SAS today announced that it has raised €130 million, or about 150 million, to commercialize its technology. The Series A round was led by publicly traded chipma

Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles

HardwareDGX agent

arXiv:2505.08222v3 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) offer a cost-effective solution for scientific missions such as underwater tracking. Reinforcement learning (RL) has

Speedrunning Tabular Foundation Model Pretraining

HardwareDGX agent

arXiv:2606.03681v1 Announce Type: new Abstract: Pretraining cost is a major bottleneck for research on tabular foundation models, slowing the iteration cycle for new architectures, priors, and optimiz

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

To Boldly Go: The Case for Space Datacenters

HardwareDGX agent

This article explores the potential viability and benefits of locating data centers in space rather than on Earth, likely discussing advantages such as reduced cooling costs due to the vacuum environm

Towards Compact Autonomous Driving Perception with Balanced Learning and Multi-sensor Fusion

HardwareDGX agent

arXiv:2606.02979v1 Announce Type: cross Abstract: We present a novel compact deep multi-task learning model to handle various autonomous driving perception tasks in one forward pass. The model perform

2 Jun 2026

APB-V: Accelerating Long-Video Understanding via Sequence-Parallelism-aware Approximate Attention

HardwareDGX agent

arXiv:2601.21444v2 Announce Type: replace-cross Abstract: The efficiency of long-video inference remains a critical bottleneck, mainly due to the dense computation in the prefill stage of Large Multim

b9471

Local AiDGX agent

B9471 is a build release of llama.cpp, an open-source C/C++ implementation for LLM inference that enables running large language models on consumer hardware with optimized performance. This intermedia

BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding

HardwareDGX agent

arXiv:2606.00144v1 Announce Type: cross Abstract: Speculative decoding speeds up autoregressive decoding by using a drafter to propose multiple tokens that a verifier validates in parallel. In resourc

Cohort-Scale Neural Atlases of Ultrasound Video

HardwareDGX agent

arXiv:2606.00890v1 Announce Type: new Abstract: Ultrasound is the most widely used real-time imaging modality in clinical practice, yet per-frame video annotation remains a major bottleneck: expert la

CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving

HardwareDGX agent

arXiv:2603.28768v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has recently emerged as the mainstream architecture for efficiently scaling large language models while maintaining n

Deploy Agentic-Ready AI at the Edge with Memory Efficiency in NVIDIA JetPack 7.2

HardwareDGX agent

NVIDIA JetPack 7.2 enables deployment of AI agents to edge devices with optimized memory and performance for real-world applications. The release directly supports one-command deployment of NVIDIA Nem

Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap

HardwareDGX agent

arXiv:2512.10236v2 Announce Type: replace-cross Abstract: Modern ML workloads demand distributing training and inference across multiple GPUs. However, these parallelization techniques often suffer fr

DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing

HardwareDGX agent

arXiv:2511.04791v2 Announce Type: replace Abstract: Modern LLM serving systems must sustain high throughput while meeting strict latency SLOs across two distinct inference phases: compute-intensive pr

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

HardwareDGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

First technical Deepdive on M3 on the internet😎

HardwareDGX agent

First technical Deepdive on M3 on the internet😎 MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major sparse atte

Industrial Software Leaders Build Secure, Autonomous AI Engineers With NVIDIA NemoClaw

HardwareDGX agent

Accelerated computing has revolutionized industrial engineering, compressing simulation times from weeks to hours. Today’s remaining challenges sit in the end-to-end workflow surrounding the simulatio

Introducing workspaces for Lambda Cloud

HardwareDGX agent

Lambda workspaces help teams organize cloud resources, control access, and separate dev, staging, and production in shared GPU environments. A junior researcher kills a production training run. A cont

Join the livestream to hear from our team members @karan4d and @yoniebans live from @nvidia! https://www.youtube.com/watch?v=pgQDbRMa2Eg

HardwareDGX agent

Nous Research is hosting a livestream featuring team members Karan and Yoni speaking from NVIDIA, likely discussing AI research, model development, or collaboration between Nous Research and NVIDIA. T

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

HardwareDGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

LithoGRPO: Fast Inverse Lithography via GRPO Reinforced Flow Matching

HardwareDGX agent

arXiv:2606.00228v1 Announce Type: new Abstract: In semiconductor manufacturing, lithography projects circuit layouts onto silicon wafers through an optical mask. As circuit features shrink below the w

Lodestar: An Online-Learning LLM Inference Router

HardwareDGX agent

arXiv:2606.00946v1 Announce Type: cross Abstract: Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for G

Marvell shares closed up 32.52% on Tuesday after Jensen Huang hailed the chipmaker as the 'next trillion-dollar company' at Computex (Sawdah Bhaimiya/CNBC)

HardwareDGX agent

Sawdah Bhaimiya / CNBC: Marvell shares closed up 32.52% on Tuesday after Jensen Huang hailed the chipmaker as the “next trillion-dollar company” at Computex — Nvidia's CEO Jensen Huang hailed Marvell

Microsoft announces Surface RTX Spark AI supercomputer development box

HardwareDGX agent

Microsoft Corp. today announced it’s bringing agentic development capabilities into the hands of developers with a new desktop form factor supercomputer called the Surface RTX Spark Dev Box. Developed

MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major…

HardwareDGX agent

MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major sparse attention, paged MSA decode, optimized index scoring

NVIDIA Jetson Brings Agentic AI to the Physical World

HardwareDGX agent

Agentic AI is getting physical. At COMPUTEX on Tuesday, NVIDIA announced NVIDIA JetPack 7.2 and NVIDIA NemoClaw support on NVIDIA Jetson. JetPack 7.2 brings agentic AI skills, Yocto project support, N

Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving

HardwareDGX agent

arXiv:2606.01839v1 Announce Type: cross Abstract: LLM-based agents resolve a user task through many turns of dependent inference and tool calls, producing a workload whose total cost is unknown when t

Open by Design: How NVIDIA and DigitalOcean Are Building the Stack for the Always-On Agentic Era

HardwareDGX agent

NVIDIA and DigitalOcean are collaborating to develop an open infrastructure stack designed to support autonomous AI agents that operate continuously. The initiative emphasizes open-source principles a

Opus 4.8

HardwareDGX agent

Opus 4.8 likely refers to a software release, update, or feature announcement covered by Ben's Bites, a technology newsletter focused on AI and software developments. Without access to the specific ar

PortBERT: Navigating the Depths of Portuguese Language Models

HardwareDGX agent

arXiv:2606.02100v1 Announce Type: new Abstract: Transformer models dominate modern NLP, but efficient, language-specific models remain scarce. In Portuguese, most focus on scale or accuracy, often neg

Practical Aspects on Solving Differential Equations Using Deep Learning: A Primer

HardwareDGX agent

arXiv:2408.11266v5 Announce Type: replace Abstract: Deep learning is now common across many scientific fields, including the study of partial differential equations. This article provides a brief, acc

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

HardwareDGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

Procurement records show at least seven Chinese universities that support the country's military and defense industry are seeking access to Nvidia's H200 chips (Bloomberg)

HardwareDGX agent

Bloomberg: Procurement records show at least seven Chinese universities that support the country's military and defense industry are seeking access to Nvidia's H200 chips — At least seven Chinese univ

Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills

HardwareDGX agent

arXiv:2503.05641v4 Announce Type: replace-cross Abstract: Combining existing pre-trained LLMs is a promising approach for diverse reasoning tasks. However, task-level expert selection is often too coa

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

HardwareDGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

SUPREME: A Multi-GPU Framework for Reproducible Image Unlearning Method Evaluation

HardwareDGX agent

arXiv:2606.00380v1 Announce Type: cross Abstract: Machine unlearning removes the influence of specific training data from a trained model without retraining it from scratch. Evaluating an unlearning m

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVID…

HardwareDGX agent

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVIDIA NemoClaw and OpenShell, connect it to Slack, Outlook, Git

The future of design doesn’t interrupt your process. @NousResearch Hermes keeps designers in creative flow - Rhino → ComfyUI → @Blender in o…

HardwareDGX agent

The future of design doesn’t interrupt your process. @NousResearch Hermes keeps designers in creative flow - Rhino → ComfyUI → @Blender in one seamless agent workflow. No context switching. Just visio

v0.30.1

Local AiDGX agent

Ollama 0.30 provides improved compatibility and performance using llama.cpp, augments the MLX engine on Apple Silicon for broader hardware support, and brings support for a wider range of models inclu

1 Jun 2026

$200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀

HardwareDGX agent

200m investment in world model AI R&D and jobs - a warm British welcome to your European HQ, @runwayml @c_valenzuelab 🇬🇧🚀 ANOTHER big AI Lab is massively investing in London. This time it is @runwayml

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition

HardwareDGX agent

A lot more to come! We are excited to work closely with NVIDIA on the Cosmos Coalition Introducing the Cosmos Coalition A new global initiative with NVIDIA and leading AI labs to build and open-source

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent.

HardwareDGX agent

A new beginning of PC starts with @NVIDIARTXSpark, supercharging what's possible in Hermes Agent. This is the NVIDIA RTX Spark Superchip. A new beginning for personal computers. Designed for creators,

Accelerate LLM model loading and increase context windows with GPUDirect on Amazon FSx for Lustre and TurboQuant

HardwareDGX agent

If you’re iterating on deploying large language models (LLMs) on AWS GPU instances, you’ve probably noticed the larger the model to be loaded into GPU High Bandwidth Memory (HBM), the longer the painf

AI companies wouldn't enslave engineers, or steal chips from Nvidia, or illegally tap a power line. So why, asks @AGSNYT, do they feel entit…

HardwareDGX agent

AI companies wouldn't enslave engineers, or steal chips from Nvidia, or illegally tap a power line. So why, asks @AGSNYT, do they feel entitled to 'a brazen theft of intellectual property' from news o

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning

HardwareDGX agent

arXiv:2605.31481v1 Announce Type: new Abstract: As robot control shifts toward large-scale reinforcement learning with in-loop dynamics computation, the community's reliance on CPU-bound libraries suc

Box says it created 13 new AI-focused roles, like AI architect and AI solutions manager, and plans to grow its staff to 3,000 by early 2027, up from 2,900 (Kalley Huang/New York Times)

HardwareDGX agent

Kalley Huang / New York Times: Box says it created 13 new AI-focused roles, like AI architect and AI solutions manager, and plans to grow its staff to 3,000 by early 2027, up from 2,900 — Box, a Silic

CoMo3R-SLAM: Collaborative Monocular Dense SLAM with Learned 3D Reconstruction Priors for Outdoor Multi-Agent Systems

HardwareDGX agent

arXiv:2605.30488v1 Announce Type: new Abstract: Collaborative dense SLAM is essential for multi-robot teams to achieve scalable and consistent 3D perception across large-scale outdoor environments. Ex

Cross-Entropy Optimization of Physically Grounded Task and Motion Plans

HardwareDGX agent

arXiv:2512.11571v2 Announce Type: replace Abstract: Autonomously performing tasks often requires robots to plan high-level discrete actions and continuous low-level motions to realize them. Previous T

Deep-learning-based low-energy trigger algorithms for the Hyper-Kamiokande experiment

HardwareDGX agent

arXiv:2605.31391v1 Announce Type: cross Abstract: Modern machine learning techniques have become increasingly important in particle physics because of their powerful pattern-recognition capabilities,

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

HardwareDGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

Develop Physical AI Reasoning, World, and Action Models with NVIDIA Cosmos 3

HardwareDGX agent

NVIDIA Cosmos 3 is a frontier foundation model for physical AI that combines physical reasoning, world generation, and action generation within a single open model. The model uses a Mixture-of-Transfo

DSD-GS: Dynamic-Static Decomposition of Gaussian Splatting for Efficient and High-Fidelity Dynamic Scene Reconstruction

HardwareDGX agent

arXiv:2605.30863v1 Announce Type: new Abstract: Dynamic scene reconstruction and novel view synthesis are fundamental to next-generation visual intelligence applications such as virtual reality, robot

Elastic ViTs from Pretrained Models without Retraining

HardwareDGX agent

arXiv:2510.17700v2 Announce Type: replace Abstract: Vision foundation models achieve remarkable performance but are only available in a limited set of pre-determined sizes, forcing sub-optimal deploym

Five thoughts from Nvidia CEO Jensen Huang’s GTC Taipei 2026 keynote

HardwareDGX agent

Useful artificial intelligence has arrived, and if Nvidia Chief Executive Jensen Huang is right, it is about to reshape not only data centers but also the structure of the global economy and the tech

Gradient-Free Training of Spiking Neural Networks via Low-Rank Evolution Strategies

ResearchDGX agent

arXiv:2605.30361v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer compelling energy efficiency on neuromorphic hardware, yet their training remains challenging because the discret

🆕Grok Imagine’s Video Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://www.latent.space/p/video…

HardwareDGX agent

🆕Grok Imagine’s Video Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://www.latent.space/p/video-agents @EthanHe_42, former @xai world model lead and @nvidia

HetCCL: Enabling Collective Communication For Mixed-Vendor Heterogeneous Clusters

ResearchDGX agent

arXiv:2605.31000v1 Announce Type: cross Abstract: Training Large Language Models (LLMs) on heterogeneous clusters presents significant challenges for collective communication, as hardware from multipl

How Cosmos 3 Helps Physical AI Think Before It Acts

HardwareDGX agent

NVIDIA Cosmos 3 uses a mixture-of-transformers architecture that pairs a reasoning transformer with a generation transformer, enabling it to understand object interactions and spatial-temporal relatio

https://x.com/huggingface/status/2061495792246915131

HardwareDGX agent

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost 60,000 followers

Intel: Our upcoming AI chip will be cheaper, run cooler than Nvidia, AMD options

HardwareDGX agent

Intel plans to ship its Crescent Island AI GPU in 2026, targeting Nvidia and AMD on cost and power efficiency. The chip uses cheaper LPDDR5X memory and is designed to run in air-cooled server racks ra

It's cool to see that the http://hf.co/blog has become progressively a major source of learning and news about AI for the community. Just lo…

HardwareDGX agent

It's cool to see that the http://hf.co/blog has become progressively a major source of learning and news about AI for the community. Just look at the recent content created & major announcements & tut

← Previous
1…2122232425…75
Next →