AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
30 Jun 2026

CLEAR-MoE: Shared-Basis Expert Extraction from Frozen Vision Transformers via Calibration-Driven Layer Selection

HardwareDGX agent

arXiv:2606.28516v1 Announce Type: new Abstract: We present CLEAR-MoE, a four-phase post-training pipeline that converts a frozen pretrained Vision Transformer (ViT) into a sparse Mixture-of-Experts (M

ConCise: Training-Free Conclusion-Chain State Compression for Cost-Efficient Multi-Step RAG Services

HardwareDGX agent

arXiv:2606.28361v1 Announce Type: cross Abstract: Multi-step retrieval-augmented generation (RAG) has been widely deployed as LLM-powered web services for complex question answering, where iterative r

DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model

HardwareDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.30292v1 Announce Type: cross Abstract: We present DreamForge-World 0.1 Preview, a preview foundational world model for real-time interactive world simulation. The system adapts the LongLive

Enhanced Diffusion Sampling: Efficient Rare Event Sampling and Free Energy Calculation with Diffusion Models

HardwareDGX agent

arXiv:2602.16634v2 Announce Type: replace-cross Abstract: The rare-event sampling problem has long been the central limiting factor in molecular dynamics (MD), especially in biomolecular simulation. R

Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

HardwareDGX agent

arXiv:2606.29082v1 Announce Type: new Abstract: Would experience designing faster GPU kernels also help close in on a long-standing open mathematical conjecture? Large Language Models (LLMs) integrate

Featuring the man of the moment @dylan522p @shaunmmaguire in today's Training Data episode. Nobody is a more trusted industry insider to the…

HardwareDGX agent

Featuring the man of the moment @dylan522p @shaunmmaguire in today's Training Data episode. Nobody is a more trusted industry insider to the biggest infrastructure build-out in history. The story of h

GPU-Accelerated Inverse Structural Anastylosis from Block Collapse Dynamics

HardwareDGX agent

arXiv:2606.28394v1 Announce Type: new Abstract: The physical anastylosis of collapsed architectural monuments -- the meticulous reassembly of fallen stone elements into their original structural confi

GPU Parallelization Strategies for Forward and Backward Propagation in Shallow Neural Networks: A CUDA-Based Comparative Study

HardwareDGX agent

arXiv:2606.30497v1 Announce Type: cross Abstract: We present a comparative study of CUDA optimization strategies applied to forward and backward propagation in a shallow neural network. Three stacked

HARD-KV: Head-Adaptive Regularization for Decoding-time KV Compression

HardwareDGX agent

arXiv:2606.28831v1 Announce Type: cross Abstract: Long-context LLM inference faces a fundamental conflict: head-adaptive compression algorithms (e.g., Top-p nucleus sampling) offer superior accuracy b

How Jaiveer Singh Is Helping Robots — and Developers — Move Faster

HardwareDGX agent

When Jaiveer Singh talks about robots, he doesn’t begin with spectacle. He begins with infrastructure: the boards inside machines, the software that lets developers see through a robot’s cameras and t

How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost

HardwareDGX agent

As organizations move from AI pilots to production AI factories, infrastructure decisions have shifted from peak chip specifications to cost per token: how many useful tokens they can deliver per doll

How Outpost VFX Uses AWS to Accelerate AI Model Training for Visual Effects

HardwareDGX agent

In this post, we explore how Outpost VFX achieved 8x faster training speeds using AWS infrastructure to transform their face replacement workflow, the technical architecture they implemented to overco

Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning

HardwareDGX agent

Editor’s note: This post is part of Into the Omniverse, a series focused on how developers, 3D practitioners, and enterprises can transform their workflows using the latest advances in OpenUSD and NVI

KernelSight-LM: A Kernel-Level LLM Inference Simulator

Local AiDGX agent

arXiv:2606.28565v1 Announce Type: cross Abstract: As large language models (LLMs) move into production serving, practitioners must rapidly evaluate inference performance across diverse hardware, model

Knowing in Advance When an Evolutionary Outer Loop Will Not Help: A Pre-Registered Cheap-Baseline Screening Rule

HardwareDGX agent

arXiv:2606.29119v1 Announce Type: cross Abstract: We introduce a pre-registered screening rule that decides, before any implementation, whether an evolutionary / population / lifecycle outer loop over

New instances bring AWS’ Graviton5 CPUs to high-performance cloud workloads

HardwareDGX agent

Amazon Web Services Inc.‘s next-generation silicon is working its way deeper into the company’s cloud compute infrastructure offerings with the debut of the new Amazon EC2 C9g and EC2 C9gd instances t

Powering AI agents: CoreWeave’s validation of Nvidia Vera Rubin signals new chapter for rack-scale computing

HardwareDGX agent

Agentic AI is driving key technology providers to rethink the computing architecture required to run rapidly expanding autonomous systems. In response to this challenge, two leading tech companies hav

Project gallery and whitepaper: https://research.nvidia.com/labs/gear/aspire/ ASPIRE is a great collaboration between NVIDIA GEAR lab, UMich…

HardwareDGX agent

Project gallery and whitepaper: https://research.nvidia.com/labs/gear/aspire/ ASPIRE is a great collaboration between NVIDIA GEAR lab, UMich, Berkeley, and CMU. Kudos to all the coauthors who pour the

RAGA: Real Time Ray Traced Gaussian Shadow Casting for 3DGS Avatar-Scene Interaction

HardwareDGX agent

arXiv:2606.29329v1 Announce Type: new Abstract: We study the problem of physically plausible shadow casting when animating 3D Gaussian Splatting (3DGS) avatars, either individually or in multi-avatar

Robust and Efficient Monocular 3D Gaussian SLAM for Kilometer-Scale Outdoor Scenes

HardwareDGX agent

arXiv:2606.30436v1 Announce Type: new Abstract: Scaling monocular 3D Gaussian Splatting (3DGS) SLAM to kilometer-level outdoor environments poses two tightly coupled challenges: fragile long-term pose

Spanning the Visual Analogy Space with a Weight Basis of LoRAs

HardwareDGX agent

arXiv:2602.15727v2 Announce Type: replace-cross Abstract: Visual analogy learning enables image editing via demonstration rather than textual description, allowing users to specify complex transformat

SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces

HardwareDGX agent

arXiv:2403.07711v5 Announce Type: replace-cross Abstract: Given the remarkable achievements in image generation through diffusion models, the research community has shown increasing interest in extend

Stathera nabs $35M to make vacuum-sealed silicon oscillators for AI chips

HardwareDGX agent

Chip component supplier Stathera Inc. today announced that it has closed a 35 million funding round. Semiconductor-focused fund Maverick Silicon led the Series A deal. It was joined by chipmaker Media

SwarmX: Agentic Scheduling for Low-Latency Agentic Systems

HardwareDGX agent

arXiv:2606.21401v2 Announce Type: replace-cross Abstract: Agentic AI applications compose multiple model calls and tool executions, creating new scheduling challenges for GPU-CPU clusters. Their infer

TokenBudgeting: Our Conversations with Enterprises on Token Spend

HardwareDGX agent

TokenBudgeting examines enterprise spending patterns and cost management strategies related to AI token consumption, based on SemiAnalysis's direct conversations with corporate customers. The analysis

Transolver-3: Scaling Up Transformer Solvers to Industrial-Scale Geometries

HardwareDGX agent

arXiv:2602.04940v2 Announce Type: replace Abstract: Deep learning has emerged as a transformative tool for the neural surrogate modeling of partial differential equations (PDEs), known as neural PDE s

29 Jun 2026

DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers

HardwareDGX agent

arXiv:2601.16956v1 Announce Type: cross Abstract: The rapid growth of Large Transformer-based models, specifically Large Language Models (LLMs), now scaling to trillions of parameters, has necessitate

Everyone is concerned about Chinese models being banned The entire ecosystem of successful AI startups with real revenue are all catching pr…

HardwareDGX agent

Dylan Patel discusses concerns within the AI startup ecosystem about potential Chinese AI model bans and their impact on companies generating significant revenue. The post appears to address industry-

Firefly Aerospace Operates NVIDIA Jetson in Lunar Orbit for the First Time

HardwareDGX agent

Firefly Aerospace partnered with NVIDIA to deploy an NVIDIA Jetson module on its Elytra spacecraft in lunar orbit, enabling on-orbit AI processing for the Ocula Moon imaging service. The Jetson module

Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference and infrastr…

HardwareDGX agent

Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference and infrastructure. After that: cocktails and a DJ. Oh la la! See you th

How to Govern Autonomous Agents in Enterprise AI Factories

HardwareDGX agent

Enterprise autonomous AI agents inspect code, run tests, search documents, and operate for extended periods on behalf of users , requiring governance frameworks to control access and actions. The Ente

Import AI 463: Self-improving robots; a 10k Chinese GPU cluster; and an elegiac essay for the human era

HardwareDGX agent

This newsletter issue covers three major topics in AI development: advances in self-improving robotic systems, details about a large-scale Chinese GPU cluster with 10,000 units for AI training, and a

Nvidia CEO Jensen “@Tesla stack is the most advanced autonomous vehicle stack in the world. I’m fairly certain they were already using end-t…

HardwareDGX agent

Nvidia CEO Jensen “@Tesla stack is the most advanced autonomous vehicle stack in the world. I’m fairly certain they were already using end-to-end AI. Whether their AI did reasoning or not in somewhat

One day left to submit your projects for the Hermes Agent Accelerated Business Hackathon presented by @NVIDIAAI × @stripe × @NousResearch! S…

HardwareDGX agent

One day left to submit your projects for the Hermes Agent Accelerated Business Hackathon presented by @NVIDIAAI × @stripe × @NousResearch! Submissions close 11:59 PM PT tomorrow, June 30th. Last minut

Source: Taiwanese authorities raided Super Micro's Taiwan office as part of a probe into the alleged smuggling of Nvidia chips to China; SMCI closed down 8.1% (Debby Wu/Bloomberg)

HardwareDGX agent

Debby Wu / Bloomberg: Source: Taiwanese authorities raided Super Micro's Taiwan office as part of a probe into the alleged smuggling of Nvidia chips to China; SMCI closed down 8.1% — Super Micro Compu

Test-Input Generation for Tensor Programs: What Actually Finds Kernel Bugs

HardwareDGX agent

arXiv:2606.27396v1 Announce Type: cross Abstract: Test-input generation for tensor kernels is folkloric. Most projects pick a representative shape and dtype, run a fixed-shape allclose-style check, an

The Context-Ready Transformer

HardwareDGX agent

arXiv:2606.27538v1 Announce Type: cross Abstract: We introduce the context-ready transformer, a new recurrent neural network architecture built from a D-layer transformer block that pre-contextualizes

Trend spotting: HPE, Aible and Nvidia teach AI agents to spot the next big signal

HardwareDGX agent

Many organizations today are floating in a ton of data. This is good for powering AI agents, but less favorable when it comes to identifying which signals from all of that data actually matter. The so

28 Jun 2026

Australia-based Firmus partners with Nvidia to build its first data center in Batam, Indonesia; the 360 MW Nvidia DSX AI factory campus is developed with DayOne (Bloomberg)

HardwareDGX agent

Bloomberg: Australia-based Firmus partners with Nvidia to build its first data center in Batam, Indonesia; the 360 MW Nvidia DSX AI factory campus is developed with DayOne — Firmus Technologies Pty Lt

people in SF be like what's ur superpower 'i won the imo' 'i got 3900 on codeforces' 'I can beat all the leetcoding' Mine is that I can comm…

HardwareDGX agent

This post humorously contrasts the competitive achievements people in San Francisco boast about—such as winning the International Mathematical Olympiad, achieving high competitive programming ratings,

27 Jun 2026

A look at advanced chip packaging, now more reliant on TSMC and its partners in Taiwan than ever, and the efforts to address this bottleneck in the US (Don Clark/New York Times)

HardwareDGX agent

Don Clark / New York Times: A look at advanced chip packaging, now more reliant on TSMC and its partners in Taiwan than ever, and the efforts to address this bottleneck in the US — A silicon wafer ref

AI executives and lobbyists say they are seeking regulatory clarity from the Trump administration but are wary of pressing for answers, fearing retaliation (Politico)

HardwareDGX agent

Politico: AI executives and lobbyists say they are seeking regulatory clarity from the Trump administration but are wary of pressing for answers, fearing retaliation — The unpredictability was on disp

How AI is shaping the 2026 US midterms, as public anger grows against data center expansion and the AI industry emerges as one of the biggest financial backers (Bloomberg)

HardwareDGX agent

Bloomberg: How AI is shaping the 2026 US midterms, as public anger grows against data center expansion and the AI industry emerges as one of the biggest financial backers — From data center backlash t

NEW paper from NVIDIA. (bookmark it) Speed-of-light performance analysis tells you the theoretical floor of a workload, but teams still deri…

HardwareDGX agent

NEW paper from NVIDIA. (bookmark it) Speed-of-light performance analysis tells you the theoretical floor of a workload, but teams still derive it by hand and freeze it. SOLAR automates the whole thing

You're a negative Nancy I'm positive Patel

HardwareDGX agent

Dylan Patel responds to criticism or skepticism (represented as 'negative Nancy') with an optimistic counterargument (playing on his own name as 'positive Patel'), likely defending a bullish outlook o

26 Jun 2026

Adding my voice to the chorus who are saddened by @Om’s too-early passing.

HardwareDGX agent

Adding my voice to the chorus who are saddened by @Om’s too-early passing. Just heartbroken to hear of the passing of the incredible Om Malik. @Om was always a pioneer, a deep thinker, and a truly ori

AWS hikes prices for Nvidia GPUs in its EC2 Capacity Blocks service, which let businesses rent AI compute in advance, by 20%; Trainium chip pricing is unchanged (Catherine Perloff/The Information)

HardwareDGX agent

Catherine Perloff / The Information: AWS hikes prices for Nvidia GPUs in its EC2 Capacity Blocks service, which let businesses rent AI compute in advance, by 20%; Trainium chip pricing is unchanged —

b9816

Local AiDGX agent

B9816 is a release build number for llama.cpp, an open-source project that enables efficient large language model inference in C/C++ on consumer hardware. This release likely contains bug fixes, perfo

Blackwell Approachability and Gradient Equilibrium are Equivalent

HardwareDGX agent

arXiv:2606.27315v1 Announce Type: new Abstract: Gradient equilibrium (GEQ) is a recently introduced online optimization framework that generalizes first-order stationarity from offline optimization an

BREAKING NEWS: OpenAI’s Technical Lead for Compute has returned to OpenAI after joining Meta in April 2026. What is going on at Meta that pe…

HardwareDGX agent

BREAKING NEWS: OpenAI’s Technical Lead for Compute has returned to OpenAI after joining Meta in April 2026. What is going on at Meta that people are quitting after only a couple of months? Did Anuj ge

CARVE: Content-Aware Recurrent with Value Efficiency for Chunk-Parallel Linear Attention

HardwareDGX agent

arXiv:2606.27229v1 Announce Type: cross Abstract: Recurrent models must forget in order to remember, yet the state of the art decides what to erase without consulting what is stored -- the gate sees o

CAT-Q: Cost-efficient and Accurate Ternary Quantization for LLMs

HardwareDGX agent

arXiv:2606.26650v1 Announce Type: cross Abstract: In this paper, we present CAT-Q, Cost-efficient and Accurate Ternary Quantization, for compressing and accelerating LLMs. Unlike existing state-of-the

Compiler-Driven Approximation Tuning for Hyperdimensional Computing

ResearchDGX agent

arXiv:2606.26547v1 Announce Type: cross Abstract: As Moore's law reaches its physical and economic limits, domain-specific approaches are increasingly employed to accelerate machine learning workloads

DASH: Faster Shampoo via Batched Block Preconditioning and Efficient Inverse-Root Solvers

HardwareDGX agent

arXiv:2602.02016v2 Announce Type: replace Abstract: Shampoo is one of the leading approximate second-order optimizers: a variant of it has won the MLCommons AlgoPerf competition, and it has been shown

Event-based Gaze Control System for Accurate Real-time Spin Estimation in Professional Ball Games

HardwareDGX agent

arXiv:2606.26780v1 Announce Type: new Abstract: Spin plays a crucial role in many ball sports due to its effect on the trajectory of the ball. Vision-based estimation of the ball's spin during a game

Hierarchical Muon: Tiled Newton-Schulz Updates for Efficient Muon Optimization

HardwareDGX agent

arXiv:2606.27216v1 Announce Type: cross Abstract: Muon-type optimizers construct update directions for dense neural-network weights by applying a finite Newton-Schulz map to momentum-gradient matrices

OctoSense: Self-Supervised Learning for Multimodal Robot Perception

HardwareDGX agent

arXiv:2606.27317v1 Announce Type: new Abstract: We present OctoSense, an open-source sensor platform with stereo RGB and event cameras, LiDAR, a thermal camera, an inertial measurement unit, RTK-corre

Otter Weather: Skillful and Computationally Efficient Medium-Range Weather Forecasting

HardwareDGX agent

arXiv:2606.26421v1 Announce Type: new Abstract: State-of-the-art medium-range AI weather models can outperform traditional Numerical Weather Prediction (NWP) but require massive training budgets. This

ProtoKV: Streaming Video Understanding under Delayed Query with Summary-State Memory

HardwareDGX agent

arXiv:2606.26762v1 Announce Type: new Abstract: Streaming video understanding (SVU) must answer queries that arrive asynchronously while visual tokens stream continuously under strict GPU-memory and q

“Scale cannot solve AI’s fundamental problem with accuracy” If my argument @financialtimes, excerpted below, is remotely correct, hyperscali…

HardwareDGX agent

“Scale cannot solve AI’s fundamental problem with accuracy” If my argument @financialtimes, excerpted below, is remotely correct, hyperscaling will prove to be among the biggest financial blunders in

← Previous
1…1617181920…75
Next →