AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
Hardware

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI wou…

DGX agent

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI would become a commodity. now they think they have made some bi

hardwaregary-marcus--x
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

Jensen Huang called Fireworks 'the TSMC of AI factories' at GTC 2026. Here's the @nvidia CEO's full conversation with our own, @lqiao:

DGX agent

Jensen Huang, NVIDIA's CEO, compared Fireworks to 'the TSMC of AI factories' during a conversation at GTC 2026, highlighting Fireworks' role as a foundational infrastructure provider in the AI industr

hardwaresonya-huang--x
29 May 2026
Model Releases

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

DGX agent

arXiv:2510.24606v2 Announce Type: replace Abstract: The quadratic cost of attention limits the scalability of long-context LLMs, especially under limited hardware memory budgets. While attention is of

model-releasesarxiv-cs-cl
29 May 2026
Hardware

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

DGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

hardwarenvidia-developer
29 May 2026
Hardware

ScheduleStream: Temporal Planning with Samplers for GPU-Accelerated Multi-Arm Task and Motion Planning & Scheduling

DGX agent

arXiv:2511.04758v2 Announce Type: replace-cross Abstract: Bimanual and humanoid robots are appealing because of their human-like ability to leverage multiple arms to efficiently complete tasks. Howeve

hardwarearxiv-cs-ai
29 May 2026
Hardware

Sources: ByteDance has partnered with chipmaker InnoStar to develop an AI inference chip modeled after Groq's LPUs, which are built to run AI models at low cost (The Information)

DGX agent

The Information: Sources: ByteDance has partnered with chipmaker InnoStar to develop an AI inference chip modeled after Groq's LPUs, which are built to run AI models at low cost — TikTok owner ByteDan

hardwaretechmeme
29 May 2026
Hardware

stop asking for rationality @jeffreyleefunk! it won’t happen here! (actually, it will … eventually ...)

DGX agent

stop asking for rationality @jeffreyleefunk! it won’t happen here! (actually, it will … eventually ...) @GaryMarcus As companies reduce AI use in effort to achieve productivity improvements in most co

hardwaregary-marcus--x
29 May 2026
Hardware

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more rel…

DGX agent

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more reliable CPUs or even on serverless GPU (that cost less than a

hardwareclem-delangue--x
29 May 2026
Hardware

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in …

DGX agent

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in under 10 seconds. This deep dive shows the systems work behi

hardwaretogether-ai--x
29 May 2026
Hardware

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

DGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

hardwarearxiv-cs-ai
29 May 2026
Hardware

What to expect at Computex 2026: AI chips, budget PCs competing with the MacBook Neo, Nvidia entering laptop SoC market with the rumored N1X chip, and more (PCMag)

DGX agent

PCMag: What to expect at Computex 2026: AI chips, budget PCs competing with the MacBook Neo, Nvidia entering laptop SoC market with the rumored N1X chip, and more — The theme for Computex 2026, accord

hardwaretechmeme
29 May 2026
Hardware

A Unified Structured Query Understanding Framework for Industrial Semantic Search

DGX agent

arXiv:2605.27441v1 Announce Type: cross Abstract: Query understanding in large-scale industrial search systems is typically implemented as a cascade of disparate, task-specific components. While indiv

hardwarearxiv-cs-lg
28 May 2026
Hardware

and see, perhaps further reflection of the death of tokenmaxxing

DGX agent

and see, perhaps further reflection of the death of tokenmaxxing The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic

hardwaregary-marcus--x
28 May 2026
Hardware

Heterogeneous Parallelism for Multimodal Large Language Model Training

DGX agent

arXiv:2605.27678v1 Announce Type: new Abstract: Foundation model training is becoming multimodal, from post-training pipelines to large-scale pretraining. As modality coverage broadens, context window

hardwarearxiv-cs-lg
28 May 2026
Hardware

Intel makes a bid for handheld gaming PCs with new Arc G3 processors

DGX agent

Intel introduced the Arc G-Series processors, a new family designed for next-generation handheld gaming systems, launching with Arc G3 and Arc G3 Extreme processors running Windows 11 based on Intel C

hardwarears-technica
28 May 2026
Hardware

IoT Image Processing in Hybrid Cloud Environments: Use Case

DGX agent

This article likely explores practical applications of image processing for IoT devices deployed across hybrid cloud architectures, discussing how organizations can leverage both on-premises and cloud

hardwarevultr
28 May 2026
Hardware

Learning When to Optimize: Verified Optimization Skills from Expert GPU-Kernel Lineages

DGX agent

arXiv:2605.28213v1 Announce Type: new Abstract: LLM-based agents are increasingly used to generate GPU kernels, but they often know what optimizations to try without knowing when those optimizations a

hardwarearxiv-cs-ai
28 May 2026
Hardware

Long Live the Librarian! A Persistent Search Sub-Agent for Energy-Efficient Multi-Agent Software Engineering Systems

DGX agent

arXiv:2605.27787v1 Announce Type: cross Abstract: Multi-agent systems (MAS) have substantially advanced autonomous software engineering (SWE), but their growing inference energy demands raise sustaina

hardwarearxiv-cs-cl
28 May 2026
Hardware

NVIDIA Research Advances Robotics From Simulation to the Real World

DGX agent

Robotics is entering a new phase: moving from controlled demos and scripted automation toward generalizable, reliable embodied autonomy in the real world. At the International Conference on Robotics a

hardwarenvidia-blog
28 May 2026
Hardware

OpenURMA: A Clean-Room Open Implementation of the Unified Bus Protocol

DGX agent

arXiv:2605.28717v1 Announce Type: new Abstract: Modern datacenter RDMA is bottlenecked at the network interface, not the wire. A NIC running RoCE or InfiniBand holds per-connection state for every (ap

hardwarearxiv-cs-ai
28 May 2026
Hardware

Source: the Shanghai Futures Exchange is in the early stages of designing futures contracts for AI tokens; US exchanges are set to launch GPU compute futures (Reuters)

DGX agent

Reuters: Source: the Shanghai Futures Exchange is in the early stages of designing futures contracts for AI tokens; US exchanges are set to launch GPU compute futures — China is designing a futures ma

hardwaretechmeme
28 May 2026
Hardware

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

DGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

hardwarearxiv-cs-ai
28 May 2026
Hardware

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: eve…

DGX agent

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights

hardwareclem-delangue--x
28 May 2026
Hardware

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic ass…

DGX agent

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic asset in tech. When the underlying commodity that powers your ent

hardwaregary-marcus--x
28 May 2026
Hardware

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

DGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

hardwareclem-delangue--x
28 May 2026
Hardware

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

DGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

hardwareclem-delangue--x
28 May 2026
Hardware

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation

DGX agent

arXiv:2605.27582v1 Announce Type: cross Abstract: Embodied navigation requires an agent to map language and visual observations to a stream of spatial actions that drive a real robot through environme

hardwarearxiv-cs-cv
28 May 2026
Hardware

Utilidata raises $40M more to optimize data center power use

DGX agent

Utilidata Inc. today disclosed that it has raised 40 million in funding from Renown Capital Partners and Keyframe Capital. The cash infusion comes as an extension to a 60.3 million Series C round the

hardwaresiliconangle
28 May 2026
Research

Zero-shot Quantum Neural Architecture Search

DGX agent

arXiv:2605.27410v1 Announce Type: cross Abstract: Variational Quantum Algorithms (VQAs) are a leading approach to exploiting near-term quantum hardware, leveraging parameterized quantum circuits and c

researcharxiv-cs-lg
28 May 2026
Hardware

AI training data provider Human Archive raises $8.2M

DGX agent

Artificial intelligence training data provider Human Archive Inc. today announced that it has raised 8.2 million in funding. Wing Venture Capital, NVP Capital, Y Combinator headlined the consortium th

hardwaresiliconangle
27 May 2026
Hardware

Anthropic Growth and Bedrock Mix Drive AWS Margins Higher While Peers Lag

DGX agent

AWS experienced margin expansion driven by growth in Anthropic-related services and an improved product mix from Amazon Bedrock, its generative AI platform, while competing cloud providers failed to a

hardwaresemianalysis
27 May 2026
Hardware

AssetGen: Deployable 3D Asset Generation at Interactive Speed

DGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

hardwarearxiv-cs-ai
27 May 2026
Hardware

Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 (Nikkei Asia)

DGX agent

Nikkei Asia: Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 — TAIPEI — Nvidia is now spend

hardwaretechmeme
27 May 2026
Hardware

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search

DGX agent

arXiv:2605.26636v1 Announce Type: cross Abstract: We introduce JetViT, a novel family of hybrid-architecture Vision Transformer (ViT) models that match the accuracy of state-of-the-art full-attention

hardwarearxiv-cs-ai
27 May 2026
Hardware

Multi-Robot Box Transport over Different Surfaces with Decentralized Role-based Proportional Control

DGX agent

arXiv:2605.26430v1 Announce Type: new Abstract: Collaborative transport of objects via pushing by multiple robots has many applications, ranging from construction and warehouse environments to post di

hardwarearxiv-cs-ro
27 May 2026
Hardware

NightSight: Passive Computation for Navigation in Dark Using Events

DGX agent

arXiv:2605.26330v1 Announce Type: new Abstract: Small aerial robots are particularly well-suited for search and rescue in confined and hazardous environments due to their agility, low cost, and abilit

hardwarearxiv-cs-ro
27 May 2026
Hardware

Nvidia bets $150B on Taiwan as Trump's plan to make US an AI hub backfires

DGX agent

Nvidia CEO Jensen Huang announced the chip company plans to invest approximately 150 billion annually in Taiwan, terming it the 'epicentre' of the AI revolution . This represents a dramatic increase f

hardwarears-technica
27 May 2026
Hardware

NVIDIA Blackwell Sets STAC-AI Record for LLM Inference in Finance

DGX agent

NVIDIA's Blackwell GB200 NVL72 architecture achieved the fastest-ever results on the STAC-AI benchmark for financial LLM inference, delivering up to 3.2x single-GPU performance improvements over the p

hardwarenvidia-developer
27 May 2026
Hardware

NVIDIA Dynamo Snapshot: Fast Startup for Inference Workloads on Kubernetes

DGX agent

NVIDIA Dynamo Snapshot introduces ModelExpress for 7x faster startup via checkpoint restore and weight streaming with NVIDIA NVLink and NIXL , enabling efficient model initialization for inference wor

hardwarenvidia-developer
27 May 2026
Hardware

Nvidia kills Windows XP-era Control Panel 'after 20 years of dedicated service'

DGX agent

Nvidia has officially retired the NVIDIA Control Panel after 20 years of service, shifting all functionality to the newer Nvidia App. The legacy Control Panel is no longer included in default GeForce

hardwarears-technica
27 May 2026
Hardware

Qrita: High-performance Top-k and Top-p using Pivot-based Truncation and Selection

DGX agent

arXiv:2602.01518v2 Announce Type: replace Abstract: Despite their importance in model sampling, efficient implementation of Top-k and Top-p algorithms for large vocabularies remains a significant chal

hardwarearxiv-cs-ai
27 May 2026
Hardware

RT-Lynx: Putting the GEMM Sparsity In a Right Way for Diffusion Models

DGX agent

arXiv:2605.26632v1 Announce Type: new Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs. While prior work has reduced this cos

hardwarearxiv-cs-lg
27 May 2026
Hardware

SIA: Self Improving AI with Harness & Weight Updates

DGX agent

arXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l

hardwarearxiv-cs-ai
27 May 2026
Hardware

SIKA-GP: Accelerating Gaussian Process Inference with Sparse Inducing Kernel Approximations for Bayesian Deep Learning

DGX agent

arXiv:2605.26509v1 Announce Type: new Abstract: Gaussian processes (GPs) provide a principled Bayesian framework for uncertainty estimation, but their computational complexity severely limits scalabil

hardwarearxiv-cs-lg
27 May 2026
Hardware

Sources: Taiwan suspects three people smuggled at least one Nvidia chip shipment to China after first exporting them to Japan; Taiwan detained them last week (Bloomberg)

DGX agent

Bloomberg: Sources: Taiwan suspects three people smuggled at least one Nvidia chip shipment to China after first exporting them to Japan; Taiwan detained them last week — Taiwan prosecutors suspect th

hardwaretechmeme
27 May 2026
Hardware

Tensormesh taps Nvidia, AMD and CoreWeave for funding to fix AI model memory problems

DGX agent

Tensormesh Inc. has hit upon a way to make artificial intelligence inference more efficient by eliminating the need for redundant computations, and its technology is so convincing that several of AI i

hardwaresiliconangle
27 May 2026
Hardware

The AI GPU market has belonged to hyperscalers, but AMD’s MI350P is coming for the enterprise

DGX agent

Enterprise GPU access is expanding as Advanced Micro Devices Inc. lowers barriers with the air-cooled Instinct MI350P GPU and a ready-to-run software stack designed for standard enterprise servers. Th

hardwaresiliconangle
27 May 2026
Hardware

Towards Feedback-to-Plan Decisions for Self-Evolving LLM Agents in CUDA Kernel Generation

DGX agent

arXiv:2605.26720v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong empirical gains as self-evolving agents for CUDA kernel generation, driven by feedback-conditioned planni

hardwarearxiv-cs-ai
27 May 2026
← Previous
1…2930313233…94
Next →