AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
29 May 2026

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI wou…

HardwareDGX agent

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI would become a commodity. now they think they have made some bi

Jensen Huang called Fireworks 'the TSMC of AI factories' at GTC 2026. Here's the @nvidia CEO's full conversation with our own, @lqiao:

HardwareDGX agent

Jensen Huang, NVIDIA's CEO, compared Fireworks to 'the TSMC of AI factories' during a conversation at GTC 2026, highlighting Fireworks' role as a foundational infrastructure provider in the AI industr

Long-Context Modeling with Dynamic Hierarchical Sparse Attention for Memory-Constrained LLM Inference

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

arXiv:2510.24606v2 Announce Type: replace Abstract: The quadratic cost of attention limits the scalability of long-context LLMs, especially under limited hardware memory budgets. While attention is of

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

HardwareDGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

ScheduleStream: Temporal Planning with Samplers for GPU-Accelerated Multi-Arm Task and Motion Planning & Scheduling

HardwareDGX agent

arXiv:2511.04758v2 Announce Type: replace-cross Abstract: Bimanual and humanoid robots are appealing because of their human-like ability to leverage multiple arms to efficiently complete tasks. Howeve

Sources: ByteDance has partnered with chipmaker InnoStar to develop an AI inference chip modeled after Groq's LPUs, which are built to run AI models at low cost (The Information)

HardwareDGX agent

The Information: Sources: ByteDance has partnered with chipmaker InnoStar to develop an AI inference chip modeled after Groq's LPUs, which are built to run AI models at low cost — TikTok owner ByteDan

stop asking for rationality @jeffreyleefunk! it won’t happen here! (actually, it will … eventually ...)

HardwareDGX agent

stop asking for rationality @jeffreyleefunk! it won’t happen here! (actually, it will … eventually ...) @GaryMarcus As companies reduce AI use in effort to achieve productivity improvements in most co

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more rel…

HardwareDGX agent

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more reliable CPUs or even on serverless GPU (that cost less than a

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in …

HardwareDGX agent

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in under 10 seconds. This deep dive shows the systems work behi

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

HardwareDGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

What to expect at Computex 2026: AI chips, budget PCs competing with the MacBook Neo, Nvidia entering laptop SoC market with the rumored N1X chip, and more (PCMag)

HardwareDGX agent

PCMag: What to expect at Computex 2026: AI chips, budget PCs competing with the MacBook Neo, Nvidia entering laptop SoC market with the rumored N1X chip, and more — The theme for Computex 2026, accord

28 May 2026

A Unified Structured Query Understanding Framework for Industrial Semantic Search

HardwareDGX agent

arXiv:2605.27441v1 Announce Type: cross Abstract: Query understanding in large-scale industrial search systems is typically implemented as a cascade of disparate, task-specific components. While indiv

and see, perhaps further reflection of the death of tokenmaxxing

HardwareDGX agent

and see, perhaps further reflection of the death of tokenmaxxing The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic

Heterogeneous Parallelism for Multimodal Large Language Model Training

HardwareDGX agent

arXiv:2605.27678v1 Announce Type: new Abstract: Foundation model training is becoming multimodal, from post-training pipelines to large-scale pretraining. As modality coverage broadens, context window

Intel makes a bid for handheld gaming PCs with new Arc G3 processors

HardwareDGX agent

Intel introduced the Arc G-Series processors, a new family designed for next-generation handheld gaming systems, launching with Arc G3 and Arc G3 Extreme processors running Windows 11 based on Intel C

IoT Image Processing in Hybrid Cloud Environments: Use Case

HardwareDGX agent

This article likely explores practical applications of image processing for IoT devices deployed across hybrid cloud architectures, discussing how organizations can leverage both on-premises and cloud

Learning When to Optimize: Verified Optimization Skills from Expert GPU-Kernel Lineages

HardwareDGX agent

arXiv:2605.28213v1 Announce Type: new Abstract: LLM-based agents are increasingly used to generate GPU kernels, but they often know what optimizations to try without knowing when those optimizations a

Long Live the Librarian! A Persistent Search Sub-Agent for Energy-Efficient Multi-Agent Software Engineering Systems

HardwareDGX agent

arXiv:2605.27787v1 Announce Type: cross Abstract: Multi-agent systems (MAS) have substantially advanced autonomous software engineering (SWE), but their growing inference energy demands raise sustaina

NVIDIA Research Advances Robotics From Simulation to the Real World

HardwareDGX agent

Robotics is entering a new phase: moving from controlled demos and scripted automation toward generalizable, reliable embodied autonomy in the real world. At the International Conference on Robotics a

OpenURMA: A Clean-Room Open Implementation of the Unified Bus Protocol

HardwareDGX agent

arXiv:2605.28717v1 Announce Type: new Abstract: Modern datacenter RDMA is bottlenecked at the network interface, not the wire. A NIC running RoCE or InfiniBand holds per-connection state for every (ap

Source: the Shanghai Futures Exchange is in the early stages of designing futures contracts for AI tokens; US exchanges are set to launch GPU compute futures (Reuters)

HardwareDGX agent

Reuters: Source: the Shanghai Futures Exchange is in the early stages of designing futures contracts for AI tokens; US exchanges are set to launch GPU compute futures — China is designing a futures ma

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

HardwareDGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: eve…

HardwareDGX agent

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic ass…

HardwareDGX agent

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic asset in tech. When the underlying commodity that powers your ent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

HardwareDGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

HardwareDGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation

HardwareDGX agent

arXiv:2605.27582v1 Announce Type: cross Abstract: Embodied navigation requires an agent to map language and visual observations to a stream of spatial actions that drive a real robot through environme

Utilidata raises $40M more to optimize data center power use

HardwareDGX agent

Utilidata Inc. today disclosed that it has raised 40 million in funding from Renown Capital Partners and Keyframe Capital. The cash infusion comes as an extension to a 60.3 million Series C round the

Zero-shot Quantum Neural Architecture Search

ResearchDGX agent

arXiv:2605.27410v1 Announce Type: cross Abstract: Variational Quantum Algorithms (VQAs) are a leading approach to exploiting near-term quantum hardware, leveraging parameterized quantum circuits and c

27 May 2026

AI training data provider Human Archive raises $8.2M

HardwareDGX agent

Artificial intelligence training data provider Human Archive Inc. today announced that it has raised 8.2 million in funding. Wing Venture Capital, NVP Capital, Y Combinator headlined the consortium th

Anthropic Growth and Bedrock Mix Drive AWS Margins Higher While Peers Lag

HardwareDGX agent

AWS experienced margin expansion driven by growth in Anthropic-related services and an improved product mix from Amazon Bedrock, its generative AI platform, while competing cloud providers failed to a

AssetGen: Deployable 3D Asset Generation at Interactive Speed

HardwareDGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 (Nikkei Asia)

HardwareDGX agent

Nikkei Asia: Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 — TAIPEI — Nvidia is now spend

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search

HardwareDGX agent

arXiv:2605.26636v1 Announce Type: cross Abstract: We introduce JetViT, a novel family of hybrid-architecture Vision Transformer (ViT) models that match the accuracy of state-of-the-art full-attention

Multi-Robot Box Transport over Different Surfaces with Decentralized Role-based Proportional Control

HardwareDGX agent

arXiv:2605.26430v1 Announce Type: new Abstract: Collaborative transport of objects via pushing by multiple robots has many applications, ranging from construction and warehouse environments to post di

NightSight: Passive Computation for Navigation in Dark Using Events

HardwareDGX agent

arXiv:2605.26330v1 Announce Type: new Abstract: Small aerial robots are particularly well-suited for search and rescue in confined and hazardous environments due to their agility, low cost, and abilit

Nvidia bets $150B on Taiwan as Trump's plan to make US an AI hub backfires

HardwareDGX agent

Nvidia CEO Jensen Huang announced the chip company plans to invest approximately 150 billion annually in Taiwan, terming it the 'epicentre' of the AI revolution . This represents a dramatic increase f

NVIDIA Blackwell Sets STAC-AI Record for LLM Inference in Finance

HardwareDGX agent

NVIDIA's Blackwell GB200 NVL72 architecture achieved the fastest-ever results on the STAC-AI benchmark for financial LLM inference, delivering up to 3.2x single-GPU performance improvements over the p

NVIDIA Dynamo Snapshot: Fast Startup for Inference Workloads on Kubernetes

HardwareDGX agent

NVIDIA Dynamo Snapshot introduces ModelExpress for 7x faster startup via checkpoint restore and weight streaming with NVIDIA NVLink and NIXL , enabling efficient model initialization for inference wor

Nvidia kills Windows XP-era Control Panel 'after 20 years of dedicated service'

HardwareDGX agent

Nvidia has officially retired the NVIDIA Control Panel after 20 years of service, shifting all functionality to the newer Nvidia App. The legacy Control Panel is no longer included in default GeForce

Qrita: High-performance Top-k and Top-p using Pivot-based Truncation and Selection

HardwareDGX agent

arXiv:2602.01518v2 Announce Type: replace Abstract: Despite their importance in model sampling, efficient implementation of Top-k and Top-p algorithms for large vocabularies remains a significant chal

RT-Lynx: Putting the GEMM Sparsity In a Right Way for Diffusion Models

HardwareDGX agent

arXiv:2605.26632v1 Announce Type: new Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs. While prior work has reduced this cos

SIA: Self Improving AI with Harness & Weight Updates

HardwareDGX agent

arXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l

SIKA-GP: Accelerating Gaussian Process Inference with Sparse Inducing Kernel Approximations for Bayesian Deep Learning

HardwareDGX agent

arXiv:2605.26509v1 Announce Type: new Abstract: Gaussian processes (GPs) provide a principled Bayesian framework for uncertainty estimation, but their computational complexity severely limits scalabil

Sources: Taiwan suspects three people smuggled at least one Nvidia chip shipment to China after first exporting them to Japan; Taiwan detained them last week (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Taiwan suspects three people smuggled at least one Nvidia chip shipment to China after first exporting them to Japan; Taiwan detained them last week — Taiwan prosecutors suspect th

Tensormesh taps Nvidia, AMD and CoreWeave for funding to fix AI model memory problems

HardwareDGX agent

Tensormesh Inc. has hit upon a way to make artificial intelligence inference more efficient by eliminating the need for redundant computations, and its technology is so convincing that several of AI i

The AI GPU market has belonged to hyperscalers, but AMD’s MI350P is coming for the enterprise

HardwareDGX agent

Enterprise GPU access is expanding as Advanced Micro Devices Inc. lowers barriers with the air-cooled Instinct MI350P GPU and a ready-to-run software stack designed for standard enterprise servers. Th

Towards Feedback-to-Plan Decisions for Self-Evolving LLM Agents in CUDA Kernel Generation

HardwareDGX agent

arXiv:2605.26720v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong empirical gains as self-evolving agents for CUDA kernel generation, driven by feedback-conditioned planni

Towards Shared Embodied Intelligence in Humanoid Robots through Optimization Development and Testing of the Human Aware ergoCub Robot

ResearchDGX agent

arXiv:2605.26991v1 Announce Type: new Abstract: Collaboration is central to human behavior, enabling tasks beyond individual capability. This ability arises from coordinating actions through internal

We're open-sourcing the Unigram tokenizer we rebuilt to reduce CPU utilization by 5-6x. Small rerankers and embedders run in single-digit mi…

HardwareDGX agent

We're open-sourcing the Unigram tokenizer we rebuilt to reduce CPU utilization by 5-6x. Small rerankers and embedders run in single-digit milliseconds on GPU, making CPU tokenization a meaningful shar

What’s New for Game Developers in NVIDIA RTX: DLSS 4.5 for UE5 and Multilingual AI Characters

HardwareDGX agent

NVIDIA DLSS 4.5 introduces Dynamic Multi Frame Generation and a second-generation transformer model for Super Resolution to enhance game image quality and frame rates for Unreal Engine 5. The new Tens

26 May 2026

3D-printable humanoid legs let robotics experiments run wild

IndustryDGX agent

Researchers at UC Berkeley have developed Berkeley Humanoid Lite, a low-cost, open-source robot made of 3D-printed parts , which keeps hardware costs under $5,000 with a modular design allowing easy c

A universal vision transformer for fast calorimeter simulations

HardwareDGX agent

arXiv:2601.05289v2 Announce Type: replace-cross Abstract: The high-dimensional complex nature of detectors makes fast calorimeter simulations a prime application for modern generative machine learning

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 …

HardwareDGX agent

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 all the way to 24.04 the other day). Asked Codex on my Mac t

AgentIR: A Workload-Adaptive Cascade Retrieval Substrate for Long-Term Conversational Memory

HardwareDGX agent

arXiv:2605.25092v1 Announce Type: cross Abstract: Long-term conversational memory is a retrieval workload classical IR was not built for: the index grows during the query stream, query types shift int

AI readiness in telecommunications

HardwareDGX agent

This Databricks blog post likely examines the prerequisites and preparedness required for telecommunications companies to successfully implement and deploy AI solutions. It probable covers organizatio

As agentic AI surges, CPUs and air-cooled infrastructure move to the fore

HardwareDGX agent

Agents are turning up the heat for enterprises, turning air-cooled AI infrastructure into a boardroom priority. While GPUs have largely dominated the AI conversation, CPUs are increasingly coming into

Beyond the Aggregation Dilemma: Prior-Retaining Decoupled Learning for Multimodal Graphs

HardwareDGX agent

arXiv:2605.24684v1 Announce Type: cross Abstract: Multimodal Attributed Graph Learning (MAGL) integrates intrinsic node attributes with structural topology via graph aggregation. However, as pretraine

Build high-performance generative AI systems with Strands Agents, NVIDIA NIM, and Amazon Bedrock AgentCore

HardwareDGX agent

In this post you'll learn how to build a multi-agent campaign review system that demonstrates parallel reasoning, context persistence, and traceable execution paths using an integrated architecture th

Develop High-Performance GPU Kernels in C++ with NVIDIA CUDA Tile

HardwareDGX agent

This NVIDIA Developer article guides developers on creating optimized GPU kernels using C++ with NVIDIA's CUDA Tile technology, which enables efficient parallel computation on graphics processors. The

← Previous
1…2324252627…75
Next →