AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
1 Jun 2026

Welcome NVIDIA Cosmos 3: The First Open Omni-model for Physical AI Reasoning and Action

HardwareDGX agent

NVIDIA Cosmos 3 is an open-source omni-model designed for physical AI reasoning and action tasks, representing an advancement in multimodal AI systems. The model integrates multiple modalities to enab

whoah - Grace + Blackwell chips in a laptop. @Microsoft + @NVIDIA teaming up to take on 6 years of total dominance of Apple Silicon

HardwareDGX agent

Microsoft and NVIDIA are partnering to develop laptop chips based on Grace and Blackwell architectures to compete with Apple Silicon's six-year market dominance. This collaboration aims to challenge A

Windows before Jensen revolutionized the PC were dark days Now we have NX1 and all problems will be solved

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

This post from Dylan Patel discusses how PC graphics capabilities were limited before Jensen (likely referring to Jensen Huang of NVIDIA) introduced revolutionary GPU technology, and suggests that the

Wirescreen analysis of 3,800 Chinese military procurement records finds 500+ instances since 2019 where the PLA sought Nvidia chips, including the A100 and A800 (New York Times)

HardwareDGX agent

New York Times: Wirescreen analysis of 3,800 Chinese military procurement records finds 500+ instances since 2019 where the PLA sought Nvidia chips, including the A100 and A800 — An analysis of six ye

31 May 2026

A look at AMD CEO Lisa Su's and Nvidia CEO Jensen Huang's contrasting China playbooks, with Su keeping a lower profile; China accounts for ~20% of AMD's revenue (Reuters)

HardwareDGX agent

Reuters: A look at AMD CEO Lisa Su's and Nvidia CEO Jensen Huang's contrasting China playbooks, with Su keeping a lower profile; China accounts for ~20% of AMD's revenue — When AMD CEO Lisa Su arrived

is the word for this “rug pull”?

HardwareDGX agent

is the word for this “rug pull”? 🚨 SPACEX MAY BE THE BIGGEST INSIDER CASHOUT IN MARKET HISTORY. SpaceX is expected to IPO at a valuation of up to $2 TRILLION. That would instantly make it larger than

Sources: Microsoft and Nvidia will unveil the first Windows PCs powered by Nvidia SoCs, including devices from Surface and Dell, at Computex and Build 2026 (Ina Fried/Axios)

HardwareDGX agent

Ina Fried / Axios: Sources: Microsoft and Nvidia will unveil the first Windows PCs powered by Nvidia SoCs, including devices from Surface and Dell, at Computex and Build 2026 — The company best known

卧槽!你随便扔一首 YouTube 歌,它就能直接拆成 6 条独立音轨——人声、鼓、贝斯、吉他、钢琴、其他……全给你扒出来! 这工具叫 StemDeck,GitHub 开源项目,本地跑全程不上传、不注册、不花钱!它基于 Demucs 模型,拆完后给你一个接近 DAW 的界面: 静…

HardwareDGX agent

卧槽!你随便扔一首 YouTube 歌,它就能直接拆成 6 条独立音轨——人声、鼓、贝斯、吉他、钢琴、其他……全给你扒出来! 这工具叫 StemDeck,GitHub 开源项目,本地跑全程不上传、不注册、不花钱!它基于 Demucs 模型,拆完后给你一个接近 DAW 的界面: 静音、独奏、拉推子、缩波形、设循环、单轨下载、混音导出……想怎么玩就怎么玩。硬核亮点: 本地全速运行:自动识别 GPU,A

30 May 2026

A profile of OpenAI President Greg Brockman, who is overseeing product in his new role and has become an important ambassador for AI to the Trump administration (Wall Street Journal)

HardwareDGX agent

Wall Street Journal: A profile of OpenAI President Greg Brockman, who is overseeing product in his new role and has become an important ambassador for AI to the Trump administration — After years in t

As robotaxi companies attempt to scale in the US, they face increasing scrutiny and mounting criticism from drivers, law enforcement, and local governments (Sean McLain/Wall Street Journal)

HardwareDGX agent

Sean McLain / Wall Street Journal: As robotaxi companies attempt to scale in the US, they face increasing scrutiny and mounting criticism from drivers, law enforcement, and local governments — As auto

We took the Hot Wings Challenge to NVIDIA GTC 🌶️ @realDanFu (VP of Kernels) and @sarung (VP of Customer Success) answered some questions ar…

HardwareDGX agent

We took the Hot Wings Challenge to NVIDIA GTC 🌶️ @realDanFu (VP of Kernels) and @sarung (VP of Customer Success) answered some questions around AI, one spicy wing at a time. Some people sweat. Some pe

When you go into something with apprehension, you're going to end up with repulsion

HardwareDGX agent

This post likely discusses how approaching a situation with fear or negative expectations can create a self-fulfilling prophecy, resulting in aversion or disgust toward that outcome. Dylan Patel appea

29 May 2026

AI Dark Output: The Visible Cost of Invisible Output

HardwareDGX agent

Why AI's increasing output is going to be one of the hardest economic measurement problems in history. AI 'Dark Output' could end up being the majority of economic activity, but a challenge to measure

AI Dark Output: The Visible Cost of Invisible Output Why AI's increasing output is going to be one of the hardest economic measurement probl…

HardwareDGX agent

AI Dark Output: The Visible Cost of Invisible Output Why AI's increasing output is going to be one of the hardest economic measurement problems in history. AI 'Dark Output' could end up being the majo

BadBlocks: Low-Cost and Stealthy Backdoor Attacks Tailored for Text-to-Image Diffusion Models

HardwareDGX agent

arXiv:2508.03221v5 Announce Type: replace-cross Abstract: Despite the remarkable progress of diffusion models in image generation, recent studies reveal their vulnerability to backdoor attacks via cov

Bastion: Budget-Aware Speculative Decoding with Tree-structured Block Diffusion Drafting

HardwareDGX agent

arXiv:2605.29727v1 Announce Type: new Abstract: Block-diffusion drafters have recently emerged as a powerful alternative for speculative decoding by predicting multiple future-token distributions in a

CA-AC-MPC: CUDA-Accelerated Actor-Critic Model Predictive Control

HardwareDGX agent

arXiv:2605.29155v1 Announce Type: cross Abstract: In the literature, actor-critic model predictive control (AC-MPC) integrates MPC with reinforcement learning to enable high-performance control of com

Comprehensive observability for Amazon SageMaker AI LLM inference: From GPU utilization to LLM quality

HardwareDGX agent

This post demonstrates a comprehensive observability solution using Amazon Managed Grafana dashboards that provides a holistic view of both quality and quantity for LLMs served on Amazon SageMaker AI

DELOS: Detecting Shallow Transits in Kepler Photometry Using a Contrastive-Learning Framework

HardwareDGX agent

arXiv:2605.29428v1 Announce Type: cross Abstract: We present DEtection in phase-folded Light curves with cOntrastive Scoring (DELOS), a contrastive-learning-based framework designed to search for shal

DFlash: Block Diffusion for Flash Speculative Decoding

HardwareDGX agent

arXiv:2602.06036v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) deliver strong performance but require inherently sequential decoding, leading to high inference latency

EarlyTom: Early Token Compression Completes Fast Video Understanding

HardwareDGX agent

arXiv:2605.30010v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have demonstrated strong capabilities in video understanding tasks. However, their practical deployment is stil

Elon Musk: “I don't think most people understand just how quickly machine intelligence is advancing. It's much faster than almost anyone rea…

HardwareDGX agent

Elon Musk: “I don't think most people understand just how quickly machine intelligence is advancing. It's much faster than almost anyone realizes, even within Silicon Valley and certainly outside Sili

ESAM++: Efficient Online 3D Perception on the Edge

HardwareDGX agent

arXiv:2605.29505v1 Announce Type: new Abstract: Online 3D scene perception in real time is essential for robotics, AR/VR, and autonomous systems, particularly in edge computing scenarios where computa

Hot take on what comes next, after the sudden decline of tokenmaxxing: - OpenAI will struggle - with the decline of tokenmaxxing Anthropic w…

HardwareDGX agent

Hot take on what comes next, after the sudden decline of tokenmaxxing: - OpenAI will struggle - with the decline of tokenmaxxing Anthropic will struggle (aside from this quarter) to make a profit - Go

How Together AI built the world’s fastest speech-to-text stack

HardwareDGX agent

Together AI developed an optimized speech-to-text system focused on achieving the fastest processing speeds through technical innovations in their inference stack and model optimization. The approach

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI wou…

HardwareDGX agent

i warned these guys of exactly this problem - no moat because everyone is training on same data - two years ago. and warned them that AI would become a commodity. now they think they have made some bi

Jensen Huang called Fireworks 'the TSMC of AI factories' at GTC 2026. Here's the @nvidia CEO's full conversation with our own, @lqiao:

HardwareDGX agent

Jensen Huang, NVIDIA's CEO, compared Fireworks to 'the TSMC of AI factories' during a conversation at GTC 2026, highlighting Fireworks' role as a foundational infrastructure provider in the AI industr

Run Step 3.7 Flash on NVIDIA GPUs with Enterprise-Ready Multimodal AI

HardwareDGX agent

Step 3.7 Flash is a 198B-parameter Mixture-of-Experts vision-language model designed for enterprise-scale production workloads, featuring native image and video input, a 256k context window, and confi

ScheduleStream: Temporal Planning with Samplers for GPU-Accelerated Multi-Arm Task and Motion Planning & Scheduling

HardwareDGX agent

arXiv:2511.04758v2 Announce Type: replace-cross Abstract: Bimanual and humanoid robots are appealing because of their human-like ability to leverage multiple arms to efficiently complete tasks. Howeve

Sources: ByteDance has partnered with chipmaker InnoStar to develop an AI inference chip modeled after Groq's LPUs, which are built to run AI models at low cost (The Information)

HardwareDGX agent

The Information: Sources: ByteDance has partnered with chipmaker InnoStar to develop an AI inference chip modeled after Groq's LPUs, which are built to run AI models at low cost — TikTok owner ByteDan

stop asking for rationality @jeffreyleefunk! it won’t happen here! (actually, it will … eventually ...)

HardwareDGX agent

stop asking for rationality @jeffreyleefunk! it won’t happen here! (actually, it will … eventually ...) @GaryMarcus As companies reduce AI use in effort to achieve productivity improvements in most co

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more rel…

HardwareDGX agent

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more reliable CPUs or even on serverless GPU (that cost less than a

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in …

HardwareDGX agent

Together AI serves the two fastest STT models measured by @ArtificialAnlys NVIDIA Parakeet-TDT 0.6B v3 can transcribe 20 hours of speech in under 10 seconds. This deep dive shows the systems work behi

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

HardwareDGX agent

arXiv:2605.30351v1 Announce Type: cross Abstract: Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by chang

What is the best used or refurbished laptop with GPU for open source Imege generation?

HardwareDGX agent

This Reddit discussion in the StableDiffusion community addresses recommendations for affordable, used or refurbished laptops equipped with GPUs suitable for running open-source image generation model

What to expect at Computex 2026: AI chips, budget PCs competing with the MacBook Neo, Nvidia entering laptop SoC market with the rumored N1X chip, and more (PCMag)

HardwareDGX agent

PCMag: What to expect at Computex 2026: AI chips, budget PCs competing with the MacBook Neo, Nvidia entering laptop SoC market with the rumored N1X chip, and more — The theme for Computex 2026, accord

28 May 2026

A Unified Structured Query Understanding Framework for Industrial Semantic Search

HardwareDGX agent

arXiv:2605.27441v1 Announce Type: cross Abstract: Query understanding in large-scale industrial search systems is typically implemented as a cascade of disparate, task-specific components. While indiv

and see, perhaps further reflection of the death of tokenmaxxing

HardwareDGX agent

and see, perhaps further reflection of the death of tokenmaxxing The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic

H100 price: down H200 price: down tokenmaxxing: dead RoI for most customers: AWOL you can’t defy gravity forever

HardwareDGX agent

Gary Marcus discusses declining prices for NVIDIA's H100 and H200 AI chips, suggesting that inflated valuations and returns on investment for AI hardware customers are becoming unsustainable as the ma

Heterogeneous Parallelism for Multimodal Large Language Model Training

HardwareDGX agent

arXiv:2605.27678v1 Announce Type: new Abstract: Foundation model training is becoming multimodal, from post-training pipelines to large-scale pretraining. As modality coverage broadens, context window

Intel makes a bid for handheld gaming PCs with new Arc G3 processors

HardwareDGX agent

Intel introduced the Arc G-Series processors, a new family designed for next-generation handheld gaming systems, launching with Arc G3 and Arc G3 Extreme processors running Windows 11 based on Intel C

IoT Image Processing in Hybrid Cloud Environments: Use Case

HardwareDGX agent

This article likely explores practical applications of image processing for IoT devices deployed across hybrid cloud architectures, discussing how organizations can leverage both on-premises and cloud

Learning When to Optimize: Verified Optimization Skills from Expert GPU-Kernel Lineages

HardwareDGX agent

arXiv:2605.28213v1 Announce Type: new Abstract: LLM-based agents are increasingly used to generate GPU kernels, but they often know what optimizations to try without knowing when those optimizations a

Long Live the Librarian! A Persistent Search Sub-Agent for Energy-Efficient Multi-Agent Software Engineering Systems

HardwareDGX agent

arXiv:2605.27787v1 Announce Type: cross Abstract: Multi-agent systems (MAS) have substantially advanced autonomous software engineering (SWE), but their growing inference energy demands raise sustaina

NVIDIA Research Advances Robotics From Simulation to the Real World

HardwareDGX agent

Robotics is entering a new phase: moving from controlled demos and scripted automation toward generalizable, reliable embodied autonomy in the real world. At the International Conference on Robotics a

Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩 https://huggingface.co/nvidia/GLM-5.1-NVFP4

HardwareDGX agent

NVIDIA has released GLM-5.1-NVFP4, a quantized version of the GLM-5.1 model, now available on Hugging Face. The model appears to use NVFP4 (NVIDIA's floating-point 4-bit) quantization format, designed

OpenURMA: A Clean-Room Open Implementation of the Unified Bus Protocol

HardwareDGX agent

arXiv:2605.28717v1 Announce Type: new Abstract: Modern datacenter RDMA is bottlenecked at the network interface, not the wire. A NIC running RoCE or InfiniBand holds per-connection state for every (ap

Source: the Shanghai Futures Exchange is in the early stages of designing futures contracts for AI tokens; US exchanges are set to launch GPU compute futures (Reuters)

HardwareDGX agent

Reuters: Source: the Shanghai Futures Exchange is in the early stages of designing futures contracts for AI tokens; US exchanges are set to launch GPU compute futures — China is designing a futures ma

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

HardwareDGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: eve…

HardwareDGX agent

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic ass…

HardwareDGX agent

The price to rent an Nvidia H200 just collapsed from 7/hr to 4/hr in three weeks. A -40% drop in the cost of the single most strategic asset in tech. When the underlying commodity that powers your ent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

HardwareDGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

HardwareDGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation

HardwareDGX agent

arXiv:2605.27582v1 Announce Type: cross Abstract: Embodied navigation requires an agent to map language and visual observations to a stream of spatial actions that drive a real robot through environme

Utilidata raises $40M more to optimize data center power use

HardwareDGX agent

Utilidata Inc. today disclosed that it has raised 40 million in funding from Renown Capital Partners and Keyframe Capital. The cash infusion comes as an extension to a 60.3 million Series C round the

27 May 2026

AI training data provider Human Archive raises $8.2M

HardwareDGX agent

Artificial intelligence training data provider Human Archive Inc. today announced that it has raised 8.2 million in funding. Wing Venture Capital, NVP Capital, Y Combinator headlined the consortium th

Anthropic Growth and Bedrock Mix Drive AWS Margins Higher While Peers Lag

HardwareDGX agent

AWS experienced margin expansion driven by growth in Anthropic-related services and an improved product mix from Amazon Bedrock, its generative AI platform, while competing cloud providers failed to a

AssetGen: Deployable 3D Asset Generation at Interactive Speed

HardwareDGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 (Nikkei Asia)

HardwareDGX agent

Nikkei Asia: Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 — TAIPEI — Nvidia is now spend

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search

HardwareDGX agent

arXiv:2605.26636v1 Announce Type: cross Abstract: We introduce JetViT, a novel family of hybrid-architecture Vision Transformer (ViT) models that match the accuracy of state-of-the-art full-attention

← Previous
1…1415161718…29
Next →