AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
5 May 2026

Effective Capacitance Modeling Using Graph Neural Networks

HardwareDGX agent

arXiv:2507.03787v2 Announce Type: replace Abstract: Static timing analysis is a crucial stage in the VLSI design flow that verifies the timing correctness of circuits. Timing analysis depends on the p

Fast Log-Domain Sinkhorn Optimal Transport with Warp-Level GPU Reductions

HardwareDGX agent

arXiv:2605.00837v1 Announce Type: new Abstract: Entropic regularized optimal transport (OT) via the Sinkhorn algorithm has become a fundamental tool in machine learning, yet existing implementations e

How to Build In-Vehicle AI Agents with NVIDIA: From Cloud to Car

HardwareDGX agent

The automotive cockpit is undergoing a fundamental shift from rule-based interfaces to agentic, multimodal AI systems capable of reasoning, planning, and acting. This is enabled by an agentic AI pipel

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Inference is 80-90% of the lifetime cost of a production AI system. Most AI-native teams are leaving performance and margin on the table. He…

HardwareDGX agent

Inference is 80-90% of the lifetime cost of a production AI system. Most AI-native teams are leaving performance and margin on the table. Here’s how Together AI, the AI Native Cloud, fixes that on @nv

InfiniteDiffusion: Bridging Learned Fidelity and Procedural Utility for Open-World Terrain Generation

HardwareDGX agent

arXiv:2512.08309v4 Announce Type: replace Abstract: For decades, procedural worlds have been built on procedural noise functions such as Perlin noise, which are fast and infinite, yet fundamentally li

Lambda assembles leadership team to power gigawatt-scale AI infrastructure for the superintelligence era

HardwareDGX agent

Lambda Labs has assembled a new leadership team to develop and manage large-scale AI infrastructure capable of supporting gigawatt-level power consumption for advanced AI models. The company is positi

LEAP: Layer-wise Exit-Aware Pretraining for Efficient Transformer Inference

HardwareDGX agent

arXiv:2605.01058v1 Announce Type: cross Abstract: Layer-aligned distillation and convergence-based early exit represent two predominant computational efficiency paradigms for transformer inference; ye

LLM Output Detectability and Task Performance Can be Jointly Optimized

HardwareDGX agent

arXiv:2605.01350v1 Announce Type: new Abstract: Detecting machine-generated text is essential for transparency and accountability when deploying large language models (LLMs). Among detection approache

Maistros: A Greek Large Language Model Adapted Through Knowledge Distillation From Large Reasoning Models

HardwareDGX agent

arXiv:2605.01870v1 Announce Type: new Abstract: Large Language Models (LLMs) have substantially advanced the field of Natural Language Processing (NLP), achieving state-of-the-art performance across a

MusicInfuser: Making Video Diffusion Listen and Dance

HardwareDGX agent

arXiv:2503.14505v3 Announce Type: replace Abstract: We introduce MusicInfuser, an approach that aligns pre-trained text-to-video diffusion models to generate high-quality dance videos synchronized wit

Need for Speed: Zero-Shot Depth Completion with Single-Step Diffusion

HardwareDGX agent

arXiv:2603.10584v2 Announce Type: replace Abstract: We introduce Marigold-SSD, a single-step, late-fusion depth completion framework that leverages strong diffusion priors while eliminating the costly

No GPU utilization

HardwareDGX agent

Ollama not using GPU is commonly diagnosed by running `ollama ps`—if it shows 100% CPU, Ollama isn't detecting the GPU. The issue typically has multiple causes, with common ones including driver probl

Nscale agrees to spend €695M on infrastructure in Portugal in an expansion of its Microsoft partnership, supplying 66K+ Nvidia Rubin GPUs from late 2027 (Paula Doenecke/Bloomberg)

HardwareDGX agent

Paula Doenecke / Bloomberg: Nscale agrees to spend €695M on infrastructure in Portugal in an expansion of its Microsoft partnership, supplying 66K+ Nvidia Rubin GPUs from late 2027 — Nscale Global Hol

Nvidia, AMD back $100M round for AI tooling startup RadixArk

HardwareDGX agent

RadixArk Inc., a startup that provides tools for artificial intelligence developers, has raised 100 million from a group of high-profile backers. Nvidia Corp.’s NVentures fund led the seed round with

NVIDIA and ServiceNow Partner on New Autonomous AI Agents for Enterprises

HardwareDGX agent

Enterprise AI has learned to generate. It has learned to reason. Now companies are asking the next question: How should AI act? Early agent systems have shown what’s possible, moving beyond simple pro

Rethinking Network Topologies for Cost-Effective Mixture-of-Experts LLM Serving

HardwareDGX agent

arXiv:2605.00254v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) architectures have turned LLM serving into a cluster-scale workload in which communication consumes a considerable portion of

SplitZip: Ultra Fast Lossless KV Compression for Disaggregated LLM Serving

HardwareDGX agent

arXiv:2605.01708v1 Announce Type: cross Abstract: Contemporary systems serving large language models (LLMs) have adopted prefill-decode disaggregation to better load-balance between the compute-bound

Stochastic Sparse Attention for Memory-Bound Inference

HardwareDGX agent

arXiv:2605.01910v1 Announce Type: new Abstract: Autoregressive decoding becomes bandwidth-limited at long contexts, as generating each token requires reading all n_k key and value vectors from KV cach

The extit{Silicon Society} Cookbook: Design Space of LLM-based Social Simulations

HardwareDGX agent

arXiv:2605.00197v1 Announce Type: cross Abstract: Studies attempting to simulate human behavior with extit{Silicon Societies} grow in numbers while LLM-only social networks have started appearing outs

toodles from mickey mouse clubhouse was weirdly ahead of its time wake phrase: mickey mouse clubhouse launched in 2006, and “oh toodles” tra…

Model ReleasesDGX agent

toodles from mickey mouse clubhouse was weirdly ahead of its time wake phrase: mickey mouse clubhouse launched in 2006, and “oh toodles” trained toddlers on the assistant wake phrase years before siri

We cut model cold-start times from minutes to seconds. 60x faster, saving thousands of GPU-minutes and hundreds of terabytes of transfers ev…

HardwareDGX agent

We cut model cold-start times from minutes to seconds. 60x faster, saving thousands of GPU-minutes and hundreds of terabytes of transfers every day. It turns out GPUs already holding weights are faste

4 May 2026

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments

HardwareDGX agent

arXiv:2605.00650v1 Announce Type: new Abstract: Fine-tuning LLMs is necessary for various dedicated downstream tasks, but classic backpropagation-based fine-tuning methods require substantial GPU memo

Boosting multimodal inference performance by >10% with a single Python dictionary

HardwareDGX agent

This Modal blog post describes a performance optimization technique for multimodal AI inference that achieves over 10% improvement through a simple Python dictionary-based approach. The article likely

Cerebras seeks a valuation of up to 26.62B in its US IPO, aiming to raise 3.5B by selling 28M shares at 115 to 125 apiece in its second attempt to go public (Reuters)

HardwareDGX agent

Reuters: Cerebras seeks a valuation of up to 26.62B in its US IPO, aiming to raise 3.5B by selling 28M shares at 115 to 125 apiece in its second attempt to go public — Nvidia-rival Cerebras is seeking

Down payment on a EUV machine!

HardwareDGX agent

Dylan Patel likely announced or commented on a significant financial commitment or investment milestone related to an EUV (Extreme Ultraviolet) lithography machine, which are critical semiconductor ma

Lattice Semiconductor agrees to acquire AMI for $1.65B in cash and stock; Georgia-based AMI provides firmware and infrastructure manageability for cloud and AI (Mike Rogoway/Oregonian)

HardwareDGX agent

Mike Rogoway / Oregonian: Lattice Semiconductor agrees to acquire AMI for 1.65B in cash and stock; Georgia-based AMI provides firmware and infrastructure manageability for cloud and AI — Hillsboro-bas

Legislators and experts criticize the EU's €20B sovereign compute data center plan, questioning whether there is demand and the plan's reliance on Nvidia GPUs (Pieter Haeck/Politico)

HardwareDGX agent

Pieter Haeck / Politico: Legislators and experts criticize the EU's €20B sovereign compute data center plan, questioning whether there is demand and the plan's reliance on Nvidia GPUs — Critics argue

MiniVLA-Nav v1: A Multi-Scene Simulation Dataset for Language-Conditioned Robot Navigation

HardwareDGX agent

arXiv:2605.00397v1 Announce Type: new Abstract: We present MiniVLA-Nav v1, a simulation dataset for Language-Conditioned Object Approach (LCOA) navigation: given a short natural-language instruction,

Optimize Supply Chain Decision Systems Using NVIDIA cuOpt Agent Skills

HardwareDGX agent

NVIDIA cuOpt Agent Skills integrate LLM reasoning with GPU-accelerated solvers to enable AI agents to translate natural language supply chain problems into optimized mathematical models, with skills e

The case against OpenAI is getting markedly stronger now that Musk is off the stand. Why? Musk’s lawyer is interrogating OpenAI founder Greg…

HardwareDGX agent

The case against OpenAI is getting markedly stronger now that Musk is off the stand. Why? Musk’s lawyer is interrogating OpenAI founder Greg Brockman, making clear that OpenAI sold its mission as a no

3 May 2026

Analysis: Asian suppliers account for ~90% of Nvidia's production costs, up from 65% in 2025, as latest wave of collaborations shifts from chips to physical AI (Abhishek Vishnoi/Bloomberg)

HardwareDGX agent

Abhishek Vishnoi / Bloomberg: Analysis: Asian suppliers account for ~90% of Nvidia's production costs, up from 65% in 2025, as latest wave of collaborations shifts from chips to physical AI — The list

Are babies and relationships measured on the same time scale? After 18-24 months you start measuring in years.

HardwareDGX agent

This post draws a comparison between how parents measure their babies' development and relationship milestones, noting that both shift from month-based to year-based metrics around the 18-24 month mar

NVIDIA's New AI Turns One Photo Into A World That Never Breaks

HardwareDGX agent

NVIDIA's Instant Splat is a breakthrough AI technique that synthesizes missing information from a few images to create immersive 3D models of virtual worlds. The method produces high-quality 3D models

People hate on SF men for group think, but every girl in SF has Tabi's now It's an epidemic

HardwareDGX agent

This post humorously critiques conformity in San Francisco by observing that while men in the tech scene are often stereotyped as groupthink followers, women in SF have similarly adopted Tabis (Japane

Qwen3.6 vs gpt-oss:120b on Apple Silicon — three Qwen variants benchmarked, plus what works and where it does not

Model ReleasesDGX agent

This post benchmarks three Qwen3.6 model variants against gpt-oss:120b when running on Apple Silicon hardware, evaluating their performance characteristics and practical usability. It documents both t

The number of jobs in the future is endless because the problems to solve are endless. Jobs multiply as we get more complex. No AI or human …

HardwareDGX agent

The number of jobs in the future is endless because the problems to solve are endless. Jobs multiply as we get more complex. No AI or human can solve all problems and all the work to do in the Univers

Whenever I make a group chat green, I now get to proudly proclaim it as goblin mode

HardwareDGX agent

The post humorously references Apple's iMessage feature that displays group chats in green when they include non-iPhone users, joking that activating 'goblin mode' (a playful term for behaving mischie

2 May 2026

CUDA V.13?

HardwareDGX agent

Ollama's MLX engine runs on NVIDIA GPUs via CUDA v13 on Windows and Linux. Users have reported that Ollama crashes on RTX 3060 with cuda_v13 in versions 0.13.0 through 0.15.6, while deleting the cuda_

Intel Inside the Micro Revolution: 8008 Origins

HardwareDGX agent

This article explores the origins and development of Intel's 8008 microprocessor, a pioneering chip that played a crucial role in the early microcomputer revolution of the 1970s. It likely examines th

Sources: Anthropic is in early talks to buy AI inference chips from UK-based Fractile when they become available in 2027 (The Information)

HardwareDGX agent

The Information: Sources: Anthropic is in early talks to buy AI inference chips from UK-based Fractile when they become available in 2027 — As Anthropic's sales explode, straining the servers it uses,

1 May 2026

AI Value Capture - The Shift To Model Labs Vera Rubin VR NVL72: V for Value - Rubin delivers a step jump in performance per TCO. ROI accruin…

HardwareDGX agent

AI Value Capture - The Shift To Model Labs Vera Rubin VR NVL72: V for Value - Rubin delivers a step jump in performance per TCO. ROI accruing to users, Neoclouds, Hyperscalers, AI Labs, Memory Vendors

Benchmarking Deep Learning Models for Object Detection on Edge Computing Devices

HardwareDGX agent

arXiv:2409.16808v1 Announce Type: cross Abstract: Modern applications, such as autonomous vehicles, require deploying deep learning algorithms on resource-constrained edge devices for real-time image

By popular request, YOLO-mode has landed on ml-intern https://smolagents-ml-intern.hf.space This enables the intern to execute long-running …

HardwareDGX agent

By popular request, YOLO-mode has landed on ml-intern https://smolagents-ml-intern.hf.space This enables the intern to execute long-running tasks like launching parallel ablations to determine the opt

First Silicon Valley captured the US Government, and now it is capturing “independent” media.

HardwareDGX agent

First Silicon Valley captured the US Government, and now it is capturing “independent” media. 🟡 Semafor is launching a new initiative: Silicon Valley & The World, bringing together leaders building ar

FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems

HardwareDGX agent

arXiv:2604.28156v1 Announce Type: cross Abstract: We present FlexiTac, a low-cost, open-source, and scalable piezoresistive tactile sensing solution designed for robotic end-effectors. FlexiTac is a p

GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2

HardwareDGX agent

arXiv:2509.16248v3 Announce Type: replace-cross Abstract: This paper presents GRAPHMEND, a high-level compiler technique that eliminates FX graph breaks in PyTorch 2 programs. Although PyTorch 2 intro

Marketing Architects Scales Real-Time CTV Advertising with Vultr Infrastructure

HardwareDGX agent

Discover how Marketing Architects scaled real-time CTV advertising with Vultr, doubling throughput, reducing latency by 50%, and cutting cloud costs by up to 50% with high-performance Kubernetes and c

MARS: Efficient, Adaptive Co-Scheduling for Heterogeneous Agentic Systems

HardwareDGX agent

arXiv:2604.26963v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as the execution core of autonomous agents rather than as standalone text generators. Agentic w

More money for worse work that you have to fix, good stuff this AI thing, thanks Nvidia.

HardwareDGX agent

Gary Marcus critiques AI systems (particularly Nvidia's offerings) for producing lower-quality outputs that require significant correction and refinement despite increased computational investment. Th

Pentagon inks AI procurement deals with seven companies, leaves out Anthropic

HardwareDGX agent

The U.S. Defense Department today announced that it has inked artificial intelligence procurement contracts with seven tech firms. The group includes Amazon Web Service Inc., Google LLC, Microsoft Cor

Pentagon strikes classified AI deals with OpenAI, Google, and Nvidia — but not Anthropic

HardwareDGX agent

The Pentagon has struck deals with OpenAI, Google, Microsoft, Amazon, Nvidia, Elon Musk's xAI, and the startup Reflection, allowing the agency to use their AI tools in classified settings, according t

Performance-Driven QUBO for Recommender Systems on Quantum Annealers

SafetyDGX agent

arXiv:2410.15272v3 Announce Type: replace-cross Abstract: Quantum annealers offer a promising hardware platform for solving combinatorial optimization problems, especially those formulated as Quadrati

Sources: Huawei expects AI chip revenue to hit ~12B in 2026, up 60% from 7.5B in 2025, as orders for its Ascend 950PR chip surge and Nvidia stalls in China (Zijing Wu/Financial Times)

HardwareDGX agent

Zijing Wu / Financial Times: Sources: Huawei expects AI chip revenue to hit ~12B in 2026, up 60% from 7.5B in 2025, as orders for its Ascend 950PR chip surge and Nvidia stalls in China — Chinese tech

Sources: the US DOD strikes agreements with Nvidia, Microsoft, Reflection AI, and AWS to use their AI tools on classified military networks for 'lawful' use (Katrina Manson/Bloomberg)

HardwareDGX agent

Katrina Manson / Bloomberg: Sources: the US DOD strikes agreements with Nvidia, Microsoft, Reflection AI, and AWS to use their AI tools on classified military networks for “lawful” use — The Pentagon

Strait: Perceiving Priority and Interference in ML Inference Serving

HardwareDGX agent

arXiv:2604.28175v1 Announce Type: new Abstract: Machine learning (ML) inference serving systems host deep neural network (DNN) models and schedule incoming inference requests across deployed GPUs. How

ZipCCL: Efficient Lossless Data Compression of Communication Collectives for Accelerating LLM Training

HardwareDGX agent

arXiv:2604.27844v1 Announce Type: cross Abstract: Communication has emerged as a critical bottleneck in the distributed training of large language models (LLMs). While numerous approaches have been pr

30 Apr 2026

10 years ago I took Justin's C++ 11 class at Google. He is a gifted teacher. Watch this talk where he explains the 500B AI build out from 1s…

HardwareDGX agent

10 years ago I took Justin's C++ 11 class at Google. He is a gifted teacher. Watch this talk where he explains the 500B AI build out from 1st principles. One gpu to world wide racks and why some of th

A projection-based framework for gradient-free and parallel learning

HardwareDGX agent

arXiv:2506.05878v2 Announce Type: replace Abstract: We present a feasibility-seeking approach to neural network training. This mathematical optimization framework is distinct from conventional gradien

AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control

HardwareDGX agent

arXiv:2601.11568v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) is highly memory-intensive due to optimizer state overhead. The FRUGAL framework mitigates this with gra

Automating GPU Kernel Translation with AI Agents: cuTile Python to cuTile.jl

HardwareDGX agent

The TileGym project developed an AI-driven skill-based workflow that encodes 17 critical translation rules, static validation scripts, and example kernels, enabling automated conversion of cuTile Pyth

← Previous
1…2930313233…75
Next →