AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,479 results
Hardware

AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization

DGX agent

arXiv:2604.18137v1 Announce Type: cross Abstract: Processing-in-Memory (PIM) architectures offer a promising solution to the memory bottlenecks in data-intensive machine learning, yet often overlook t

hardwarearxiv-cs-lg
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

DGX agent

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Building the foundation for AI across the public sector through our partner ecosystem

DGX agent

The demand for AI within the public sector has never been higher. Practitioners and CXO’s are looking for ways to harness AI to improve mission outcomes, enhance security, and streamline operations. H

hardwaregoogle-cloud-ai
21 Apr 2026
Hardware

Capacity without conflict: A guide to multi-tenant GPU cluster design for AI-native teams

DGX agent

This guide addresses the design and management of multi-tenant GPU clusters optimized for AI teams, focusing on strategies to maximize resource utilization while minimizing contention and conflicts be

hardwaretogether-ai-blog
21 Apr 2026
Hardware

Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing

DGX agent

arXiv:2604.18170v1 Announce Type: new Abstract: LLMs edit text and code by autoregressively regenerating the full output, even when most tokens appear verbatim in the input. We study Copy-as-Decode, a

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

“Cursor has also given SpaceX the right to acquire Cursor later this year for 60 billion or pay 10 billion for our work together.” persona…

DGX agent

“Cursor has also given SpaceX the right to acquire Cursor later this year for 60 billion or pay 10 billion for our work together.” personally this is the most exciting option pricing deal of the year,

hardwareswyx--x
21 Apr 2026
Hardware

Excited to partner with the SpaceX team to scale up Composer. A meaningful step on our path to build the best place to code with AI.

DGX agent

Excited to partner with the SpaceX team to scale up Composer. A meaningful step on our path to build the best place to code with AI. SpaceXAI and @cursor_ai are now working closely together to create

hardwareelon-musk--x
21 Apr 2026
Hardware

FLASH: Fast Learning via GPU-Accelerated Simulation for High-Fidelity Deformable Manipulation in Minutes

DGX agent

arXiv:2604.17513v1 Announce Type: new Abstract: Simulation frameworks such as Isaac Sim have enabled scalable robot learning for locomotion and rigid-body manipulation; however, contact-rich simulatio

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

FlexiCache: Leveraging Temporal Stability of Attention Heads for Efficient KV Cache Management

DGX agent

arXiv:2511.00868v2 Announce Type: replace Abstract: Large Language Model (LLM) serving is increasingly constrained by the growing size of the key-value (KV) cache, which scales with both context lengt

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Earth System Prediction

DGX agent

arXiv:2604.16590v1 Announce Type: new Abstract: Accurate weather and climate prediction relies on data assimilation (DA), which estimates the Earth system state by integrating observations with models

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)!

DGX agent

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)! Introducing ml-intern, the agent that just automated the post-training team @hugg

hardwareclem-delangue--x
21 Apr 2026
Hardware

Hi i'm dwarkesh! Grew up all over the US, now sf-based and always down to nerd out about AI, science & history :) a lil about me: 🟠 Host of…

DGX agent

Hi i'm dwarkesh! Grew up all over the US, now sf-based and always down to nerd out about AI, science & history :) a lil about me: 🟠 Host of the dwarkesh podcast 🟠 Studied at UT Austin 🟠 Just published

hardwaredylan-patel--x
21 Apr 2026
Hardware

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

DGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

hardwareclem-delangue--x
21 Apr 2026
Hardware

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

DGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

hardwareclem-delangue--x
21 Apr 2026
Hardware

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

DGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

hardwareclem-delangue--x
21 Apr 2026
Research

LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization

DGX agent

arXiv:2604.18117v1 Announce Type: new Abstract: Post-training quantization (PTQ) is essential for deploying large diffusion transformers on resource-constrained hardware, but aggressive 4-bit quantiza

researcharxiv-cs-lg
21 Apr 2026
Hardware

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

DGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

Neptune: Advanced ML Operator Fusion for Locality and Parallelism on GPUs

DGX agent

arXiv:2510.08726v2 Announce Type: replace-cross Abstract: Operator fusion has become a key optimization for deep learning, which combines multiple deep learning operators to improve data reuse and red

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning

DGX agent

arXiv:2509.18169v3 Announce Type: replace-cross Abstract: Tasks on complex systems require high-precision numerical computation to support decisions, but current large language models (LLMs) cannot in

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

POLAR: Online Learning for LoRA Adapter Caching and Routing in Edge LLM Serving

DGX agent

arXiv:2604.16583v1 Announce Type: new Abstract: Edge deployment of large language models (LLMs) increasingly relies on libraries of lightweight LoRA adapters, yet GPU/DRAM can keep only a small reside

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Probabilistic Programs of Thought

DGX agent

arXiv:2604.17290v1 Announce Type: new Abstract: LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to rea

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

DGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’…

DGX agent

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’s leading product and distribution to expert software engine

hardwareelon-musk--x
21 Apr 2026
Industry

Apple says Tim Cook, as executive chairman, 'will assist with certain aspects of the company, including engaging with policymakers around the world' (Marcus Mendes/9to5Mac)

DGX agent

Marcus Mendes / 9to5Mac: Apple says Tim Cook, as executive chairman, “will assist with certain aspects of the company, including engaging with policymakers around the world” — Apple confirmed today th

industrytechmeme
20 Apr 2026
Hardware

AutoFlows++: Hierarchical Message Flow Mining for System on Chip Designs

DGX agent

arXiv:2604.15359v1 Announce Type: cross Abstract: Understanding communication behavior in modern system-on-chip (SoC) designs is critical for functional verification, performance analysis, and post-si

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Autonomous AI at Scale: Adobe Agents Unlock Breakthrough Creative Intelligence With NVIDIA and WPP

DGX agent

AI agents are transforming how work gets done across all industries, accelerating everything from content creation to decision-making. NVIDIA’s expanded strategic collaborations with Adobe and WPP are

hardwarenvidia-blog
20 Apr 2026
Hardware

Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~$2.6B in its IPO, the largest in the city so far this year (Eunice Xu/South China Morning Post)

DGX agent

Eunice Xu / South China Morning Post: Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~2.6B in its IPO, the largest in the city so far this year — Nvidia suppl

hardwaretechmeme
20 Apr 2026
Hardware

Lossless Compression via Chained Lightweight Neural Predictors with Information Inheritance

DGX agent

arXiv:2604.15472v1 Announce Type: cross Abstract: This paper is dedicated to lossless data compression with probability estimation using neural networks. First, we propose a probability estimation arc

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Mitigating Indirect AGENTS.md Injection Attacks in Agentic Environments

DGX agent

NVIDIA researchers discovered a vulnerability in AI coding assistants where malicious software dependencies can inject harmful instructions into AGENTS.md configuration files, allowing attackers to re

hardwarenvidia-developer
20 Apr 2026
Hardware

NVIDIA and Partners Showcase the Future of AI-Driven Manufacturing at Hannover Messe 2026

DGX agent

Manufacturing is at an inflection point. Across every major industrial economy, the pressure to do more with less — due to faster design cycles, leaner operations and strain on skilled labor pools — i

hardwarenvidia-blog
20 Apr 2026
Hardware

PyLO: Towards Accessible Learned Optimizers in PyTorch

DGX agent

arXiv:2506.10315v3 Announce Type: replace Abstract: Learned optimizers have been an active research topic over the past decade, with increasing progress toward practical, general-purpose optimizers th

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

SK hynix says it has begun mass production of the 192GB SOCAMM2, a next-gen LPDDR5X low-power DRAM module designed particularly for Nvidia's Vera Rubin (The Korea Herald)

DGX agent

The Korea Herald: SK hynix says it has begun mass production of the 192GB SOCAMM2, a next-gen LPDDR5X low-power DRAM module designed particularly for Nvidia's Vera Rubin — SK hynix Inc. said Monday it

hardwaretechmeme
20 Apr 2026
Hardware

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, …

DGX agent

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, multi-modal hierarchical caching, and prefill-decode disaggr

hardwarekimi-moonshot--x
20 Apr 2026
Hardware

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed…

DGX agent

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed with building bigger LLMs. Trillions of parameters. Billion

hardwareyann-lecun--x
20 Apr 2026
Hardware

A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research (Abram Brown/The Information)

DGX agent

Abram Brown / The Information: A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research — For Dylan

hardwaretechmeme
19 Apr 2026
Hardware

A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M in 2026 (Emily Shugerman/The San Francisco ...)

DGX agent

Emily Shugerman / The San Francisco Standard: A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M i

hardwaretechmeme
19 Apr 2026
Hardware

Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models (Qianer Liu/The Information)

DGX agent

Qianer Liu / The Information: Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models — Google is in talk

hardwaretechmeme
19 Apr 2026
Hardware

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge back…

DGX agent

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge backlash against AI, he cobbles together this cope - to try to p

hardwaregary-marcus--x
19 Apr 2026
Hardware

Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection (Reuters)

DGX agent

Reuters: Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection — Hesai (2525.HK), China's leadi

hardwaretechmeme
18 Apr 2026
Hardware

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on…

DGX agent

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on @latentspacepod with Alessio Fanelli (@FanaHOVA) and @swyx

hardwareswyx--x
17 Apr 2026
Hardware

A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the interview (Zvi Mowshowitz/Don't Worry About the Vase)

DGX agent

Zvi Mowshowitz / Don't Worry About the Vase: A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the inter

hardwaretechmeme
17 Apr 2026
Hardware

Accelerate Clean, Modular, Nuclear Reactor Design with AI Physics

DGX agent

NVIDIA released PhysicsNeMo, an open-source AI framework designed to speed up nuclear reactor design through physics-based simulations. The framework uses machine learning-based multiphysics emulators

hardwarenvidia-developer
17 Apr 2026
Hardware

AdaSplash-2: Faster Differentiable Sparse Attention

DGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle

DGX agent

Quantum computing keeps gaining momentum even though practically it’s years away from wide commercialization, and World Quantum Day April 14 provided an excuse for a lot of announcements. Among them:

hardwaresiliconangle
17 Apr 2026
Hardware

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

DGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

cuRoboV2: Dynamics-Aware Motion Generation with Depth-Fused Distance Fields for High-DoF Robots

DGX agent

arXiv:2603.05493v2 Announce Type: replace Abstract: Effective robot autonomy requires motion generation that is safe, feasible, and reactive. Current methods are fragmented: fast planners output physi

hardwarearxiv-cs-ro
17 Apr 2026
Hardware

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

DGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

hardwarearxiv-cs-cl
17 Apr 2026
← Previous
1…4041424344…94
Next →