AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
22 Apr 2026

Earlier this month, the inaugural Runway AI Summit brought together over 700 leaders across media, entertainment, gaming and advertising. Th…

HardwareDGX agent

Earlier this month, the inaugural Runway AI Summit brought together over 700 leaders across media, entertainment, gaming and advertising. The day was filled with firesides, panels and keynotes from le

Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs

HardwareDGX agent

arXiv:2604.18788v1 Announce Type: new Abstract: Apple Neural Engine (ANE) is a dedicated neural processing unit (NPU) present in every Apple Silicon chip. Mixture-of-Experts (MoE) LLMs improve inferen

From Rainforests to Recycling Plants: 5 Ways NVIDIA AI Is Protecting the Planet

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

NVIDIA AI and accelerated computing are advancing sustainability, climate science and energy efficiency through five key applications. These include RecycleOS, an AI and robotics solution that helps r

'Funny and distressingly realistic...propelled by awesome characters and inventive twists”— @andyweirauthor Silicon Valley invents the time …

HardwareDGX agent

'Funny and distressingly realistic...propelled by awesome characters and inventive twists”— @andyweirauthor Silicon Valley invents the time machine in my upcoming book PARADOX INC, now available for p

GPU Compass – open-source, real-time GPU pricing across 20+ clouds [P]

HardwareDGX agent

GPU Compass is an open-source tool that tracks and displays real-time GPU pricing information across over 20 cloud providers. The platform likely helps machine learning practitioners and researchers c

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation

HardwareDGX agent

arXiv:2604.19167v1 Announce Type: cross Abstract: Deploying large language models (LLMs) in resource-constrained environments is hindered by heavy computational and memory requirements. We present LBL

NVIDIA and Google Cloud Collaborate to Advance Agentic and Physical AI

HardwareDGX agent

NVIDIA and Google Cloud have collaborated for more than a decade, co‑engineering a full‑stack AI platform that spans every technology layer — from performance‑optimized libraries and frameworks to ent

PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment

HardwareDGX agent

arXiv:2604.19129v1 Announce Type: new Abstract: Existing facial reenactment methods struggle with a trade-off between expressiveness and fine-grained controllability. Holistic facial reenactment model

Preserving Clusters in Error-Bounded Lossy Compression of Particle Data

HardwareDGX agent

arXiv:2604.18801v1 Announce Type: new Abstract: Lossy compression is widely used to reduce storage and I/O costs for large-scale particle datasets in scientific applications such as cosmology, molecul

Silicon Aware Neural Networks

HardwareDGX agent

arXiv:2604.19334v1 Announce Type: new Abstract: Recent work in the machine learning literature has demonstrated that deep learning can train neural networks made of discrete logic gate functions to pe

Source: Mira Murati's TML signed a deal with Google Cloud, valued in single-digit billions, to access Google's latest AI systems built on Nvidia's GB300 chips (Rebecca Bellan/TechCrunch)

HardwareDGX agent

Rebecca Bellan / TechCrunch: Source: Mira Murati's TML signed a deal with Google Cloud, valued in single-digit billions, to access Google's latest AI systems built on Nvidia's GB300 chips — Former Ope

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

HardwareDGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters

HardwareDGX agent

arXiv:2509.18831v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have significantly improved image and video synthesis. In addition, several concept control methods have b

Two new TPUs to power the next wave of AI training and inference at Google

HardwareDGX agent

Google LLC introduced two new custom silicon chips for artificial intelligence today at Google Cloud Next 2026, unveiling two distinct Tensor Processor Unit architectures built for training and infere

Vast Data, which makes software infrastructure for managing large amounts of data with a focus on AI applications, raised a 1B Series F at a 30B valuation (Kai Nicol-Schwarz/CNBC)

HardwareDGX agent

Kai Nicol-Schwarz / CNBC: Vast Data, which makes software infrastructure for managing large amounts of data with a focus on AI applications, raised a 1B Series F at a 30B valuation — Vast Data announc

We're launching two specialized TPUs for the agentic era.

HardwareDGX agent

Google announced Ironwood, its seventh-generation TPU that is twice as power efficient as the previous generation, alongside specialized hardware designed to support the emerging agentic AI era. Ironw

21 Apr 2026

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation

HardwareDGX agent

arXiv:2604.18348v1 Announce Type: new Abstract: Video diffusion transformers (DiTs) suffer from prohibitive inference latency due to quadratic attention complexity. Existing sparse attention methods e

AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization

HardwareDGX agent

arXiv:2604.18137v1 Announce Type: cross Abstract: Processing-in-Memory (PIM) architectures offer a promising solution to the memory bottlenecks in data-intensive machine learning, yet often overlook t

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

HardwareDGX agent

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

Building the foundation for AI across the public sector through our partner ecosystem

HardwareDGX agent

The demand for AI within the public sector has never been higher. Practitioners and CXO’s are looking for ways to harness AI to improve mission outcomes, enhance security, and streamline operations. H

Capacity without conflict: A guide to multi-tenant GPU cluster design for AI-native teams

HardwareDGX agent

This guide addresses the design and management of multi-tenant GPU clusters optimized for AI teams, focusing on strategies to maximize resource utilization while minimizing contention and conflicts be

Condense, Don't Just Prune: Enhancing Efficiency and Performance in MoE Layer Pruning

HardwareDGX agent

arXiv:2412.00069v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has garnered significant attention for its ability to scale up neural networks while utilizing the same or even fewer

Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing

HardwareDGX agent

arXiv:2604.18170v1 Announce Type: new Abstract: LLMs edit text and code by autoregressively regenerating the full output, even when most tokens appear verbatim in the input. We study Copy-as-Decode, a

“Cursor has also given SpaceX the right to acquire Cursor later this year for 60 billion or pay 10 billion for our work together.” persona…

HardwareDGX agent

“Cursor has also given SpaceX the right to acquire Cursor later this year for 60 billion or pay 10 billion for our work together.” personally this is the most exciting option pricing deal of the year,

Enabling AI ASICs for Zero Knowledge Proof

HardwareDGX agent

arXiv:2604.17808v1 Announce Type: cross Abstract: Zero-knowledge proof (ZKP) provers remain costly because multi-scalar multiplication (MSM) and number-theoretic transforms (NTTs) dominate runtime as

Excited to partner with the SpaceX team to scale up Composer. A meaningful step on our path to build the best place to code with AI.

HardwareDGX agent

Excited to partner with the SpaceX team to scale up Composer. A meaningful step on our path to build the best place to code with AI. SpaceXAI and @cursor_ai are now working closely together to create

FLASH: Fast Learning via GPU-Accelerated Simulation for High-Fidelity Deformable Manipulation in Minutes

HardwareDGX agent

arXiv:2604.17513v1 Announce Type: new Abstract: Simulation frameworks such as Isaac Sim have enabled scalable robot learning for locomotion and rigid-body manipulation; however, contact-rich simulatio

FlexiCache: Leveraging Temporal Stability of Attention Heads for Efficient KV Cache Management

HardwareDGX agent

arXiv:2511.00868v2 Announce Type: replace Abstract: Large Language Model (LLM) serving is increasingly constrained by the growing size of the key-value (KV) cache, which scales with both context lengt

Framework’s first eGPUs turn its laptop into a desktop PC

HardwareDGX agent

Remember when Framework made the first laptop where you can easily upgrade its entire internal video card in three minutes flat? The company's getting into the external graphics game, too. As promised

Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Earth System Prediction

HardwareDGX agent

arXiv:2604.16590v1 Announce Type: new Abstract: Accurate weather and climate prediction relies on data assimilation (DA), which estimates the Earth system state by integrating observations with models

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)!

HardwareDGX agent

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)! Introducing ml-intern, the agent that just automated the post-training team @hugg

Hi i'm dwarkesh! Grew up all over the US, now sf-based and always down to nerd out about AI, science & history :) a lil about me: 🟠 Host of…

HardwareDGX agent

Hi i'm dwarkesh! Grew up all over the US, now sf-based and always down to nerd out about AI, science & history :) a lil about me: 🟠 Host of the dwarkesh podcast 🟠 Studied at UT Austin 🟠 Just published

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

HardwareDGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

HardwareDGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

HardwareDGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

HardwareDGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

Neptune: Advanced ML Operator Fusion for Locality and Parallelism on GPUs

HardwareDGX agent

arXiv:2510.08726v2 Announce Type: replace-cross Abstract: Operator fusion has become a key optimization for deep learning, which combines multiple deep learning operators to improve data reuse and red

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning

HardwareDGX agent

arXiv:2509.18169v3 Announce Type: replace-cross Abstract: Tasks on complex systems require high-precision numerical computation to support decisions, but current large language models (LLMs) cannot in

POLAR: Online Learning for LoRA Adapter Caching and Routing in Edge LLM Serving

HardwareDGX agent

arXiv:2604.16583v1 Announce Type: new Abstract: Edge deployment of large language models (LLMs) increasingly relies on libraries of lightweight LoRA adapters, yet GPU/DRAM can keep only a small reside

Probabilistic Programs of Thought

HardwareDGX agent

arXiv:2604.17290v1 Announce Type: new Abstract: LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to rea

RACE Attention: A Strictly Linear-Time Attention Layer for Training on Outrageously Large Contexts

HardwareDGX agent

arXiv:2510.04008v5 Announce Type: replace Abstract: Softmax Attention has a quadratic time complexity in sequence length, which becomes prohibitive to run at long contexts, even with highly optimized

RainFusion2.0: Temporal-Spatial Awareness and Hardware-Efficient Block-wise Sparse Attention

HardwareDGX agent

arXiv:2512.24086v2 Announce Type: replace Abstract: In video and image generation tasks, Diffusion Transformer (DiT) models incur extremely high computational costs due to attention mechanisms, which

Real-Time Structural Detection for Indoor Navigation from 3D LiDAR Using Bird's-Eye-View Images

HardwareDGX agent

arXiv:2603.19830v2 Announce Type: replace Abstract: Efficient structural perception is essential for mapping and autonomous navigation on resource-constrained robots. Existing 3D methods are computati

Scalable and Adaptive Parallel Training of Graph Transformer on Large Graphs

HardwareDGX agent

arXiv:2604.16715v1 Announce Type: cross Abstract: Graph foundation models have demonstrated remarkable adaptability across diverse downstream tasks through large-scale pretraining on graphs. However,

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

HardwareDGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’…

HardwareDGX agent

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’s leading product and distribution to expert software engine

Web-Gewu: A Browser-Based Interactive Playground for Robot Reinforcement Learning

HardwareDGX agent

arXiv:2604.17050v1 Announce Type: new Abstract: With the rapid development of embodied intelligence, robotics education faces a dual challenge: high computational barriers and cumbersome environment c

20 Apr 2026

AutoFlows++: Hierarchical Message Flow Mining for System on Chip Designs

HardwareDGX agent

arXiv:2604.15359v1 Announce Type: cross Abstract: Understanding communication behavior in modern system-on-chip (SoC) designs is critical for functional verification, performance analysis, and post-si

Autonomous AI at Scale: Adobe Agents Unlock Breakthrough Creative Intelligence With NVIDIA and WPP

HardwareDGX agent

AI agents are transforming how work gets done across all industries, accelerating everything from content creation to decision-making. NVIDIA’s expanded strategic collaborations with Adobe and WPP are

Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~$2.6B in its IPO, the largest in the city so far this year (Eunice Xu/South China Morning Post)

HardwareDGX agent

Eunice Xu / South China Morning Post: Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~2.6B in its IPO, the largest in the city so far this year — Nvidia suppl

CPU Optimization of a Monocular 3D Biomechanics Pipeline for Low-Resource Deployment

HardwareDGX agent

arXiv:2604.15665v1 Announce Type: new Abstract: Markerless 3D movement analysis from monocular video enables accessible biomechanical assessment in clinical and sports settings. However, most research

How Much Do GPU Clusters Really Cost?

HardwareDGX agent

This article from SemiAnalysis analyzes the total cost of ownership for GPU clusters, likely covering hardware expenses, infrastructure requirements, power consumption, and operational overhead beyond

Johny Sroji, the GOAT who made Apple silicon program what it is, is now in charge of all hardware at Apple as Chief Hardware Officer

HardwareDGX agent

Johny Sroji, the GOAT who made Apple silicon program what it is, is now in charge of all hardware at Apple as Chief Hardware Officer Media Johny Srouji Taking Over as Apple's Chief Hardware Officer as

Lossless Compression via Chained Lightweight Neural Predictors with Information Inheritance

HardwareDGX agent

arXiv:2604.15472v1 Announce Type: cross Abstract: This paper is dedicated to lossless data compression with probability estimation using neural networks. First, we propose a probability estimation arc

Mitigating Indirect AGENTS.md Injection Attacks in Agentic Environments

HardwareDGX agent

NVIDIA researchers discovered a vulnerability in AI coding assistants where malicious software dependencies can inject harmful instructions into AGENTS.md configuration files, allowing attackers to re

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

HardwareDGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

NVIDIA and Partners Showcase the Future of AI-Driven Manufacturing at Hannover Messe 2026

HardwareDGX agent

Manufacturing is at an inflection point. Across every major industrial economy, the pressure to do more with less — due to faster design cycles, leaner operations and strain on skilled labor pools — i

PyLO: Towards Accessible Learned Optimizers in PyTorch

HardwareDGX agent

arXiv:2506.10315v3 Announce Type: replace Abstract: Learned optimizers have been an active research topic over the past decade, with increasing progress toward practical, general-purpose optimizers th

Scalable Posterior Uncertainty for Flexible Density-Based Clustering

HardwareDGX agent

arXiv:2603.03188v2 Announce Type: replace-cross Abstract: We introduce a novel framework for uncertainty quantification in clustering that combines martingale posterior distributions with density-base

Silicon Valley has forgotten what normal people want

HardwareDGX agent

One of the most mortifying things about knowing a lot of techies is listening to them tell me excitedly about some very important discovery that they believe they have made. Recently, I ran into an ac

← Previous
1…242526272829
Next →