AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
Hardware

Stream-CQSA: Avoiding Out-of-Memory in Attention Computation via Flexible Workload Scheduling

DGX agent

arXiv:2604.20819v1 Announce Type: new Abstract: The scalability of long-context large language models is fundamentally limited by the quadratic memory cost of exact self-attention, which often leads t

hardwarearxiv-cs-lg
23 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Tag, You’re It: GeForce NOW Levels Up Game Discovery With Xbox Game Pass and Ubisoft+ Labels

DGX agent

GeForce NOW is doubling down on what matters most: gamers. This week’s upgrades bring smarter libraries, making it easier than ever for gamers to turn a PC collection into a cloud-powered flex. It sta

hardwarenvidia-blog
23 Apr 2026
Hardware

Temporally Extended Mixture-of-Experts Models

DGX agent

arXiv:2604.20156v1 Announce Type: new Abstract: Mixture-of-Experts models, now popular for scaling capacity at fixed inference speed, switch experts at nearly every token. Once a model outgrows availa

hardwarearxiv-cs-lg
23 Apr 2026
Hardware

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind…

DGX agent

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind paper on test-time compute scaling. Here's what the agent d

hardwareclem-delangue--x
23 Apr 2026
Hardware

US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases (Alexandra Alper/Reuters)

DGX agent

Alexandra Alper / Reuters: US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases — Nvidia's (

hardwaretechmeme
23 Apr 2026
Hardware

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a gre…

DGX agent

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a great business strategy 🙈 So now we have 2 free opus sessions/d

hardwareclem-delangue--x
23 Apr 2026
Hardware

We tried a new thing with NVIDIA to roll out Codex across a whole company and it was awesome to see it work. Let us know if you'd like to do…

DGX agent

Sam Altman posted about OpenAI's successful pilot program with NVIDIA to deploy Codex (an AI code generation model) across an entire company, highlighting positive results from the enterprise rollout.

hardwaresam-altman--x
23 Apr 2026
Hardware

AdaGScale: Viewpoint-Adaptive Gaussian Scaling in 3D Gaussian Splatting to Reduce Gaussian-Tile Pairs

DGX agent

arXiv:2604.18980v1 Announce Type: new Abstract: Reducing the number of Gaussian-tile pairs is one of the most promising approaches to improve 3D Gaussian Splatting (3D-GS) rendering speed on GPUs. How

hardwarearxiv-cs-cv
22 Apr 2026
Hardware

Advancing Emerging Optimizers for Accelerated LLM Training with NVIDIA Megatron

DGX agent

This blog post discusses higher-order optimization algorithms such as Shampoo that have been applied in neural network training for at least a decade. The article explores how these emerging optimizer

hardwarenvidia-developer
22 Apr 2026
Hardware

ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants

DGX agent

arXiv:2604.18616v1 Announce Type: cross Abstract: LLM-based coding agents can generate functionally correct GPU kernels, yet their performance remains far below hand-optimized libraries on critical co

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

Compute is king. I’m excited about the next phase of growth for Cursor, as their close partner.

DGX agent

Compute is king. I’m excited about the next phase of growth for Cursor, as their close partner. SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge

hardwaresonya-huang--x
22 Apr 2026
Hardware

@dwarkesh_sp @dylan522p No catfishing here, we're pixelmaxxing!

DGX agent

This post appears to be a humorous take on AI optimization and technology trends, likely referencing 'pixel-maxxing' (maximizing pixel-level performance or visual quality) in contrast to 'catfishing'

hardwaredylan-patel--x
22 Apr 2026
Hardware

Earlier this month, the inaugural Runway AI Summit brought together over 700 leaders across media, entertainment, gaming and advertising. Th…

DGX agent

Earlier this month, the inaugural Runway AI Summit brought together over 700 leaders across media, entertainment, gaming and advertising. The day was filled with firesides, panels and keynotes from le

hardwarecristobal-valenzuela--x
22 Apr 2026
Hardware

Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs

DGX agent

arXiv:2604.18788v1 Announce Type: new Abstract: Apple Neural Engine (ANE) is a dedicated neural processing unit (NPU) present in every Apple Silicon chip. Mixture-of-Experts (MoE) LLMs improve inferen

hardwarearxiv-cs-lg
22 Apr 2026
Hardware

From Rainforests to Recycling Plants: 5 Ways NVIDIA AI Is Protecting the Planet

DGX agent

NVIDIA AI and accelerated computing are advancing sustainability, climate science and energy efficiency through five key applications. These include RecycleOS, an AI and robotics solution that helps r

hardwarenvidia-blog
22 Apr 2026
Hardware

'Funny and distressingly realistic...propelled by awesome characters and inventive twists”— @andyweirauthor Silicon Valley invents the time …

DGX agent

'Funny and distressingly realistic...propelled by awesome characters and inventive twists”— @andyweirauthor Silicon Valley invents the time machine in my upcoming book PARADOX INC, now available for p

hardwareswyx--x
22 Apr 2026
Hardware

GPU Compass – open-source, real-time GPU pricing across 20+ clouds [P]

DGX agent

GPU Compass is an open-source tool that tracks and displays real-time GPU pricing information across over 20 cloud providers. The platform likely helps machine learning practitioners and researchers c

hardwarer-machinelearning
22 Apr 2026
Hardware

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation

DGX agent

arXiv:2604.19167v1 Announce Type: cross Abstract: Deploying large language models (LLMs) in resource-constrained environments is hindered by heavy computational and memory requirements. We present LBL

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

NVIDIA and Google Cloud Collaborate to Advance Agentic and Physical AI

DGX agent

NVIDIA and Google Cloud have collaborated for more than a decade, co‑engineering a full‑stack AI platform that spans every technology layer — from performance‑optimized libraries and frameworks to ent

hardwarenvidia-blog
22 Apr 2026
Hardware

PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment

DGX agent

arXiv:2604.19129v1 Announce Type: new Abstract: Existing facial reenactment methods struggle with a trade-off between expressiveness and fine-grained controllability. Holistic facial reenactment model

hardwarearxiv-cs-cv
22 Apr 2026
Hardware

Preserving Clusters in Error-Bounded Lossy Compression of Particle Data

DGX agent

arXiv:2604.18801v1 Announce Type: new Abstract: Lossy compression is widely used to reduce storage and I/O costs for large-scale particle datasets in scientific applications such as cosmology, molecul

hardwarearxiv-cs-lg
22 Apr 2026
Hardware

Silicon Aware Neural Networks

DGX agent

arXiv:2604.19334v1 Announce Type: new Abstract: Recent work in the machine learning literature has demonstrated that deep learning can train neural networks made of discrete logic gate functions to pe

hardwarearxiv-cs-cv
22 Apr 2026
Hardware

Source: Mira Murati's TML signed a deal with Google Cloud, valued in single-digit billions, to access Google's latest AI systems built on Nvidia's GB300 chips (Rebecca Bellan/TechCrunch)

DGX agent

Rebecca Bellan / TechCrunch: Source: Mira Murati's TML signed a deal with Google Cloud, valued in single-digit billions, to access Google's latest AI systems built on Nvidia's GB300 chips — Former Ope

hardwaretechmeme
22 Apr 2026
Hardware

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

DGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters

DGX agent

arXiv:2509.18831v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have significantly improved image and video synthesis. In addition, several concept control methods have b

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

Two new TPUs to power the next wave of AI training and inference at Google

DGX agent

Google LLC introduced two new custom silicon chips for artificial intelligence today at Google Cloud Next 2026, unveiling two distinct Tensor Processor Unit architectures built for training and infere

hardwaresiliconangle
22 Apr 2026
Hardware

Vast Data, which makes software infrastructure for managing large amounts of data with a focus on AI applications, raised a 1B Series F at a 30B valuation (Kai Nicol-Schwarz/CNBC)

DGX agent

Kai Nicol-Schwarz / CNBC: Vast Data, which makes software infrastructure for managing large amounts of data with a focus on AI applications, raised a 1B Series F at a 30B valuation — Vast Data announc

hardwaretechmeme
22 Apr 2026
Hardware

We're launching two specialized TPUs for the agentic era.

DGX agent

Google announced Ironwood, its seventh-generation TPU that is twice as power efficient as the previous generation, alongside specialized hardware designed to support the emerging agentic AI era. Ironw

hardwaregoogle-ai
22 Apr 2026
Hardware

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation

DGX agent

arXiv:2604.18348v1 Announce Type: new Abstract: Video diffusion transformers (DiTs) suffer from prohibitive inference latency due to quadratic attention complexity. Existing sparse attention methods e

hardwarearxiv-cs-cv
21 Apr 2026
Hardware

AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization

DGX agent

arXiv:2604.18137v1 Announce Type: cross Abstract: Processing-in-Memory (PIM) architectures offer a promising solution to the memory bottlenecks in data-intensive machine learning, yet often overlook t

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

DGX agent

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Building the foundation for AI across the public sector through our partner ecosystem

DGX agent

The demand for AI within the public sector has never been higher. Practitioners and CXO’s are looking for ways to harness AI to improve mission outcomes, enhance security, and streamline operations. H

hardwaregoogle-cloud-ai
21 Apr 2026
Hardware

Capacity without conflict: A guide to multi-tenant GPU cluster design for AI-native teams

DGX agent

This guide addresses the design and management of multi-tenant GPU clusters optimized for AI teams, focusing on strategies to maximize resource utilization while minimizing contention and conflicts be

hardwaretogether-ai-blog
21 Apr 2026
Hardware

Condense, Don't Just Prune: Enhancing Efficiency and Performance in MoE Layer Pruning

DGX agent

arXiv:2412.00069v3 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has garnered significant attention for its ability to scale up neural networks while utilizing the same or even fewer

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing

DGX agent

arXiv:2604.18170v1 Announce Type: new Abstract: LLMs edit text and code by autoregressively regenerating the full output, even when most tokens appear verbatim in the input. We study Copy-as-Decode, a

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

“Cursor has also given SpaceX the right to acquire Cursor later this year for 60 billion or pay 10 billion for our work together.” persona…

DGX agent

“Cursor has also given SpaceX the right to acquire Cursor later this year for 60 billion or pay 10 billion for our work together.” personally this is the most exciting option pricing deal of the year,

hardwareswyx--x
21 Apr 2026
Hardware

Enabling AI ASICs for Zero Knowledge Proof

DGX agent

arXiv:2604.17808v1 Announce Type: cross Abstract: Zero-knowledge proof (ZKP) provers remain costly because multi-scalar multiplication (MSM) and number-theoretic transforms (NTTs) dominate runtime as

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

Excited to partner with the SpaceX team to scale up Composer. A meaningful step on our path to build the best place to code with AI.

DGX agent

Excited to partner with the SpaceX team to scale up Composer. A meaningful step on our path to build the best place to code with AI. SpaceXAI and @cursor_ai are now working closely together to create

hardwareelon-musk--x
21 Apr 2026
Hardware

FLASH: Fast Learning via GPU-Accelerated Simulation for High-Fidelity Deformable Manipulation in Minutes

DGX agent

arXiv:2604.17513v1 Announce Type: new Abstract: Simulation frameworks such as Isaac Sim have enabled scalable robot learning for locomotion and rigid-body manipulation; however, contact-rich simulatio

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

FlexiCache: Leveraging Temporal Stability of Attention Heads for Efficient KV Cache Management

DGX agent

arXiv:2511.00868v2 Announce Type: replace Abstract: Large Language Model (LLM) serving is increasingly constrained by the growing size of the key-value (KV) cache, which scales with both context lengt

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Framework’s first eGPUs turn its laptop into a desktop PC

DGX agent

Remember when Framework made the first laptop where you can easily upgrade its entire internal video card in three minutes flat? The company's getting into the external graphics game, too. As promised

hardwarethe-verge-ai
21 Apr 2026
Hardware

Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Earth System Prediction

DGX agent

arXiv:2604.16590v1 Announce Type: new Abstract: Accurate weather and climate prediction relies on data assimilation (DA), which estimates the Earth system state by integrating observations with models

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)!

DGX agent

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)! Introducing ml-intern, the agent that just automated the post-training team @hugg

hardwareclem-delangue--x
21 Apr 2026
Hardware

Hi i'm dwarkesh! Grew up all over the US, now sf-based and always down to nerd out about AI, science & history :) a lil about me: 🟠 Host of…

DGX agent

Hi i'm dwarkesh! Grew up all over the US, now sf-based and always down to nerd out about AI, science & history :) a lil about me: 🟠 Host of the dwarkesh podcast 🟠 Studied at UT Austin 🟠 Just published

hardwaredylan-patel--x
21 Apr 2026
Hardware

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

DGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

hardwareclem-delangue--x
21 Apr 2026
Hardware

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

DGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

hardwareclem-delangue--x
21 Apr 2026
Hardware

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

DGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

hardwareclem-delangue--x
21 Apr 2026
Hardware

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

DGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

hardwarearxiv-cs-ro
21 Apr 2026
← Previous
1…3031323334…37
Next →