AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
18 May 2026

Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs

HardwareDGX agent

The first NVIDIA Vera CPUs arrived at three of the world's leading AI labs on Friday — Anthropic in San Francisco, OpenAI in Mission Bay, SpaceXAI in Palo Alto — followed by a delivery to Oracle Cloud

16 May 2026

Reduce your GPU power limit

HardwareDGX agent

Setting a GPU power limit using nvidia-smi reduces heat output by approximately 20% with only a 5-8% inference speed loss. Undervolting the GPU can reduce power consumption by 5-15% with zero performa

SF vibes are frenetic over the huge divide in outcomes and career uncertainty for software engineers; over 5 years ~10K people in AI attained retirement wealth (Deedy/@deedydas)

Hardware

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

Deedy / @deedydas: SF vibes are frenetic over the huge divide in outcomes and career uncertainty for software engineers; over 5 years ~10K people in AI attained retirement wealth — The vibes in SF fee

Shenzhen-listed RoboTechnik, which claims to be the largest silicon photonics tool maker and whose stock is up 340% over the past year, files for a HK listing (Zinnia Lee/Forbes)

HardwareDGX agent

Zinnia Lee / Forbes: Shenzhen-listed RoboTechnik, which claims to be the largest silicon photonics tool maker and whose stock is up 340% over the past year, files for a HK listing — RoboTechnik Intell

15 May 2026

A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 (Rebecca Bellan/TechCrunch)

HardwareDGX agent

Rebecca Bellan / TechCrunch: A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 — Every major A

An Amortized Efficiency Threshold for Comparing Neural and Heuristic Solvers in Combinatorial Optimization

HardwareDGX agent

arXiv:2605.14624v1 Announce Type: cross Abstract: A common critique of neural combinatorial-optimization solvers is that they are less energy-efficient than CPU metaheuristics, given the operational e

BEAM: Binary Expert Activation Masking for Dynamic Routing in MoE

HardwareDGX agent

arXiv:2605.14438v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enhance the efficiency of large language models by activating only a subset of experts per token. However, standa

DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration

HardwareDGX agent

arXiv:2605.14526v1 Announce Type: cross Abstract: Differentiable simulation of soft bodies is a foundation for system identification, trajectory optimization, and Real2Sim transfer. Yet, existing meth

EMA: Efficient Model Adaptation for Learning-based Systems

HardwareDGX agent

arXiv:2605.13942v1 Announce Type: new Abstract: Machine learning (ML) is increasingly applied to optimize system performance in tasks such as resource management and network simulation. Unlike traditi

EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization

HardwareDGX agent

arXiv:2605.14249v1 Announce Type: new Abstract: We present EnergyLens, an end-to-end framework for energy-aware large language model (LLM) inference optimization. As LLMs scale, predicting and reducin

Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI

HardwareDGX agent

arXiv:2605.14665v1 Announce Type: new Abstract: Legal reasoning is not semantic similarity search. A court judgment encodes constrained symbolic reasoning: precedent propagation, procedural state tran

FALO: Fast and Accurate LiDAR 3D Object Detection on Resource-Constrained Devices

HardwareDGX agent

arXiv:2506.04499v2 Announce Type: replace Abstract: Existing LiDAR 3D object detection methods predominantely rely on sparse convolutions and/or transformers, which can be challenging to run on resour

GenAI for Energy-Efficient and Interference-Aware Compressed Sensing of GNSS Signals on a Google Edge TPU

HardwareDGX agent

arXiv:2605.14839v1 Announce Type: new Abstract: Traditional methods for classifying global navigation satellite system (GNSS) jamming signals typically involve post-processing raw or spectral data str

GridCare raises $64M to speed up AI data center projects

HardwareDGX agent

GridCare Inc., a startup that helps data center operators more quickly connect their facilities to the electrical grid, has raised 64 million in funding. Early Nvidia Corp. backer Sutter Hill Ventures

Neural Field Thermal Tomography: A Differentiable Physics Framework for Non-Destructive Evaluation

HardwareDGX agent

arXiv:2603.11045v2 Announce Type: replace-cross Abstract: Inverse problems for stiff parabolic partial differential equations (PDEs), such as the inverse heat conduction problem (IHCP), are severely i

Nvidia's future in China remains unclear after the Trump-Xi Summit; Trump says China 'chose not to' buy Nvidia chips as 'they want to try to develop their own' (Meaghan Tobin/New York Times)

HardwareDGX agent

Meaghan Tobin / New York Times: Nvidia's future in China remains unclear after the Trump-Xi Summit; Trump says China “chose not to” buy Nvidia chips as “they want to try to develop their own” — The st

Parallelizing Counterfactual Regret Minimization

HardwareDGX agent

arXiv:2605.14277v1 Announce Type: new Abstract: Parallelization has played an instrumental role in the field of artificial intelligence (AI), drastically reducing the time taken to train and evaluate

Semiconductor stocks fell globally on Friday after the Trump-Xi summit concluded without major chip deals; Nvidia closed down 4.42% and AMD closed down 5.69% (Mauro Orru/Wall Street Journal)

HardwareDGX agent

Mauro Orru / Wall Street Journal: Semiconductor stocks fell globally on Friday after the Trump-Xi summit concluded without major chip deals; Nvidia closed down 4.42% and AMD closed down 5.69% — Beijin

Seoul-based WIRobotics, which develops wearable and humanoid robots and is collaborating with Nvidia and AWS, raised a ~$68M Series B led by JB Investment (Lee Jaewoon/The Elec)

HardwareDGX agent

Lee Jaewoon / The Elec: Seoul-based WIRobotics, which develops wearable and humanoid robots and is collaborating with Nvidia and AWS, raised a ~$68M Series B led by JB Investment — Company to accelera

SurgicalMamba: Dual-Path SSD with State Regramming for Online Surgical Phase Recognition

HardwareDGX agent

arXiv:2605.14889v1 Announce Type: cross Abstract: Online surgical phase recognition (SPR) underpins context-aware operating-room systems and requires committing to a prediction at every frame from pas

Synthetic Sociality: How Generative Models Privatize the Social Fabric

HardwareDGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

Towards Resource-Efficient LLMs: End-to-End Energy Accounting of Distillation Pipelines

HardwareDGX agent

arXiv:2605.13981v1 Announce Type: cross Abstract: The rise in deployment of large language models has driven a surge in GPU demand and datacenter scaling, raising concerns about electricity use, grid

Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles

HardwareDGX agent

arXiv:2605.14578v1 Announce Type: new Abstract: Partial Dependence Plots (PDPs) visualize how changes in a single feature affect the average model prediction. They are widely used in practice to inter

XFP: Quality-Targeted Adaptive Codebook Quantization with Sparse Outlier Separation for LLM Inference

HardwareDGX agent

arXiv:2605.14844v1 Announce Type: cross Abstract: We introduce XFP, a dynamic weight quantizer for LLM inference that inverts the conventional workflow: the operator specifies reconstruction quality f

14 May 2026

almost all LLM “thought” is derivative. below just sounds like hype to me

HardwareDGX agent

almost all LLM “thought” is derivative. below just sounds like hype to me 📁 Jensen Huang, CEO of NVIDIA, says GPT was never just about generating images or text but about generating thought itself. Th

British inference chip startup Fractile bags $220M to accelerate token consumption

HardwareDGX agent

U.K.-based artificial intelligence inference chip startup Fractile Ltd. said today it has closed on a 220 million Series B round of funding. The company was founded in 2022 by the Oxford University-tr

DIVER:Diving Deeper into Distilled Data via Expressive Semantic Recovery

HardwareDGX agent

arXiv:2605.12649v1 Announce Type: new Abstract: Dataset distillation aims to synthesize a compact proxy dataset that is unreadable or non-raw from the original dataset for privacy protection and highl

Elon Musk with NVIDIA CEO Jensen Huang, Apple CEO Tim Cook, President Trump, Chinese President Xi Jinping, and members of the U.S. delegatio…

HardwareDGX agent

This post appears to show Elon Musk in a photo with NVIDIA CEO Jensen Huang, Apple CEO Tim Cook, U.S. President Trump, Chinese President Xi Jinping, and members of a U.S. delegation, though the specif

EMO: Frustratingly Easy Progressive Training of Extendable MoE

HardwareDGX agent

arXiv:2605.13247v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models offer a powerful way to scale model size without increasing compute, as per-token FLOPs depend only on k active e

FlashSampling: Fast and Memory-Efficient Exact Sampling

HardwareDGX agent

arXiv:2603.15854v2 Announce Type: replace-cross Abstract: Sampling from a categorical distribution is mathematically simple, but in large-vocabulary decoding, it often triggers extra memory traffic an

Flow Augmentation and Knowledge Distillation for Lightweight Face Presentation Attack Detection

HardwareDGX agent

arXiv:2605.13108v1 Announce Type: new Abstract: Face presentation attack detection (FacePAD) remains challenging under diverse spoofing representation, including 2D print and replay, 3D mask-based spo

Geometric Autoencoder Priors for Bayesian Inversion: Learn First Observe Later

HardwareDGX agent

arXiv:2509.19929v4 Announce Type: replace-cross Abstract: Uncertainty Quantification (UQ) is paramount for inference in engineering. A common inference task is to recover full-field information of phy

How the NVIDIA Vera Rubin Platform is Solving Agentic AI’s Scale-Up Problem

HardwareDGX agent

The NVIDIA Vera Rubin Platform is a rack-scale AI supercomputer designed to power agentic AI and reasoning models at scale by eliminating bottlenecks in communication and memory movement for efficient

Made a simple utility to share your GPU profile traces. `uvx trace-tuil <local traces> -b <hf bucket name>`

HardwareDGX agent

A utility tool was created to streamline sharing GPU profile traces by uploading them to Hugging Face buckets using the command `uvx trace-tuil -b `. This simplifies the process of distributing perfor

OP4KSR: One-Step Patch-Free 4K Super-Resolution with Periodic Artifact Suppression

HardwareDGX agent

arXiv:2605.13457v1 Announce Type: new Abstract: Diffusion-based real-world image super-resolution (Real-ISR) has achieved remarkable perceptual quality; however, directly super-resolving images to 4K

Sea You in the Cloud: ‘Subnautica 2’ Early Access Dives Onto GeForce NOW

HardwareDGX agent

Dive masks on — Subnautica 2 is making a splash on GeForce NOW day-and-date with launch, so members can plunge into the title’s brand-new alien ocean from almost any device. It leads 11 new games join

Since going on @creatine_cycle podcast 1.5 months ago, I have lost 4.8 pounds of fat and gained 4.6 pounds of muscle. Swole as a service wor…

HardwareDGX agent

Since going on @creatine_cycle podcast 1.5 months ago, I have lost 4.8 pounds of fat and gained 4.6 pounds of muscle. Swole as a service works. on today's episode of Swole as a Service i brought my go

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakee…

HardwareDGX agent

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seco

13 May 2026

A profile of California Rep. Ro Khanna, who spent years cheering on the tech industry and now supports a 5% billionaire wealth tax and stricter AI regulations (Bloomberg)

HardwareDGX agent

Bloomberg: A profile of California Rep. Ro Khanna, who spent years cheering on the tech industry and now supports a 5% billionaire wealth tax and stricter AI regulations — Ro Khanna spent years cheeri

Accelerated X-Ray Analysis for Nanoscale Imaging (XANI) of Novel Materials

HardwareDGX agent

XANI (X-ray Analysis for Nanoscale Imaging) is an accelerated computing pipeline developed by NVIDIA engineers for rapid X-ray data analysis. NVIDIA's accelerated computing enables real-time experimen

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

HardwareDGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

HardwareDGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

Cerebras — Faster Tokens Please

HardwareDGX agent

Cerebras, a company specializing in AI accelerators and wafer-scale computing systems, is discussed in terms of its approaches to improving token generation speed in large language models, which is cr

ChunkFlow: Communication-Aware Chunked Prefetching for Layerwise Offloading in Distributed Diffusion Transformer Inference

HardwareDGX agent

arXiv:2605.11335v1 Announce Type: cross Abstract: Layerwise offloading reduces the GPU memory footprint of large diffusion transformer (DiT) inference by prefetching upcoming layers from host memory,

CME Group and Silicon Data to launch AI compute futures market

HardwareDGX agent

Silicon Data, the startup that provides market intelligence for artificial intelligence compute infrastructure, will provide the price indexes for a new futures market that will allow investors to hed

Efficient Remote KV Cache Reuse with GPU-native Video Codec

HardwareDGX agent

arXiv:2602.09725v3 Announce Type: replace-cross Abstract: Remote KV cache reuse fetches KV cache for identical contexts from remote storage, avoiding recomputation, accelerating LLM inference. While i

Fast MoE Inference via Predictive Prefetching and Expert Replication

HardwareDGX agent

arXiv:2605.11537v1 Announce Type: new Abstract: The Mixture of Experts (MoE) architecture has become a fundamental building block in state-of-the-art large language models (LLMs), improving domain-spe

How the 'Facebook House' in Los Altos, Mark Zuckerberg's former residence, became a hub for Chinese AI talent, who are a key part of Silicon Valley's AI boom (Viola Zhou/Rest of World)

HardwareDGX agent

Viola Zhou / Rest of World: How the “Facebook House” in Los Altos, Mark Zuckerberg's former residence, became a hub for Chinese AI talent, who are a key part of Silicon Valley's AI boom — Chinese-born

MLCommons Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces

HardwareDGX agent

arXiv:2605.11333v1 Announce Type: cross Abstract: The fast pace of artificial intelligence~(AI) innovation demands an agile methodology for observation, reproduction and optimization of distributed ma

NVIDIA, Ineffable Intelligence Team Up to Build the Future of Reinforcement Learning Infrastructure

HardwareDGX agent

Reinforcement-learning agents — AI systems that learn by trial and error — can convert computation into new knowledge. That’s the focus of a new engineering-level collaboration between NVIDIA and Inef

Recursive Superintelligence raises $650M to build self-improving AI models

HardwareDGX agent

Recursive Superintelligence Inc., a startup that hopes to develop self-improving artificial intelligence models, launched today with 650 million in funding. Alphabet Inc.’s GV fund and Greycroft led t

Red Hat and Intel spotlight scalable AI inference as enterprises move beyond the GPU gold rush

HardwareDGX agent

As companies move from testing AI to broader adoption, the biggest challenge is building scalable AI inference systems that perform without breaking the budget. The next wave of AI won’t be won on raw

Richard Socher's Recursive Superintelligence raised 650M+ from GV, Greycroft, Nvidia, AMD, and others at a 4B valuation to pursue 'recursive self-improvement' (Cade Metz/New York Times)

HardwareDGX agent

Cade Metz / New York Times: Richard Socher's Recursive Superintelligence raised 650M+ from GV, Greycroft, Nvidia, AMD, and others at a 4B valuation to pursue “recursive self-improvement” — Recursive S

The Illusion of Power Capping in LLM Decode: A Phase-Aware Energy Characterisation Across Attention Architectures

HardwareDGX agent

arXiv:2605.11999v1 Announce Type: cross Abstract: Power capping is the standard GPU energy lever in LLM serving, and it appears to work: throughput drops, power readings fall, and energy budgets are m

To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation

HardwareDGX agent

arXiv:2412.14461v4 Announce Type: replace Abstract: Unstructured text data annotation is foundational to management research. LLMs offer a cost-effective and scalable alternative to human annotation,

Transform Video Into Instantly Searchable, Actionable Intelligence with AI Agents and Skills

HardwareDGX agent

NVIDIA's video analytics AI agents analyze and process large volumes of video data through natural language tasks to provide critical insights , powered by vision language models, large language model

TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion

HardwareDGX agent

arXiv:2605.12220v1 Announce Type: new Abstract: Safe autonomous agents and mobile robots need fast real time 3D perception, especially for vulnerable road users (VRUs) such as pedestrians. We introduc

12 May 2026

AAAC: Activation-Aware Adaptive Codebooks for 4-bit LLM Weight Quantization

HardwareDGX agent

arXiv:2605.08692v1 Announce Type: cross Abstract: Post-training weight-only quantization to 4 bits is widely used to reduce the memory and compute costs of large language model inference. Existing PTQ

After studying 300 Leetcode Hards, solving every Jane Street puzzle from the Dwarkesh ads, and watching one Horace He lecture, he finally la…

HardwareDGX agent

After studying 300 Leetcode Hards, solving every Jane Street puzzle from the Dwarkesh ads, and watching one Horace He lecture, he finally landed the $400k annualized Jane Street internship. Unfortunat

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

HardwareDGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

← Previous
1…1819202122…29
Next →