AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
19 May 2026

TierCheck: Tiered Checkpointing for Fault Tolerance in Large Language Model Training

HardwareDGX agent

arXiv:2605.17821v1 Announce Type: cross Abstract: Large Language Model (LLM) training is frequently interrupted by a heterogeneous spectrum of failures, from common GPU crashes to catastrophic cluster

Token-Space Mask Prediction for Efficient Vision Transformer Segmentation

HardwareDGX agent

arXiv:2605.18177v1 Announce Type: new Abstract: Query-based Vision Transformer segmentation models typically reconstruct dense spatial feature maps to predict masks, inheriting design patterns from co

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks

HardwareDGX agent

arXiv:2605.17170v1 Announce Type: new Abstract: Agentic workloads have emerged as a major workload for LLM inference. They differ significantly from chat-only workloads, requiring long-context process

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Truthful Calibration Errors for Multi-Class Prediction

HardwareDGX agent

arXiv:2510.06388v2 Announce Type: replace Abstract: Calibrated predictions are useful because their numerical values can be interpreted as probabilities. Calibration errors are therefore widely used t

VeriCache: Turning Lossy KV Cache into Lossless LLM Inference

HardwareDGX agent

arXiv:2605.17613v1 Announce Type: cross Abstract: The large size of the KV cache has become a major bottleneck for serving LLMs with increasing context lengths. In response, many KV cache compression

Vultr Announces Milan, Italy, as 33rd Cloud Data Center Region

HardwareDGX agent

Vultr expanded its global cloud infrastructure by opening a new data center region in Milan, Italy, marking the company's 33rd cloud data center location worldwide. This expansion provides European cu

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human geno…

HardwareDGX agent

We are releasing Carbon: a crazy fast DNA model Carbon is 275x faster than the next best model. So fast you can process the whole human genome on a single GPU in <2 days. Here are the tricks we used:

18 May 2026

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

HardwareDGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

A Unified Non-Parametric and Interpretable Point Cloud Analysis via t-FCW Graph Representation

HardwareDGX agent

arXiv:2605.15475v1 Announce Type: new Abstract: We introduce an empowered transposed Fully Connected Weighted (t-FCW) graph representation to embed point clouds into a metric space. While original t-F

AgriMind: An Ensemble Deep Learning Framework for Multi-Class Plant Disease Classification

HardwareDGX agent

arXiv:2605.16076v1 Announce Type: cross Abstract: Plant disease detection is still largely manual in Bangladesh, where extension workers eyeball leaf samples across millions of smallholdings. We built

b9208

Local AiDGX agent

B9208 is a build release of llama.cpp, an open-source C/C++ project that enables efficient large language model inference on diverse hardware platforms. The llama.cpp project focuses on optimized LLM

b9221

Local AiDGX agent

b9221 is an intermediate build release from the llama.cpp project, which is a C/C++ implementation enabling efficient LLM inference on consumer hardware. The release includes platform-specific binarie

Bridging Silicon and the Hippocampus: Algebro-Deterministic Memory 'VaCoAl' as a Substrate for Vector-HaSH and TEM

HardwareDGX agent

arXiv:2605.15652v1 Announce Type: cross Abstract: Vector-HaSH and the Tolman-Eichenbaum Machine (TEM) propose that the hippocampal-entorhinal circuit factorizes content from a prestructured grid-cell

Decart, which offers real-time generative video and GPU optimization tech, raised 300M at a ~4B valuation, up from 3.1B after raising 153M in August 2025 (Robbie Whelan/Wall Street Journal)

HardwareDGX agent

Robbie Whelan / Wall Street Journal: Decart, which offers real-time generative video and GPU optimization tech, raised 300M at a ~4B valuation, up from 3.1B after raising 153M in August 2025 — Decart'

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization

HardwareDGX agent

arXiv:2605.15824v1 Announce Type: new Abstract: Human-centric video customization, particularly at the garment level, has shown significant commercial value. However, existing approaches cannot suppor

Fine-Tuning NVIDIA Cosmos Predict 2.5 with LoRA/DoRA for Robot Video Generation

HardwareDGX agent

This guide demonstrates how to fine-tune NVIDIA's Cosmos Predict 2.5 video generation model using parameter-efficient techniques like LoRA (Low-Rank Adaptation) and DoRA (Mixture of Experts-based adap

🧬 flux-genotype: A self-evolving AI kernel that runs on CPU with Ollama — mutates its own architecture

Local AiDGX agent

Flux-genotype is a self-evolving AI kernel designed to run on CPU hardware using Ollama, featuring the capability to mutate and adapt its own architecture dynamically. This project demonstrates an app

frax: Fast Robot Kinematics and Dynamics in JAX

HardwareDGX agent

arXiv:2604.04310v2 Announce Type: replace Abstract: In robot control, planning, and learning, there is a need for rigid-body dynamics libraries that are highly performant, easy to use, and compatible

How the AI industry's intense pressure to keep up and its financial rewards are creating 'walls of resentment' between Silicon Valley workers and their spouses (Alessandra Ram/Wired)

HardwareDGX agent

Alessandra Ram / Wired: How the AI industry's intense pressure to keep up and its financial rewards are creating “walls of resentment” between Silicon Valley workers and their spouses — Are you marrie

NVIDIA CEO Jensen Huang at Dell Technologies World: “Demand Is Going Parabolic, Utterly Parabolic”

HardwareDGX agent

Agentic AI inference at one-tenth the cost per token with NVIDIA Vera Rubin NVL72. Agent sandboxes run 50% faster on NVIDIA Vera than traditional CPUs — while enterprise data queries are up to 3x fast

Pwn2Own Berlin 2026: participants earned a total of $1,298,250 for 47 vulnerabilities, with successful exploits of AI products like Codex, Cursor, and LM Studio (Eduard Kovacs/SecurityWeek)

HardwareDGX agent

Eduard Kovacs / SecurityWeek: Pwn2Own Berlin 2026: participants earned a total of $1,298,250 for 47 vulnerabilities, with successful exploits of AI products like Codex, Cursor, and LM Studio — Partici

Together with SpaceXAI, we’re training a significantly larger model from scratch, using 10x more total compute. With Colossus 2’s million H1…

HardwareDGX agent

Together with SpaceXAI, we’re training a significantly larger model from scratch, using 10x more total compute. With Colossus 2’s million H100-equivalents and our combined data and training techniques

Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs

HardwareDGX agent

The first NVIDIA Vera CPUs arrived at three of the world's leading AI labs on Friday — Anthropic in San Francisco, OpenAI in Mission Bay, SpaceXAI in Palo Alto — followed by a delivery to Oracle Cloud

16 May 2026

[AINews] Cerebras' $60B IPO: Slowly, then All at Once

ToolsDGX agent

Cerebras Systems filed for a $60 billion IPO, marking a significant milestone for the AI hardware startup that specializes in large-scale AI chip systems. The article likely discusses Cerebras' journe

Reduce your GPU power limit

HardwareDGX agent

Setting a GPU power limit using nvidia-smi reduces heat output by approximately 20% with only a 5-8% inference speed loss. Undervolting the GPU can reduce power consumption by 5-15% with zero performa

SF vibes are frenetic over the huge divide in outcomes and career uncertainty for software engineers; over 5 years ~10K people in AI attained retirement wealth (Deedy/@deedydas)

HardwareDGX agent

Deedy / @deedydas: SF vibes are frenetic over the huge divide in outcomes and career uncertainty for software engineers; over 5 years ~10K people in AI attained retirement wealth — The vibes in SF fee

Shenzhen-listed RoboTechnik, which claims to be the largest silicon photonics tool maker and whose stock is up 340% over the past year, files for a HK listing (Zinnia Lee/Forbes)

HardwareDGX agent

Zinnia Lee / Forbes: Shenzhen-listed RoboTechnik, which claims to be the largest silicon photonics tool maker and whose stock is up 340% over the past year, files for a HK listing — RoboTechnik Intell

15 May 2026

A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 (Rebecca Bellan/TechCrunch)

HardwareDGX agent

Rebecca Bellan / TechCrunch: A profile of AI video generation startup Runway, which is training models directly on observational data, is now valued at 5.3B, and added 40M in ARR in Q2 — Every major A

BEAM: Binary Expert Activation Masking for Dynamic Routing in MoE

HardwareDGX agent

arXiv:2605.14438v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enhance the efficiency of large language models by activating only a subset of experts per token. However, standa

DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration

HardwareDGX agent

arXiv:2605.14526v1 Announce Type: cross Abstract: Differentiable simulation of soft bodies is a foundation for system identification, trajectory optimization, and Real2Sim transfer. Yet, existing meth

EMA: Efficient Model Adaptation for Learning-based Systems

HardwareDGX agent

arXiv:2605.13942v1 Announce Type: new Abstract: Machine learning (ML) is increasingly applied to optimize system performance in tasks such as resource management and network simulation. Unlike traditi

EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization

HardwareDGX agent

arXiv:2605.14249v1 Announce Type: new Abstract: We present EnergyLens, an end-to-end framework for energy-aware large language model (LLM) inference optimization. As LLMs scale, predicting and reducin

GridCare raises $64M to speed up AI data center projects

HardwareDGX agent

GridCare Inc., a startup that helps data center operators more quickly connect their facilities to the electrical grid, has raised 64 million in funding. Early Nvidia Corp. backer Sutter Hill Ventures

Neural Field Thermal Tomography: A Differentiable Physics Framework for Non-Destructive Evaluation

HardwareDGX agent

arXiv:2603.11045v2 Announce Type: replace-cross Abstract: Inverse problems for stiff parabolic partial differential equations (PDEs), such as the inverse heat conduction problem (IHCP), are severely i

Nvidia's future in China remains unclear after the Trump-Xi Summit; Trump says China 'chose not to' buy Nvidia chips as 'they want to try to develop their own' (Meaghan Tobin/New York Times)

HardwareDGX agent

Meaghan Tobin / New York Times: Nvidia's future in China remains unclear after the Trump-Xi Summit; Trump says China “chose not to” buy Nvidia chips as “they want to try to develop their own” — The st

Parallelizing Counterfactual Regret Minimization

HardwareDGX agent

arXiv:2605.14277v1 Announce Type: new Abstract: Parallelization has played an instrumental role in the field of artificial intelligence (AI), drastically reducing the time taken to train and evaluate

QOuLiPo: What a quantum computer sees when it reads a book

Model ReleasesDGX agent

arXiv:2605.14188v1 Announce Type: cross Abstract: What does a book look like to a quantum computer? This paper takes eight classical works of the Renaissance and its late-antique inheritance -- from A

Semiconductor stocks fell globally on Friday after the Trump-Xi summit concluded without major chip deals; Nvidia closed down 4.42% and AMD closed down 5.69% (Mauro Orru/Wall Street Journal)

HardwareDGX agent

Mauro Orru / Wall Street Journal: Semiconductor stocks fell globally on Friday after the Trump-Xi summit concluded without major chip deals; Nvidia closed down 4.42% and AMD closed down 5.69% — Beijin

Seoul-based WIRobotics, which develops wearable and humanoid robots and is collaborating with Nvidia and AWS, raised a ~$68M Series B led by JB Investment (Lee Jaewoon/The Elec)

HardwareDGX agent

Lee Jaewoon / The Elec: Seoul-based WIRobotics, which develops wearable and humanoid robots and is collaborating with Nvidia and AWS, raised a ~$68M Series B led by JB Investment — Company to accelera

SurgicalMamba: Dual-Path SSD with State Regramming for Online Surgical Phase Recognition

HardwareDGX agent

arXiv:2605.14889v1 Announce Type: cross Abstract: Online surgical phase recognition (SPR) underpins context-aware operating-room systems and requires committing to a prediction at every frame from pas

Synthetic Sociality: How Generative Models Privatize the Social Fabric

HardwareDGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

Towards Resource-Efficient LLMs: End-to-End Energy Accounting of Distillation Pipelines

HardwareDGX agent

arXiv:2605.13981v1 Announce Type: cross Abstract: The rise in deployment of large language models has driven a surge in GPU demand and datacenter scaling, raising concerns about electricity use, grid

Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles

HardwareDGX agent

arXiv:2605.14578v1 Announce Type: new Abstract: Partial Dependence Plots (PDPs) visualize how changes in a single feature affect the average model prediction. They are widely used in practice to inter

14 May 2026

almost all LLM “thought” is derivative. below just sounds like hype to me

HardwareDGX agent

almost all LLM “thought” is derivative. below just sounds like hype to me 📁 Jensen Huang, CEO of NVIDIA, says GPT was never just about generating images or text but about generating thought itself. Th

British inference chip startup Fractile bags $220M to accelerate token consumption

HardwareDGX agent

U.K.-based artificial intelligence inference chip startup Fractile Ltd. said today it has closed on a 220 million Series B round of funding. The company was founded in 2022 by the Oxford University-tr

DIVER:Diving Deeper into Distilled Data via Expressive Semantic Recovery

HardwareDGX agent

arXiv:2605.12649v1 Announce Type: new Abstract: Dataset distillation aims to synthesize a compact proxy dataset that is unreadable or non-raw from the original dataset for privacy protection and highl

Elon Musk with NVIDIA CEO Jensen Huang, Apple CEO Tim Cook, President Trump, Chinese President Xi Jinping, and members of the U.S. delegatio…

HardwareDGX agent

This post appears to show Elon Musk in a photo with NVIDIA CEO Jensen Huang, Apple CEO Tim Cook, U.S. President Trump, Chinese President Xi Jinping, and members of a U.S. delegation, though the specif

EMO: Frustratingly Easy Progressive Training of Extendable MoE

HardwareDGX agent

arXiv:2605.13247v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models offer a powerful way to scale model size without increasing compute, as per-token FLOPs depend only on k active e

FlashSampling: Fast and Memory-Efficient Exact Sampling

HardwareDGX agent

arXiv:2603.15854v2 Announce Type: replace-cross Abstract: Sampling from a categorical distribution is mathematically simple, but in large-vocabulary decoding, it often triggers extra memory traffic an

Flow Augmentation and Knowledge Distillation for Lightweight Face Presentation Attack Detection

HardwareDGX agent

arXiv:2605.13108v1 Announce Type: new Abstract: Face presentation attack detection (FacePAD) remains challenging under diverse spoofing representation, including 2D print and replay, 3D mask-based spo

How the NVIDIA Vera Rubin Platform is Solving Agentic AI’s Scale-Up Problem

HardwareDGX agent

The NVIDIA Vera Rubin Platform is a rack-scale AI supercomputer designed to power agentic AI and reasoning models at scale by eliminating bottlenecks in communication and memory movement for efficient

Made a simple utility to share your GPU profile traces. `uvx trace-tuil <local traces> -b <hf bucket name>`

HardwareDGX agent

A utility tool was created to streamline sharing GPU profile traces by uploading them to Hugging Face buckets using the command `uvx trace-tuil -b `. This simplifies the process of distributing perfor

OP4KSR: One-Step Patch-Free 4K Super-Resolution with Periodic Artifact Suppression

HardwareDGX agent

arXiv:2605.13457v1 Announce Type: new Abstract: Diffusion-based real-world image super-resolution (Real-ISR) has achieved remarkable perceptual quality; however, directly super-resolving images to 4K

Sea You in the Cloud: ‘Subnautica 2’ Early Access Dives Onto GeForce NOW

HardwareDGX agent

Dive masks on — Subnautica 2 is making a splash on GeForce NOW day-and-date with launch, so members can plunge into the title’s brand-new alien ocean from almost any device. It leads 11 new games join

Since going on @creatine_cycle podcast 1.5 months ago, I have lost 4.8 pounds of fat and gained 4.6 pounds of muscle. Swole as a service wor…

HardwareDGX agent

Since going on @creatine_cycle podcast 1.5 months ago, I have lost 4.8 pounds of fat and gained 4.6 pounds of muscle. Swole as a service works. on today's episode of Swole as a Service i brought my go

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakee…

HardwareDGX agent

Together AI STT models now hold the top two spots for transcription speed on the @ArtificialAnlys Speech to Text leaderboard. NVIDIA Parakeet TDT 0.6B V3 on Together AI ranks #1, transcribing 303 seco

13 May 2026

A profile of California Rep. Ro Khanna, who spent years cheering on the tech industry and now supports a 5% billionaire wealth tax and stricter AI regulations (Bloomberg)

HardwareDGX agent

Bloomberg: A profile of California Rep. Ro Khanna, who spent years cheering on the tech industry and now supports a 5% billionaire wealth tax and stricter AI regulations — Ro Khanna spent years cheeri

Accelerated X-Ray Analysis for Nanoscale Imaging (XANI) of Novel Materials

HardwareDGX agent

XANI (X-ray Analysis for Nanoscale Imaging) is an accelerated computing pipeline developed by NVIDIA engineers for rapid X-ray data analysis. NVIDIA's accelerated computing enables real-time experimen

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

HardwareDGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

HardwareDGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

← Previous
1…2627282930…75
Next →