AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlog
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,446 results
Hardware

Intrinsic-Noise Consolidation: A Doob-Barrier-Conditioned Diffusion Turns Analog Device Noise into a Continual-Learning Resource

DGX agent

arXiv:2607.06924v1 Announce Type: new Abstract: On analog neuromorphic hardware, intrinsic device noise is normally an accuracy tax. We ask whether it can instead consolidate memories. We cast per-syn

hardwarearxiv-cs-lg
9 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

Clustering-Embedded Model Predictive Path Integral Control: Avoiding Averaging-Induced Failure and Enabling Efficient Cluster Selection for Dynamic Obstacles

DGX agent

arXiv:2607.06499v1 Announce Type: new Abstract: With the widespread availability of parallel computing hardware, sampling-based motion planning methods such as Model Predictive Path Integral (MPPI) co

hardwarearxiv-cs-ro
8 Jul 2026
Safety

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

DGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

safetyarxiv-cs-ai
8 Jul 2026
Applications

Fine-Grained Computation Offload for Off-the-Shelf Servers in Tens of Lines

DGX agent

arXiv:2607.02630v1 Announce Type: cross Abstract: Hardware accelerators now sit on the critical path of online serving. GPUs, FPGAs, and increasingly remote services such as hardware security modules,

applicationsarxiv-cs-ai
7 Jul 2026
Hardware

It's wild how quickly Etched designed and got the chips out, all within 2 years. They went deep, hardcoding attention into silicon and getti…

DGX agent

It's wild how quickly Etched designed and got the chips out, all within 2 years. They went deep, hardcoding attention into silicon and getting very high MFU. This kind of hardware tailored made for LL

hardwareswyx--x
30 Jun 2026
Hardware

Don't Let a Few Network Failures Slow the Entire AllReduce

DGX agent

arXiv:2606.01680v1 Announce Type: cross Abstract: Network failures are among the most frequent hardware faults in large-scale GPU clusters and a leading cause of training-job interruptions. Modern col

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation

DGX agent

arXiv:2601.22813v2 Announce Type: replace Abstract: The NVFP4 lower-precision format, supported in hardware by NVIDIA Blackwell GPUs, promises to allow, for the first time, end-to-end fully-quantized

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts

DGX agent

arXiv:2605.30359v1 Announce Type: cross Abstract: Generating high-performance GPU kernels remains challenging due to the need for both correctness and hardware-aware optimization. While large language

hardwarearxiv-cs-lg
1 Jun 2026
Hardware

H100 price: down H200 price: down tokenmaxxing: dead RoI for most customers: AWOL you can’t defy gravity forever

DGX agent

Gary Marcus discusses declining prices for NVIDIA's H100 and H200 AI chips, suggesting that inflated valuations and returns on investment for AI hardware customers are becoming unsustainable as the ma

hardwaregary-marcus--x
28 May 2026
Hardware

AutoMCU: Feasibility-First MCU Neural Network Customization via LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.21560v1 Announce Type: new Abstract: Deploying neural networks on microcontroller units (MCUs) is critical for edge intelligence but remains challenging due to tight memory, storage, and co

hardwarearxiv-cs-lg
23 May 2026
Hardware

EDA Market Primer - Market Dynamics, Cadence, Synopsys, Siemens, China EDA Rise

DGX agent

EDA Market size, Share, Business Models, Drivers, Changing Customer Base, Competitive Dynamics Across Synopsys, Cadence, and Siemens, China EDA, IP, Hardware, CoT, Lock-In Economics, Disruptive Forces

hardwaresemianalysis
21 May 2026
Hardware

Instant GPU Efficiency Visibility at Fleet Scale

DGX agent

arXiv:2605.20799v1 Announce Type: cross Abstract: We present Overall FLOP Utilization (OFU), a hardware-level, precision-agnostic GPU efficiency metric for AI workloads on HPC systems, derived from tw

hardwarearxiv-cs-lg
21 May 2026
Hardware

TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload

DGX agent

arXiv:2605.20179v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have emerged as a competitive alternative to autoregressive (AR) models, offering better hardware utilization an

hardwarearxiv-cs-cl
20 May 2026
Hardware

Towards Multi-Model LLM Schedulers: Empirical Insights into Offloading and Preemption

DGX agent

arXiv:2605.19593v1 Announce Type: new Abstract: Modern deployments of Large Language Models (LLMs) increasingly require serving multiple models with diverse architectures, sizes, and specialization on

hardwarearxiv-cs-ai
20 May 2026
Hardware

LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models

DGX agent

arXiv:2605.17289v1 Announce Type: cross Abstract: Unstructured sparsity is now natively accelerated by recent GPU kernels and dataflow hardware, shifting the bottleneck from inference execution to the

hardwarearxiv-cs-ai
19 May 2026
Hardware

TrainMover: An Interruption-Resilient Runtime for ML Training

DGX agent

arXiv:2412.12636v3 Announce Type: replace-cross Abstract: Large-scale ML training jobs are frequently interrupted by hardware and software anomalies, failures, and management events. Existing solution

hardwarearxiv-cs-ai
18 May 2026
Hardware

GenAI for Energy-Efficient and Interference-Aware Compressed Sensing of GNSS Signals on a Google Edge TPU

DGX agent

arXiv:2605.14839v1 Announce Type: new Abstract: Traditional methods for classifying global navigation satellite system (GNSS) jamming signals typically involve post-processing raw or spectral data str

hardwarearxiv-cs-lg
15 May 2026
Hardware

Dooly: Configuration-Agnostic, Redundancy-Aware Profiling for LLM Inference Simulation

DGX agent

arXiv:2605.07985v1 Announce Type: cross Abstract: Selecting the optimal LLM inference configuration requires evaluation across hardware, serving engines, attention backends, and model architectures, s

hardwarearxiv-cs-ai
11 May 2026
Hardware

Nvidia up 2.6% on news that xAI is flooding the market with 220,000 secondhand GPUs.

DGX agent

Nvidia's stock rose 2.6% following news that xAI is selling 220,000 used GPUs on the secondhand market, potentially indicating a shift in AI hardware demand or xAI's operational priorities. The report

hardwaregary-marcus--x
7 May 2026
Hardware

llamacpp on Apple Silicon, once configured correctly is really rock solid! You can throw anything at it and it will answer. Impressive!

DGX agent

Llamacpp, when properly configured on Apple Silicon hardware, demonstrates robust performance and reliability for running language models. The tool can handle varied input requests effectively, making

hardwareclem-delangue--x
4 May 2026
Hardware

Most AI teams treat compute as a commodity. It's not.

DGX agent

AI compute resources should not be treated as interchangeable commodities, as different hardware configurations, providers, and architectures significantly impact training costs, inference latency, an

hardwarelambda-labs
4 May 2026
Hardware

Motivating Next-Gen Accelerators with Flexible (N:M) Activation Sparsity via Benchmarking Lightweight Post-Training Sparsification Approaches

DGX agent

arXiv:2509.22166v4 Announce Type: replace-cross Abstract: The demand for efficient large language model (LLM) inference has intensified the focus on sparsification techniques. While semi-structured (N

hardwarearxiv-cs-ai
27 Apr 2026
Hardware

We're launching two specialized TPUs for the agentic era.

DGX agent

Google announced Ironwood, its seventh-generation TPU that is twice as power efficient as the previous generation, alongside specialized hardware designed to support the emerging agentic AI era. Ironw

hardwaregoogle-ai
22 Apr 2026
Hardware

How Much Do GPU Clusters Really Cost?

DGX agent

This article from SemiAnalysis analyzes the total cost of ownership for GPU clusters, likely covering hardware expenses, infrastructure requirements, power consumption, and operational overhead beyond

hardwaresemianalysis
20 Apr 2026
Hardware

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

DGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

hardwarearxiv-cs-ro
20 Apr 2026
Hardware

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers

DGX agent

arXiv:2509.23638v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models face memory and PCIe latency bottlenecks when deployed on commodity hardware. Offloading expert weights to CPU memor

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

Fast AI Model Partition for Split Learning over Edge Networks

DGX agent

arXiv:2507.01041v4 Announce Type: replace-cross Abstract: Split learning (SL) is a distributed learning paradigm that can enable computation-intensive artificial intelligence (AI) applications by part

hardwarearxiv-cs-ai
15 Apr 2026
Hardware

GPU Acceleration of Sparse Fully Homomorphic Encrypted DNNs

DGX agent

arXiv:2604.11659v1 Announce Type: cross Abstract: Fully homomorphic encryption (FHE) has recently attracted significant attention as both a cryptographic primitive and a systems challenge. Given the l

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

Record-Remix-Replay: Hierarchical GPU Kernel Optimization using Evolutionary Search

DGX agent

arXiv:2604.11109v1 Announce Type: cross Abstract: As high-performance computing and AI workloads become increasingly dependent on GPUs, maintaining high performance across rapidly evolving hardware ge

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

LAsset: An LLM-assisted Security Asset Identification Framework for System-on-Chip (SoC) Verification

DGX agent

arXiv:2601.02624v2 Announce Type: replace-cross Abstract: The growing complexity of modern system-on-chip (SoC) and IP designs is making security assurance difficult day by day. One of the fundamental

hardwarearxiv-cs-ai
10 Apr 2026
Model Releases

QNAS: A Neural Architecture Search Framework for Accurate and Efficient Quantum Neural Networks

DGX agent

arXiv:2604.07013v1 Announce Type: cross Abstract: Designing quantum neural networks (QNNs) that are both accurate and deployable on NISQ hardware is challenging. Handcrafted ansatze must balance expre

model-releasesarxiv-cs-lg
10 Apr 2026
Hardware

CARB: A Characterization-Guided Framework for CNN Inference Cost Prediction and Deployment Screening

DGX agent

arXiv:2608.10506v1 Announce Type: cross Abstract: Accurate pre-deployment estimation of CNN inference cost--energy, latency, and peak memory--is increasingly critical as models are deployed on resourc

hardwarearxiv-cs-lg
12 Aug 2026
Hardware

EvoMem: Memory-Augmented Evolution for Code Optimization

DGX agent

arXiv:2608.10795v1 Announce Type: new Abstract: Successful mutation strategies in evolutionary code search may contain reusable knowledge that is useful beyond a single run, and in some cases may tran

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

Aero Realtime: Fully Aligned Input-Output Streams for Low-Latency Streaming Multimodal Generation

DGX agent

arXiv:2608.08469v1 Announce Type: new Abstract: Existing streaming multimodal models process observations incrementally but still follow a turn-based prefill-then-decode pattern, making them non-duple

hardwarearxiv-cs-ai
11 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 at 27+ t/s decode on Strix Halo — Vulkan + DSpark full guide

DGX agent

Been benchmarking DSv4 Flash 0731 on a Flow Z13 (Ryzen AI MAX+ 395, Radeon 8060S / gfx1151, 128GB LPDDR5X) for the past week. Figured I'd share what actually works and what doesn't — there are a lot o

model-releasesr-localllama
11 Aug 2026
Local Ai

Ego-OSCAR: Egocentric Open source Stereo CAptuRe System

DGX agent

arXiv:2608.08285v1 Announce Type: new Abstract: We present Ego-OSCAR, an open-hardware, low-cost, head-mounted stereo-inertial capture device for egocentric data collection in the wild. EgoOSCAR pairs

local-aiarxiv-cs-cv
11 Aug 2026
Hardware

🦔Nvidia announced agreements yesterday with the six biggest names in private capital, Apollo, Blackstone, BlackRock, Brookfield, Goldman Sa…

DGX agent

🦔Nvidia announced agreements yesterday with the six biggest names in private capital, Apollo, Blackstone, BlackRock, Brookfield, Goldman Sachs, and KKR, to raise over $500 billion so its own customers

hardwaregary-marcus--x
11 Aug 2026
Research

Ising Acceleration for Multi-Robot Multi-Target Planning

DGX agent

arXiv:2608.06803v1 Announce Type: cross Abstract: Ising machines are emerging as promising hardware for combinatorial optimization. With recent advances in CMOS Ising technology, they are becoming att

researcharxiv-cs-ro
10 Aug 2026
Hardware

Ultra-High Interactivity on NVIDIA GPUs? TileRT InferenceX Can TileRT software on NVIDIA GPU compete with Cerebras, Groq LPU, SambaNova? Bat…

DGX agent

Ultra-High Interactivity on NVIDIA GPUs? TileRT InferenceX Can TileRT software on NVIDIA GPU compete with Cerebras, Groq LPU, SambaNova? Batch Size 1, Disaggregated engine, High throughput prefill eng

hardwaredylan-patel--x
10 Aug 2026
Research

AFD-Ledger: Deployment Provisioning for Attention--FFN Disaggregation

DGX agent

arXiv:2608.04502v1 Announce Type: cross Abstract: Attention--Feed-Forward Network (FFN) Disaggregation (AFD) is emerging as a promising architecture for serving Mixture-of-Experts (MoE) language model

researcharxiv-cs-ai
6 Aug 2026
Model Releases

RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis

DGX agent

arXiv:2602.11506v4 Announce Type: replace-cross Abstract: The transition toward localized intelligence through Small Language Models (SLMs) has intensified the need for rigorous performance characteri

model-releasesarxiv-cs-ai
6 Aug 2026
Applications

SAMSEM -- A Generic and Scalable Approach for IC Metal Line Segmentation

DGX agent

arXiv:2603.16548v2 Announce Type: replace-cross Abstract: In light of globalized hardware supply chains, the assurance of hardware components has gained significant interest, particularly in cryptogra

applicationsarxiv-cs-cv
5 Aug 2026
Hardware

GoGoTB: Agentic RTL Verification with Specification-Grounded Coverage Closure

DGX agent

arXiv:2607.26181v1 Announce Type: new Abstract: Functional verification dominates integrated circuit (IC) front-end engineering effort, and a single missed bug that escapes to silicon can trigger a co

hardwarearxiv-cs-ai
31 Jul 2026
Hardware

Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline

DGX agent

arXiv:2607.22997v1 Announce Type: cross Abstract: Physical AI -- the integration of large vision-language-action (VLA) models with embodied agents that act in the real world -- has emerged as the next

hardwarearxiv-cs-ai
28 Jul 2026
Model Releases

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

DGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

PN-QNN: Harnessing Physical Noise as a Native Regularizer in Photonic Hybrid Quantum Neural Networks

DGX agent

arXiv:2607.20045v1 Announce Type: cross Abstract: Physical noise in near-term quantum hardware is usually treated as a nuisance to suppress. We ask whether it can instead act as a hardware-native regu

model-releasesarxiv-cs-lg
23 Jul 2026
Hardware

ReRAM-aware Model Finetuning addressing I-V Non-linearity and Retention Errors

DGX agent

arXiv:2606.17471v3 Announce Type: replace Abstract: Traditional CPU, GPU, and NPU architectures are increasingly limited by the von Neumann bottleneck. While In-Memory Computing (IMC) using ReRAM cros

hardwarearxiv-cs-lg
23 Jul 2026
Hardware

BitLogic: Training Framework for Gradient-Based FPGA-Native Neural Networks

DGX agent

arXiv:2602.07400v2 Announce Type: replace Abstract: Gradient-based LUT- and logic-gate-based neural networks (LUTNet, LogicNets, DiffLogic, PolyLUT, NeuraLUT, WARP-LUT, DWN, LILogicNet, LightLUT) repl

hardwarearxiv-cs-lg
8 Jul 2026
← Previous
1…56789…93
Next →