AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,446 results
5 Aug 2026

When Do Fewer Visual Tokens Accelerate Multimodal Inference? A Break-Even Study Across Decision Locations and Hardware

Model ReleasesDGX agent

arXiv:2608.03649v1 Announce Type: new Abstract: Fewer visual tokens do not guarantee lower end-to-end latency. We evaluate break-even with a reproducible protocol that accounts for decision overhead,

Self-Organising Digital Circuits

SafetyDGX agent

arXiv:2608.02606v1 Announce Type: new Abstract: Fault tolerance in classical computing has traditionally relied on static strategies like hardware redundancy and error-correcting codes. Biological sys

2 Aug 2026

Conclusion: r/LocalLLaMA still has brilliant open-weight research, but finding it requires wading through endless benchmark drama, non-local Discussion Points and repetitive hardware flexes.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

I let Gemma4-31b run on my laptop for like almost a day using a heavily altered pi to do a deep dive on our beloved Llama tangentially related Subreddit, and this was the conclusion. Feels pretty accu

31 Jul 2026

AgenticCANN: Automated Ascend C Operator Generation via Knowledge-Augmented Agentic Evolution

HardwareDGX agent

arXiv:2607.26661v1 Announce Type: new Abstract: Ascend C operator optimization is critical for NPU (Neural Processing Unit) inference performance but requires deep hardware expertise.While large langu

30 Jul 2026

Low-cost Embedded Breathing Rate Determination Using 802.15.4z IR-UWB Hardware for Remote Healthcare

Model ReleasesDGX agent

arXiv:2504.03772v3 Announce Type: replace-cross Abstract: Respiratory diseases account for a significant portion of global mortality. Affordable and early detection is an effective way of addressing t

Batten Down Your Packages: Mitigation Guidance for Supply Chain Compromise

Model ReleasesDGX agent

Written by: Kelli Vanderlee, Stuart Carrera For years, the cybersecurity industry's understanding of software supply chain compromise has been anchored by a few watershed events, including Russian cyb

NeoRacer: An Open, Standardized 1:12 Scale Autonomous Race Car for Benchmarking and Education

HardwareDGX agent

arXiv:2607.26855v1 Announce Type: new Abstract: Many scientific fields rely on standard benchmarks and shared platforms to improve review and reproducibility, but autonomous systems research still lac

23 Jul 2026

Self-organizing Architecture of Receptron Units: a Hardware-Aware Framework for Edge Intelligence

Local AiDGX agent

arXiv:2607.20162v1 Announce Type: new Abstract: The growing demand for intelligent processing at the edge of IoT networks is constrained by the severe computational and memory limitations of microcont

Minimize idle accelerators: Native RL job interleaving with co-operative time-slicing in llm-d

SafetyDGX agent

The math behind reinforcement learning (RL) post-training for large language models (LLMs) is notoriously unforgiving. As frontier AI labs push the boundaries of reasoning and coding models using RL p

16 Jul 2026

Industrial Dexterity Benchmark: A Hardware-Software Benchmarking Platform for Industrial Dexterous Manipulation

Model ReleasesDGX agent

arXiv:2607.14021v1 Announce Type: new Abstract: Dexterous manipulation remains a critical bottleneck in industrial automation; tasks such as cable routing, connector insertion, and precision assembly

RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation

Model ReleasesDGX agent

arXiv:2607.13897v1 Announce Type: new Abstract: The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fun

9 Jul 2026

Smart Scissor: Coupling Spatial Redundancy Reduction and CNN Compression for Embedded Hardware

Local AiDGX agent

arXiv:2607.06915v1 Announce Type: new Abstract: Scaling down the resolution of input images can greatly reduce the computational overhead of convolutional neural networks (CNNs), which is promising fo

Intrinsic-Noise Consolidation: A Doob-Barrier-Conditioned Diffusion Turns Analog Device Noise into a Continual-Learning Resource

HardwareDGX agent

arXiv:2607.06924v1 Announce Type: new Abstract: On analog neuromorphic hardware, intrinsic device noise is normally an accuracy tax. We ask whether it can instead consolidate memories. We cast per-syn

7 Jul 2026

A Spiking Sequence Generator for Polar Trajectories on Neuromorphic Hardware

ResearchDGX agent

arXiv:2607.02753v1 Announce Type: cross Abstract: Neuromorphic controllers for size, weight, and power-constrained systems require neural architectures that are both energy-efficient and interpretable

Fine-Grained Computation Offload for Off-the-Shelf Servers in Tens of Lines

ApplicationsDGX agent

arXiv:2607.02630v1 Announce Type: cross Abstract: Hardware accelerators now sit on the critical path of online serving. GPUs, FPGAs, and increasingly remote services such as hardware security modules,

29 Jun 2026

OpenAI is teasing new hardware… for Codex

IndustryDGX agent

OpenAI is releasing some sort of device related to its AI-powered coding tool, Codex, on July 15th. In a video posted to X on Monday, OpenAI shows a square-shaped device with several buttons, alongsid

23 Jun 2026

Verifiable, private AI: Google Cloud expands Confidential Computing frontiers

Model ReleasesDGX agent

Protecting sensitive data used with AI is a critical part of our commitment to providing advanced and secure cloud infrastructure. Confidential Computing cryptographically protects data in use in hard

WiSP: A Working-Set View of Mixture-of-Experts Serving on Extremely Low-Resource Hardware

Local AiDGX agent

arXiv:2606.21868v1 Announce Type: new Abstract: Modern Mixture-of-Experts (MoE) models place most of their parameters in expert layers, yet only a small fraction of those experts are used for any toke

27 May 2026

Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU

HardwareDGX agent

arXiv:2605.26118v1 Announce Type: cross Abstract: Porting deep learning algorithms to new hardware accelerators requires developers to repeatedly apply the same low-level optimizations -- quantization

Morphling: Fast, Fused, and Flexible GNN Training at Scale

HardwareDGX agent

arXiv:2512.01678v5 Announce Type: replace Abstract: Graph Neural Networks (GNNs) present a fundamental hardware challenge by fusing irregular, memory-bound graph traversals with regular, compute-inten

12 May 2026

Efficient Neural Architectures for Real-Time ECG Interpretation on Limited Hardware

Model ReleasesDGX agent

arXiv:2605.09848v1 Announce Type: new Abstract: Electrocardiogram (ECG) interpretation is essential for diagnosing a wide range of cardiac abnormalities. While deep learning has shown strong potential

11 May 2026

Frozen Backpropagation: Relaxing Weight Symmetry in Deep Spiking Neural Networks

HardwareDGX agent

arXiv:2505.13741v2 Announce Type: replace Abstract: Direct training of Spiking Neural Networks (SNNs) on neuromorphic hardware can greatly reduce energy costs compared to GPU-based training. However,

CktFormalizer: Autoformalization of Natural Language into Circuit Representations

HardwareDGX agent

arXiv:2605.07782v1 Announce Type: new Abstract: LLMs can generate hardware descriptions from natural language specifications, but the resulting Verilog often contains width mismatches, combinational l

21 Apr 2026

John Ternus’ first big problem is AI

IndustryDGX agent

Less than a year ago, Apple made headlines for a lack of AI announcements at its annual WWDC event. Ten months later, the company has announced that hardware executive John Ternus will succeed longtim

Tim Cook reshaped Apple in his own image and exits as CEO on his own terms, seemingly picking the right successor in John Ternus, much like Jobs did with Cook (John Gruber/Daring Fireball)

IndustryDGX agent

John Gruber / Daring Fireball: Tim Cook reshaped Apple in his own image and exits as CEO on his own terms, seemingly picking the right successor in John Ternus, much like Jobs did with Cook — It's a p

17 Apr 2026

Animating Petascale Time-varying Data on Commodity Hardware with LLM-assisted Scripting

ApplicationsDGX agent

arXiv:2603.07053v2 Announce Type: replace Abstract: Scientists face significant visualization challenges as time-varying datasets grow in speed and volume, often requiring specialized infrastructure a

ChatSVA: Bridging SVA Generation for Hardware Verification via Task-Specific LLMs

Model ReleasesDGX agent

arXiv:2604.02811v2 Announce Type: replace-cross Abstract: Functional verification consumes over 50% of the IC development lifecycle, where SystemVerilog Assertions (SVAs) are indispensable for formal

10 Jun 2026

LLM-Guided Neural Architecture Search for Robust Co-Design of Physical Neural Networks

ResearchDGX agent

arXiv:2606.10294v1 Announce Type: cross Abstract: Deploying neural networks on unconventional hardware demands architectures that co-optimize task accuracy and platform-specific constraints such as en

9 Jun 2026

TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks

HardwareDGX agent

arXiv:2510.16028v4 Announce Type: replace-cross Abstract: Neural networks increasingly run on hardware outside the user's control (cloud GPUs, inference marketplaces). Yet ML-as-a-Service reveals litt

DeepSeekV4 1.6T Day 0 to Day 43 Performance Over Time - Huawei, GB300 NVL72, MI355X, B200

HardwareDGX agent

This article from SemiAnalysis tracks the performance evolution of DeepSeek V4, a 1.6 trillion parameter model, over its first 43 days of training on Huawei and NVIDIA hardware including GB300 NVL72,

26 May 2026

Physical Analogue Kolmogorov-Arnold Networks based on Reconfigurable Nonlinear-Processing Units

HardwareDGX agent

arXiv:2602.07518v2 Announce Type: replace-cross Abstract: Kolmogorov-Arnold Networks (KANs) shift neural computation from linear layers to learnable nonlinear edge functions, but implementing these no

28 Apr 2026

Deployment-Aligned Low-Precision Neural Architecture Search for Spaceborne Edge AI

Local AiDGX agent

arXiv:2604.24492v1 Announce Type: cross Abstract: Designing deep networks that meet strict latency and accuracy constraints on edge accelerators increasingly relies on hardware-aware optimization, inc

20 Apr 2026

John Ternus will replace Tim Cook as Apple CEO

IndustryDGX agent

Apple announced that Tim Cook will step down as CEO on September 1, 2026, ending his 15-year tenure, with John Ternus, the senior vice president of hardware engineering, succeeding him as chief execut

11 Aug 2026

PQC in Plaintext: Google Cloud’s post-quantum cryptography roadmap

SafetyDGX agent

Securing infrastructure and services against a future cryptographically-relevant quantum computer has been a goal for Google for a decade, and we’ve dedicated ourselves to help developers by advancing

2 Jun 2026

FreqLite: A Lightweight Frequency-Decomposed Linear Model with Adaptive Reversible Normalization for Robust Long-Term Time-Series Forecasting

HardwareDGX agent

arXiv:2606.01339v1 Announce Type: cross Abstract: Long-term time-series forecasting needs models that are accurate yet efficient enough for commodity hardware. Lightweight linear forecasters are remar

Don't Let a Few Network Failures Slow the Entire AllReduce

HardwareDGX agent

arXiv:2606.01680v1 Announce Type: cross Abstract: Network failures are among the most frequent hardware faults in large-scale GPU clusters and a leading cause of training-job interruptions. Modern col

Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation

HardwareDGX agent

arXiv:2601.22813v2 Announce Type: replace Abstract: The NVFP4 lower-precision format, supported in hardware by NVIDIA Blackwell GPUs, promises to allow, for the first time, end-to-end fully-quantized

3 May 2026

torch-nvenc-compress: GPU NVENC silicon as a PCIe bandwidth multiplier — PCA + pure-ctypes Video Codec SDK wrapper. Parallel-path overlap measured at 67% of theoretical max on a real GEMM + encode workload. [P]

HardwareDGX agent

torch-nvenc-compress is a Python library that leverages GPU NVENC (NVIDIA's hardware video encoding) to optimize PCIe bandwidth utilization by compressing data during transfer. The project implements

14 Apr 2026

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

HardwareDGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

9 Apr 2026

A developer’s guide to architecting reliable GPU infrastructure at scale

Model ReleasesDGX agent

Editor’s note: This blog post outlines Google Cloud’s GPU AI/ML infrastructure reliability strategy, and will be updated with links to new community articles as they appear. As we enter the era of mul

12 Aug 2026

What do you guys do for GPU Kernels?

HardwareDGX agent

I'm trying to figure out GPU Kernel optimization on older hardware like SM80(ampere) . Is there tools you guys use? Or frameworks? Im waiting for this framework https://www.reddit.com/r/LocalLLaMA/com

7 Aug 2026

Parallel-in-Time Nonlinear Optimal Control via GPU-native Sequential Convex Programming

HardwareDGX agent

arXiv:2603.10711v3 Announce Type: replace Abstract: Real-time solution of nonlinear optimal control problems remains challenging on embedded robotic hardware, where conventional solvers often rely on

6 Aug 2026

Understanding Fault Tolerance of Adversarially Robust Pruned Models

ResearchDGX agent

arXiv:2608.04173v1 Announce Type: new Abstract: Deep neural networks (DNNs) deployed on resource-constrained neuromorphic hardware face three concurrent challenges: the need for model compression thro

4 Aug 2026

Constrained Co-Design for Photonic Bayesian Neural Networks

SafetyDGX agent

arXiv:2608.02229v1 Announce Type: new Abstract: Classical neural networks frequently produce overconfident predictions on ambiguous or out-of-distribution (OOD) data, a liability that grows with each

Why are Gamers so incredibly hostile to AI? Is it just a tiny vocal minority that spreads such toxic vitriol online?

Model ReleasesDGX agent

It's more accurate to say that many highly engaged online gamers are hostile to AI, not that 'gamers' as a whole are. Gaming is a huge community with hundreds of millions of people, and opinions vary

3 Aug 2026

OsteoCAD: A Human-in-the-Loop Cloud-Edge Framework for Bone Tumor Segmentation

HardwareDGX agent

arXiv:2607.29266v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Deep Learning (DL) have notably advanced medical image analysis, yet many health- care organizations struggle to adop

25 Jul 2026

AMD moves beyond challenger status in the race for AI platform leadership

HardwareDGX agent

AI hardware competition entered a sharper phase this week as Advanced Micro Devices Inc. used its flagship AI event to argue it isn’t merely chasing Nvidia Corp. — it intends to lead the market outrig

21 Jul 2026

Devin Outposts are now available on NVIDIA Brev. Write and profile kernels, run experiments, serve and fine-tune OSS models, and test agains…

HardwareDGX agent

Devin Outposts are now available on NVIDIA Brev. Write and profile kernels, run experiments, serve and fine-tune OSS models, and test against real hardware. Try it out: https://brev.nvidia.com/launcha

Run Devin Outposts on @nvidia Brev to give Devin access to GPU infrastructure for debugging training runs, inspecting GPU state, reproducing…

HardwareDGX agent

Run Devin Outposts on @nvidia Brev to give Devin access to GPU infrastructure for debugging training runs, inspecting GPU state, reproducing failures, and validating fixes on the hardware where the wo

15 Jul 2026

Building Faster Cryptography with Carryless Multiplication in NVIDIA CUDA 13.3

HardwareDGX agent

CUDA 13.3 introduces **clmad**, a hardware‑accelerated carryless multiply‑accumulate instruction available on all NVIDIA Ampere and newer GPUs (SM 80+), filling a long‑standing gap in GPU‑native binar

Cadence extends its AI agents beyond chips with AuraStack for circuit boards and packaging

HardwareDGX agent

Integrated circuit and electronic hardware design company Cadence Design Systems Inc. today announced a new artificial intelligence agent that assists with packaging and system design, the step after

VQCSim: When Does Compile-Once Statevector Simulation Beat Generic Quantum Frameworks?

HardwareDGX agent

arXiv:2607.11985v1 Announce Type: cross Abstract: Hybrid quantum-classical machine learning workflows repeatedly evaluate many small parametrized circuits during training and model exploration. In thi

8 Jul 2026

Clustering-Embedded Model Predictive Path Integral Control: Avoiding Averaging-Induced Failure and Enabling Efficient Cluster Selection for Dynamic Obstacles

HardwareDGX agent

arXiv:2607.06499v1 Announce Type: new Abstract: With the widespread availability of parallel computing hardware, sampling-based motion planning methods such as Model Predictive Path Integral (MPPI) co

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

SafetyDGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

30 Jun 2026

It's wild how quickly Etched designed and got the chips out, all within 2 years. They went deep, hardcoding attention into silicon and getti…

HardwareDGX agent

It's wild how quickly Etched designed and got the chips out, all within 2 years. They went deep, hardcoding attention into silicon and getting very high MFU. This kind of hardware tailored made for LL

1 Jun 2026

Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts

HardwareDGX agent

arXiv:2605.30359v1 Announce Type: cross Abstract: Generating high-performance GPU kernels remains challenging due to the need for both correctness and hardware-aware optimization. While large language

28 May 2026

H100 price: down H200 price: down tokenmaxxing: dead RoI for most customers: AWOL you can’t defy gravity forever

HardwareDGX agent

Gary Marcus discusses declining prices for NVIDIA's H100 and H200 AI chips, suggesting that inflated valuations and returns on investment for AI hardware customers are becoming unsustainable as the ma

23 May 2026

AutoMCU: Feasibility-First MCU Neural Network Customization via LLM-based Multi-Agent Systems

HardwareDGX agent

arXiv:2605.21560v1 Announce Type: new Abstract: Deploying neural networks on microcontroller units (MCUs) is critical for edge intelligence but remains challenging due to tight memory, storage, and co

21 May 2026

EDA Market Primer - Market Dynamics, Cadence, Synopsys, Siemens, China EDA Rise

HardwareDGX agent

EDA Market size, Share, Business Models, Drivers, Changing Customer Base, Competitive Dynamics Across Synopsys, Cadence, and Siemens, China EDA, IP, Hardware, CoT, Lock-In Economics, Disruptive Forces

Instant GPU Efficiency Visibility at Fleet Scale

HardwareDGX agent

arXiv:2605.20799v1 Announce Type: cross Abstract: We present Overall FLOP Utilization (OFU), a hardware-level, precision-agnostic GPU efficiency metric for AI workloads on HPC systems, derived from tw

← Previous
1…34567…75
Next →