AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
Hardware

How Together AI built the world’s fastest speech-to-text stack

DGX agent

Together AI developed an optimized speech-to-text system focused on achieving the fastest processing speeds through technical innovations in their inference stack and model optimization. The approach

hardwaretogether-ai-blog
29 May 2026
Local Ai
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Q-ANCHOR: Federated Quantum Learning with ZNE-guided Correction

DGX agent

arXiv:2605.30075v1 Announce Type: new Abstract: Quantum Federated Learning (QFL) offers a promising framework to train quantum models across distributed clients while keeping data strictly local. Due

local-aiarxiv-cs-lg
29 May 2026
Agents

SchGen: PCB Schematic Generation with Semantic-Grounded Code Representations

DGX agent

arXiv:2605.30345v1 Announce Type: new Abstract: Printed circuit board (PCB) schematic design defines nearly all electronic hardware, but it remains manual and expertise-intensive. While generative AI

agentsarxiv-cs-ai
29 May 2026
Hardware

What is the best used or refurbished laptop with GPU for open source Imege generation?

DGX agent

This Reddit discussion in the StableDiffusion community addresses recommendations for affordable, used or refurbished laptops equipped with GPUs suitable for running open-source image generation model

hardwarer-stablediffusion
29 May 2026
Model Releases

AssertLLM2: A Comprehensive LLM Benchmark for Assertion Generation from Design Specifications

DGX agent

arXiv:2605.27472v1 Announce Type: cross Abstract: Assertion-based verification (ABV) is a cornerstone of modern hardware design, yet manually translating design intent into formal SystemVerilog Assert

model-releasesarxiv-cs-ai
28 May 2026
Hardware

Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩 https://huggingface.co/nvidia/GLM-5.1-NVFP4

DGX agent

NVIDIA has released GLM-5.1-NVFP4, a quantized version of the GLM-5.1 model, now available on Hugging Face. The model appears to use NVFP4 (NVIDIA's floating-point 4-bit) quantization format, designed

hardwareclem-delangue--x
28 May 2026
Hardware

Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers

DGX agent

arXiv:2605.25346v1 Announce Type: cross Abstract: Neural network (NN) dynamics models and control policies achieve strong performance in robotics, but providing sound guarantees under uncertainty rema

hardwarearxiv-cs-ai
26 May 2026
Hardware

SA-Kura: An Energy-Efficient Systolic Array Accelerator for Locally-Coupled Kuramoto Drift in Diffusion Sampling

DGX agent

arXiv:2605.24016v1 Announce Type: cross Abstract: Diffusion inference remains costly for edge deployment, yet existing accelerators focus almost exclusively on score networks because standard drift is

hardwarearxiv-cs-ai
26 May 2026
Hardware

The Cognitive Kardashev Scale: Quantifying the Material Envelope of Civilisational Computation

DGX agent

arXiv:2605.22840v1 Announce Type: cross Abstract: How much thinking can a civilisation do? Kardashev's (1964) typology ranks civilisations by total power: planetary (Type I, ~10^16 W), stellar (Type I

hardwarearxiv-cs-ai
25 May 2026
Research

Bilateral Teleoperation with Compliant 6-DOF Pose-and-Force Sensing

DGX agent

arXiv:2605.19255v1 Announce Type: new Abstract: Existing bilateral teleoperation platforms still rely on costly rigid six-axis force/torque sensors, tightly coupled leader-follower hardware, and kiloh

researcharxiv-cs-ro
20 May 2026
Hardware

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs

DGX agent

arXiv:2605.19269v1 Announce Type: new Abstract: Transformer training systems are built around dense linear algebra, yet a nontrivial fraction of end-to-end time is spent on surrounding memory-bound op

hardwarearxiv-cs-lg
20 May 2026
Research

Receptogenesis in a Vascularized Robotic Embodiment

DGX agent

arXiv:2603.09473v2 Announce Type: replace Abstract: Equipping robotic systems with the capacity to generate extit{ex novo} hardware during operation extends control of physical adaptability. Unlike mo

researcharxiv-cs-ro
20 May 2026
Hardware

Soft Learning

DGX agent

arXiv:2605.18889v1 Announce Type: cross Abstract: Modern machine learning forces practitioners to choose between powerful but expensive deep networks and fast but limited classical algorithms. Here we

hardwarearxiv-cs-ai
20 May 2026
Hardware

Charon: A Unified and Fine-Grained Simulator for Large-Scale LLM Training and Inference

DGX agent

arXiv:2605.17164v1 Announce Type: cross Abstract: Deploying large-scale LLM training and inference with optimal performance is exceptionally challenging due to a complex design space of parallelism st

hardwarearxiv-cs-ai
19 May 2026
Hardware

NanoQuant: Efficient Sub-1-Bit Quantization of Large Language Models

DGX agent

arXiv:2602.06694v2 Announce Type: replace Abstract: Weight-only quantization has become a standard approach for efficiently serving large language models (LLMs). However, existing methods fail to effi

hardwarearxiv-cs-lg
19 May 2026
Hardware

Stable Audio 3

DGX agent

arXiv:2605.17991v1 Announce Type: cross Abstract: Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models c

hardwarearxiv-cs-ai
19 May 2026
Model Releases

A3D: Agentic AI flow for autonomous Accelerator Design

DGX agent

arXiv:2605.15237v1 Announce Type: cross Abstract: Accelerating applications through the design of hardware accelerators can significantly enhance system performance and energy efficiency. Despite adva

model-releasesarxiv-cs-ai
18 May 2026
Hardware

An Amortized Efficiency Threshold for Comparing Neural and Heuristic Solvers in Combinatorial Optimization

DGX agent

arXiv:2605.14624v1 Announce Type: cross Abstract: A common critique of neural combinatorial-optimization solvers is that they are less energy-efficient than CPU metaheuristics, given the operational e

hardwarearxiv-cs-ai
15 May 2026
Hardware

Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI

DGX agent

arXiv:2605.14665v1 Announce Type: new Abstract: Legal reasoning is not semantic similarity search. A court judgment encodes constrained symbolic reasoning: precedent propagation, procedural state tran

hardwarearxiv-cs-ai
15 May 2026
Hardware

FALO: Fast and Accurate LiDAR 3D Object Detection on Resource-Constrained Devices

DGX agent

arXiv:2506.04499v2 Announce Type: replace Abstract: Existing LiDAR 3D object detection methods predominantely rely on sparse convolutions and/or transformers, which can be challenging to run on resour

hardwarearxiv-cs-cv
15 May 2026
Research

Towards Robotic Dexterous Hand Intelligence: A Survey

DGX agent

arXiv:2605.13925v1 Announce Type: new Abstract: Robotic dexterous hands are central to contact-rich manipulation, with rapid progress driven by advances in hardware, sensing, control, simulation, and

researcharxiv-cs-ro
15 May 2026
Hardware

XFP: Quality-Targeted Adaptive Codebook Quantization with Sparse Outlier Separation for LLM Inference

DGX agent

arXiv:2605.14844v1 Announce Type: cross Abstract: We introduce XFP, a dynamic weight quantizer for LLM inference that inverts the conventional workflow: the operator specifies reconstruction quality f

hardwarearxiv-cs-ai
15 May 2026
Hardware

Geometric Autoencoder Priors for Bayesian Inversion: Learn First Observe Later

DGX agent

arXiv:2509.19929v4 Announce Type: replace-cross Abstract: Uncertainty Quantification (UQ) is paramount for inference in engineering. A common inference task is to recover full-field information of phy

hardwarearxiv-cs-lg
14 May 2026
Hardware

Cerebras — Faster Tokens Please

DGX agent

Cerebras, a company specializing in AI accelerators and wafer-scale computing systems, is discussed in terms of its approaches to improving token generation speed in large language models, which is cr

hardwaresemianalysis
13 May 2026
Hardware

MLCommons Chakra: Advancing Performance Benchmarking and Co-design using Standardized Execution Traces

DGX agent

arXiv:2605.11333v1 Announce Type: cross Abstract: The fast pace of artificial intelligence~(AI) innovation demands an agile methodology for observation, reproduction and optimization of distributed ma

hardwarearxiv-cs-lg
13 May 2026
Hardware

Geometrically Approximated Modeling for Emitter-Centric Ray-Triangle Filtering in Arbitrarily Dynamic LiDAR Simulation

DGX agent

arXiv:2605.10457v1 Announce Type: cross Abstract: Real-time Light Detection And Ranging (LiDAR) simulation must find, per emitted ray, the closest intersecting triangle even in dynamic scenes containi

hardwarearxiv-cs-ro
12 May 2026
Hardware

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale

DGX agent

arXiv:2605.10886v1 Announce Type: cross Abstract: Recent GPU generations deliver significantly higher FLOPs using lower-precision arithmetic, such as FP8. While successfully applied to large language

hardwarearxiv-cs-ai
12 May 2026
Hardware

Non-Monotonic Latency in Apple MPS Decoding: KV Cache Interactions and Execution Regimes

DGX agent

arXiv:2605.08913v1 Announce Type: cross Abstract: Autoregressive inference is typically assumed to scale predictably with decoding length, and key-value (KV) caching is widely regarded as a universall

hardwarearxiv-cs-cl
12 May 2026
Hardware

Understand and Accelerate Memory Processing Pipeline for Disaggregated LLM Inference

DGX agent

arXiv:2603.29002v2 Announce Type: replace-cross Abstract: Modern large language models (LLMs) increasingly depends on efficient long-context processing and generation mechanisms, including sparse atte

hardwarearxiv-cs-ai
12 May 2026
Hardware

Everybody knows about the toilet maker Toto and MSG umami inventor Ajinomoto that are powering AI, but have you heard of the Taiwanese kitch…

DGX agent

This post likely discusses lesser-known Taiwanese companies in the kitchen or consumer appliance sector that play significant roles in powering AI infrastructure, drawing a parallel to how established

hardwaredylan-patel--x
11 May 2026
Hardware

SemanticDialect: Semantic-Aware Mixed-Format Quantization for Video Diffusion Transformers

DGX agent

arXiv:2603.02883v3 Announce Type: replace Abstract: Diffusion Transformers (DiTs) achieve state-of-the-art video generation quality, but their substantial memory and computational footprints hinder ed

hardwarearxiv-cs-cv
11 May 2026
Hardware

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache

DGX agent

arXiv:2605.06763v1 Announce Type: new Abstract: Sparse attention improves LLM inference efficiency by selecting a subset of key-value entries, but at the cost of potential accuracy degradation. In par

hardwarearxiv-cs-lg
11 May 2026
Hardware

SpikingBrain: Spiking Brain-inspired Large Models

DGX agent

arXiv:2509.05276v4 Announce Type: replace-cross Abstract: Mainstream Transformer-based large language models face major efficiency bottlenecks: training computation scales quadratically with sequence

hardwarearxiv-cs-ai
11 May 2026
Hardware

XiYOLO: Energy-Aware Object Detection via Iterative Architecture Search and Scaling

DGX agent

arXiv:2605.06927v1 Announce Type: cross Abstract: Object detection on heterogeneous edge devices must satisfy strict energy, latency, and memory constraints while still providing reliable perception f

hardwarearxiv-cs-ai
11 May 2026
Hardware

Great collab with @SakanaAILabs on an #ICML26 paper about sparse transformer kernels + formats optimized for modern NVIDIA GPU execution. • …

DGX agent

Great collab with @SakanaAILabs on an #ICML26 paper about sparse transformer kernels + formats optimized for modern NVIDIA GPU execution. • TwELL sparse packing • Fused CUDA kernels • 20%+ inference/t

hardwaredavid-ha--x
8 May 2026
Hardware

With faster node startup for GKE, say goodbye to cold-start latency

DGX agent

We’ve rolled out a significant update to Google Kubernetes Engine (GKE) that solves one of the most annoying problems in cloud infrastructure: cold start latency. GKE now has up to 4x faster node star

hardwaregoogle-cloud-ai
8 May 2026
Hardware

Achieving Peak System and Workload Efficiency on NVIDIA GB200 NVL72 with Slurm Block Scheduling

DGX agent

This article discusses optimization techniques for maximizing efficiency on NVIDIA's GB200 NVL72 system using Slurm block scheduling, a job scheduling approach designed to improve resource utilization

hardwarenvidia-developer
7 May 2026
Hardware

Coral: Cost-Efficient Multi-LLM Serving over Heterogeneous Cloud GPUs

DGX agent

arXiv:2605.04357v1 Announce Type: cross Abstract: The usage of large language models (LLMs) has grown increasingly fragmented, with no single model dominating. Meanwhile, cloud providers offer a wide

hardwarearxiv-cs-cl
7 May 2026
Hardware

Piper: Efficient Large-Scale MoE Training via Resource Modeling and Pipelined Hybrid Parallelism

DGX agent

arXiv:2605.05049v1 Announce Type: cross Abstract: Frontier models increasingly adopt Mixture-of-Experts (MoE) architectures to achieve large-model performance at reduced cost. However, training MoE mo

hardwarearxiv-cs-lg
7 May 2026
Hardware

Nvidia partners with Corning to boost the supply of optical network components

DGX agent

Nvidia Corp. will help publicly traded glass maker Corning Inc. boost the rate at which it produces parts for optical data center networks. The partnership, which the companies announced today, will s

hardwaresiliconangle
6 May 2026
Hardware

VUDA: Breaking CUDA-Vulkan Isolation for Spatial Sharing of Compute and Graphics on the Same GPU

DGX agent

arXiv:2605.01352v1 Announce Type: cross Abstract: GPU-based simulation environments for embodied AI interleave physics simulation (CUDA) and photorealistic rendering (Vulkan) on a single device. We ob

hardwarearxiv-cs-ai
6 May 2026
Local Ai

A Synthesizable RTL Implementation of Predictive Coding Networks

DGX agent

arXiv:2603.18066v2 Announce Type: replace-cross Abstract: Backpropagation has enabled modern deep learning but is difficult to realize as an online, fully distributed hardware learning system due to g

local-aiarxiv-cs-lg
5 May 2026
Hardware

aerial-autonomy-stack -- a Faster-than-real-time, Autopilot-agnostic, ROS2 Framework to Simulate and Deploy Perception-based Drones

DGX agent

arXiv:2602.07264v2 Announce Type: replace Abstract: Unmanned aerial vehicles are rapidly transforming multiple applications, from agricultural and infrastructure monitoring to logistics and defense. I

hardwarearxiv-cs-ro
5 May 2026
Hardware

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization

DGX agent

arXiv:2602.02958v4 Announce Type: replace Abstract: Despite rapid progress in autoregressive video diffusion, an emerging system algorithm bottleneck limits both deployability and generation capabilit

hardwarearxiv-cs-lg
5 May 2026
Hardware

Silicon Valley bets $200M on AI data centers floating in the ocean

DGX agent

Panthalassa, a US startup, is developing floating data centers powered by ocean waves and cooled by seawater to address AI infrastructure's growing energy demands. Peter Thiel led a $140 million inves

hardwarears-technica
5 May 2026
Hardware

DPU or GPU for Accelerating Neural Networks Inference -- Why not both? Split CNN Inference

DGX agent

arXiv:2605.00174v1 Announce Type: cross Abstract: Video and image streaming on edge devices requires low latency. To address this, Neural Networks (NNs) are widely used, and prior work mainly focuses

hardwarearxiv-cs-cv
4 May 2026
Hardware

AI Value Capture - The Shift To Model Labs

DGX agent

This article examines how value in the AI industry is shifting away from traditional chip manufacturers toward model development labs and companies that build large language models. It likely analyzes

hardwaresemianalysis
1 May 2026
Hardware

All the Doomers and hawks are lining up behind this distillation 'attack' farce because they want to see open source banned. It's really as …

DGX agent

All the Doomers and hawks are lining up behind this distillation 'attack' farce because they want to see open source banned. It's really as simple as that. They want to take away your right to choose,

hardwareyann-lecun--x
1 May 2026
← Previous
1…910111213…93
Next →