AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Research

Ising Acceleration for Multi-Robot Multi-Target Planning

DGX agent

arXiv:2608.06803v1 Announce Type: cross Abstract: Ising machines are emerging as promising hardware for combinatorial optimization. With recent advances in CMOS Ising technology, they are becoming att

researcharxiv-cs-ro
10 Aug 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AFD-Ledger: Deployment Provisioning for Attention--FFN Disaggregation

DGX agent

arXiv:2608.04502v1 Announce Type: cross Abstract: Attention--Feed-Forward Network (FFN) Disaggregation (AFD) is emerging as a promising architecture for serving Mixture-of-Experts (MoE) language model

researcharxiv-cs-ai
6 Aug 2026
Model Releases

RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis

DGX agent

arXiv:2602.11506v4 Announce Type: replace-cross Abstract: The transition toward localized intelligence through Small Language Models (SLMs) has intensified the need for rigorous performance characteri

model-releasesarxiv-cs-ai
6 Aug 2026
Applications

SAMSEM -- A Generic and Scalable Approach for IC Metal Line Segmentation

DGX agent

arXiv:2603.16548v2 Announce Type: replace-cross Abstract: In light of globalized hardware supply chains, the assurance of hardware components has gained significant interest, particularly in cryptogra

applicationsarxiv-cs-cv
5 Aug 2026
Hardware

GoGoTB: Agentic RTL Verification with Specification-Grounded Coverage Closure

DGX agent

arXiv:2607.26181v1 Announce Type: new Abstract: Functional verification dominates integrated circuit (IC) front-end engineering effort, and a single missed bug that escapes to silicon can trigger a co

hardwarearxiv-cs-ai
31 Jul 2026
Hardware

Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline

DGX agent

arXiv:2607.22997v1 Announce Type: cross Abstract: Physical AI -- the integration of large vision-language-action (VLA) models with embodied agents that act in the real world -- has emerged as the next

hardwarearxiv-cs-ai
28 Jul 2026
Model Releases

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

DGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

PN-QNN: Harnessing Physical Noise as a Native Regularizer in Photonic Hybrid Quantum Neural Networks

DGX agent

arXiv:2607.20045v1 Announce Type: cross Abstract: Physical noise in near-term quantum hardware is usually treated as a nuisance to suppress. We ask whether it can instead act as a hardware-native regu

model-releasesarxiv-cs-lg
23 Jul 2026
Hardware

ReRAM-aware Model Finetuning addressing I-V Non-linearity and Retention Errors

DGX agent

arXiv:2606.17471v3 Announce Type: replace Abstract: Traditional CPU, GPU, and NPU architectures are increasingly limited by the von Neumann bottleneck. While In-Memory Computing (IMC) using ReRAM cros

hardwarearxiv-cs-lg
23 Jul 2026
Hardware

BitLogic: Training Framework for Gradient-Based FPGA-Native Neural Networks

DGX agent

arXiv:2602.07400v2 Announce Type: replace Abstract: Gradient-based LUT- and logic-gate-based neural networks (LUTNet, LogicNets, DiffLogic, PolyLUT, NeuraLUT, WARP-LUT, DWN, LILogicNet, LightLUT) repl

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

DBNN: Neural Spike Classification Using a Deep Binarized Neural Network

DGX agent

arXiv:2607.05590v1 Announce Type: cross Abstract: Implantable brain-computer interfaces require on-node spike sorting to reduce telemetry bandwidth and power while maintaining reliable neural decoding

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

Life Cycle Assessment of Pre-training the Lucie 7B Open-Source Large Language Model on the Jean Zay Supercomputer

DGX agent

arXiv:2607.05408v1 Announce Type: cross Abstract: The environmental impact of training large language models (LLMs) is increasingly scrutinised, yet most published estimates focus on operational energ

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation

DGX agent

arXiv:2607.06558v1 Announce Type: new Abstract: Scaling robot learning requires massive, diverse trajectory data, yet collection is currently bottlenecked by physical teleoperation, where every demons

hardwarearxiv-cs-ro
8 Jul 2026
Hardware

MLSYSIM: First-Principles Infrastructure Modeling for Machine Learning Systems

DGX agent

arXiv:2607.02558v1 Announce Type: cross Abstract: As machine learning shifts from laboratory curiosity to critical infrastructure, the systems that sustain it span an extraordinary range, from sub-mil

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

WattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs

DGX agent

arXiv:2607.02391v1 Announce Type: cross Abstract: Large Language Model (LLM) inference workloads are a rapidly growing contributor to data center energy consumption. Optimizing these deployments requi

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

CSAR: Containerized System Architecture for Robotics

DGX agent

arXiv:2606.30293v1 Announce Type: new Abstract: Robotic applications increasingly rely on distributed computational infrastructures that combine embedded devices, edge servers, and cloud resources. Th

hardwarearxiv-cs-ro
30 Jun 2026
Hardware

Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization

DGX agent

arXiv:2606.26453v1 Announce Type: new Abstract: We present KernelPro, a closed-loop multi-agent system that automatically generates, profiles, and iteratively optimizes GPU kernel code by integrating

hardwarearxiv-cs-lg
26 Jun 2026
Agents

Towards Autonomous Accelerator Design: FPGA Accelerator Generation with SECDA

DGX agent

arXiv:2606.11117v1 Announce Type: cross Abstract: Designing FPGA-based accelerators for modern artificial intelligence workloads requires exploring a large and complex hardware design space that invol

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

DGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

model-releasesarxiv-cs-lg
9 Jun 2026
Hardware

Towards Automated Kernel Generation in the Era of LLMs

DGX agent

arXiv:2601.15727v3 Announce Type: replace Abstract: The performance of modern AI systems is fundamentally constrained by the quality of their underlying GPU kernels, which translate high-level algorit

hardwarearxiv-cs-lg
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

Bit-Exact AI Inference Verification Without Performance Tradeoffs

DGX agent

arXiv:2606.00279v1 Announce Type: cross Abstract: Verifying claims about AI workloads is a pre- requisite for credible AI governance of covert adversaries (who comply with monitoring only when detecti

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Bastion: Budget-Aware Speculative Decoding with Tree-structured Block Diffusion Drafting

DGX agent

arXiv:2605.29727v1 Announce Type: new Abstract: Block-diffusion drafters have recently emerged as a powerful alternative for speculative decoding by predicting multiple future-token distributions in a

hardwarearxiv-cs-lg
29 May 2026
Hardware

Inference Time Context Sparsity: Illusion or Opportunity?

DGX agent

arXiv:2605.24168v1 Announce Type: new Abstract: Sparsity has long been a central theme in LLM efficiency, but its role in context processing remains unresolved. As LLM workloads shift toward longer co

hardwarearxiv-cs-ai
26 May 2026
Hardware

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration

DGX agent

arXiv:2605.22015v1 Announce Type: new Abstract: Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spa

hardwarearxiv-cs-cv
22 May 2026
Hardware

COBALT: Crowdsourcing Robot Learning via Cloud-Based Teleoperation with Smartphones

DGX agent

arXiv:2605.19138v1 Announce Type: cross Abstract: The scarcity of large-scale, high-quality demonstration data remains a bottleneck in scaling imitation learning for robotic manipulation. We present C

hardwarearxiv-cs-ai
20 May 2026
Hardware

GATO: GPU-Accelerated and Batched Trajectory Optimization for Scalable Edge Model Predictive Control

DGX agent

arXiv:2510.07625v2 Announce Type: replace Abstract: While Model Predictive Control (MPC) delivers strong performance across robotics applications, solving the underlying (batches of) nonlinear traject

hardwarearxiv-cs-ro
11 May 2026
Hardware

LLMSpace: Carbon Footprint Modeling for Large Language Model Inference on LEO Satellites

DGX agent

arXiv:2605.05615v2 Announce Type: replace Abstract: Large language models (LLMs) impose rapidly growing energy demands, creating an emerging energy and carbon crisis driven by large-scale inference. S

hardwarearxiv-cs-lg
11 May 2026
Hardware

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

DGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

hardwarearxiv-cs-lg
11 May 2026
Hardware

When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

DGX agent

arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol

hardwarearxiv-cs-ai
7 May 2026
Hardware

A Semantic Autonomy Framework for VLM-Integrated Indoor Mobile Robots: Hybrid Deterministic Reasoning and Cross-Robot Adaptive Memory

DGX agent

arXiv:2605.02525v1 Announce Type: new Abstract: Autonomous indoor mobile robots can navigate reliably to metric coordinates using established frameworks such as ROS 2 Navigation 2, yet they lack the a

hardwarearxiv-cs-ro
5 May 2026
Hardware

The Quantization Trap: Breaking Linear Scaling Laws in Multi-Hop Reasoning

DGX agent

arXiv:2602.13595v2 Announce Type: replace Abstract: Neural scaling laws provide a predictable recipe for AI advancement: reducing numerical precision should linearly improve computational efficiency a

hardwarearxiv-cs-ai
5 May 2026
Hardware

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization

DGX agent

arXiv:2605.02262v1 Announce Type: cross Abstract: Recently, video language models (VLMs) have been applied in various fields. However, the visual token sequence of the VLM is too long, which may cause

hardwarearxiv-cs-cl
5 May 2026
Hardware

EdgeFM: Efficient Edge Inference for Vision-Language Models

DGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

hardwarearxiv-cs-cv
1 May 2026
Model Releases

EdgeSpike: Spiking Neural Networks for Low-Power Autonomous Sensing in Edge IoT Architectures

DGX agent

arXiv:2604.27004v1 Announce Type: cross Abstract: We propose EdgeSpike, a co-designed spiking neural network (SNN) framework for autonomous low-power sensing in edge Internet of Things (IoT) architect

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics

DGX agent

arXiv:2604.24916v1 Announce Type: new Abstract: We introduce asRoBallet, to the best of our knowledge, the first successful deployment of reinforcement learning (RL) on a humanoid ballbot hardware. Hi

model-releasesarxiv-cs-ro
29 Apr 2026
Hardware

Latent Inter-Frame Pruning: A Training-Free Method Bridging Traditional Video Compression and Modern Diffusion Transformers for Efficient Generation

DGX agent

arXiv:2604.23858v1 Announce Type: new Abstract: Video generation, while capable of generating realistic videos, is computationally expensive and slow, prohibiting real-time applications. In this paper

hardwarearxiv-cs-cv
28 Apr 2026
Applications

HGQ-LUT: Fast LUT-Aware Training and Efficient Architectures for DNN Inference

DGX agent

arXiv:2604.22293v1 Announce Type: cross Abstract: Lookup-table (LUT) based neural networks can deliver ultra-low latency and excellent hardware efficiency on FPGAs by mapping arithmetic operations dir

applicationsarxiv-cs-lg
27 Apr 2026
Hardware

A Delta-Aware Orchestration Framework for Scalable Multi-Agent Edge Computing

DGX agent

arXiv:2604.20129v1 Announce Type: new Abstract: The Synergistic Collapse occurs when scaling beyond 100 agents causes superlinear performance degradation that individual optimizations cannot prevent.

hardwarearxiv-cs-lg
23 Apr 2026
Hardware

Amortized Vine Copulas for High-Dimensional Density and Information Estimation

DGX agent

arXiv:2604.20568v1 Announce Type: new Abstract: Modeling high-dimensional dependencies while keeping likelihoods tractable remains challenging. Classical vine-copula pipelines are interpretable but ca

hardwarearxiv-cs-lg
23 Apr 2026
Hardware

PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment

DGX agent

arXiv:2604.19129v1 Announce Type: new Abstract: Existing facial reenactment methods struggle with a trade-off between expressiveness and fine-grained controllability. Holistic facial reenactment model

hardwarearxiv-cs-cv
22 Apr 2026
Hardware

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

DGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation

DGX agent

arXiv:2604.18348v1 Announce Type: new Abstract: Video diffusion transformers (DiTs) suffer from prohibitive inference latency due to quadratic attention complexity. Existing sparse attention methods e

hardwarearxiv-cs-cv
21 Apr 2026
Hardware

Real-Time Structural Detection for Indoor Navigation from 3D LiDAR Using Bird's-Eye-View Images

DGX agent

arXiv:2603.19830v2 Announce Type: replace Abstract: Efficient structural perception is essential for mapping and autonomous navigation on resource-constrained robots. Existing 3D methods are computati

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

CPU Optimization of a Monocular 3D Biomechanics Pipeline for Low-Resource Deployment

DGX agent

arXiv:2604.15665v1 Announce Type: new Abstract: Markerless 3D movement analysis from monocular video enables accessible biomechanical assessment in clinical and sports settings. However, most research

hardwarearxiv-cs-cv
20 Apr 2026
Hardware

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

DGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

DGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

Automatic Charge State Tuning of 300 mm FDSOI Quantum Dots Using Neural Network Segmentation of Charge Stability Diagram

DGX agent

arXiv:2604.13662v1 Announce Type: cross Abstract: Tuning of gate-defined semiconductor quantum dots (QDs) is a major bottleneck for scaling spin qubit technologies. We present a deep learning (DL) dri

hardwarearxiv-cs-cv
16 Apr 2026
← Previous
1…34567…55
Next →