AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Local Ai

The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution

DGX agent

arXiv:2605.27599v1 Announce Type: cross Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for

local-aiarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Hardware-Aware Federated Learning for Speech Emotion Recognition

DGX agent

arXiv:2605.24712v1 Announce Type: new Abstract: Federated learning (FL) enables privacy-preserving collaborative training across distributed edge devices, but real deployments involve heterogeneous cl

researcharxiv-cs-lg
26 May 2026
Model Releases

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

DGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

DGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

model-releasesarxiv-cs-lg
18 May 2026
Research

A Hardware-Aware, Per-Layer Methodology for Post-Training Quantization of Large Language Models

DGX agent

arXiv:2605.14929v1 Announce Type: new Abstract: Scaled Outer Product (SOP) is a post-training quantization methodology for large language model weights, designed to deliver near-lossless fidelity at 4

researcharxiv-cs-lg
15 May 2026
Safety

Evolving Layer-Specific Scalar Functions for Hardware-Aware Transformer Adaptation

DGX agent

arXiv:2605.14047v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve state-of-the-art performance on challenging vision tasks, but their deployment on edge devices is severely hindered b

safetyarxiv-cs-cv
15 May 2026
Research

MobileEgo Anywhere: Open Infrastructure for long horizon egocentric data on commodity hardware

DGX agent

arXiv:2605.05945v2 Announce Type: replace-cross Abstract: The recent advancement of Vision Language Action (VLA) models has driven a critical demand for large scale egocentric datasets. However, exist

researcharxiv-cs-cl
13 May 2026
Model Releases

CUDAHercules: Benchmarking Hardware-Aware Expert-level CUDA Optimization for LLMs

DGX agent

arXiv:2605.08467v1 Announce Type: new Abstract: Large language models show promise for automated CUDA programming, however even the strongest coding models (e.g., Claude-Opus-4.6) may still fall short

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

FastOmniTMAE: Parallel Clause Learning for Scalable and Hardware-Efficient Tsetlin Embeddings

DGX agent

arXiv:2605.06982v1 Announce Type: new Abstract: Embedding models in natural language processing (NLP) increasingly rely on deep architectures such as BERT, while simpler models such as Word2Vec provid

model-releasesarxiv-cs-lg
11 May 2026
Local Ai

MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction

DGX agent

arXiv:2604.21957v1 Announce Type: cross Abstract: Recent works have demonstrated that attention-based transformer and large language model (LLM) architectures can achieve strong channel state predicti

local-aiarxiv-cs-ai
27 Apr 2026
Safety

LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval

DGX agent

arXiv:2604.18913v1 Announce Type: new Abstract: Knowledge graphs (KGs) are increasingly integrated with large language models (LLMs) to provide structured, verifiable reasoning. A core operation in th

safetyarxiv-cs-cl
22 Apr 2026
Research

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation

DGX agent

arXiv:2512.03053v2 Announce Type: replace-cross Abstract: We show for invertible problems that transform data from a source domain (for example, Logic Condition Tables (LCTs)) to a destination domain

researcharxiv-cs-ai
20 Apr 2026
Local Ai

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

DGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

local-aiarxiv-cs-ai
17 Apr 2026
Safety

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

DGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization

DGX agent

arXiv:2503.03088v4 Announce Type: replace-cross Abstract: The Segment Anything Model (SAM) has revolutionized image and video segmentation with its powerful zero-shot capabilities. However, its massiv

model-releasesarxiv-cs-lg
10 Apr 2026
Research

Deployment Is Not Destiny: Robot Recomposition in the Field with Unseen Software, Hardware, and Compute Payloads

DGX agent

arXiv:2608.11063v1 Announce Type: new Abstract: The tight coupling of subsystems in most robots, though a natural consequence of their complexity, leads to monolithic designs that are time-consuming a

researcharxiv-cs-ro
12 Aug 2026
Local Ai

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

DGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

local-aiarxiv-cs-ai
5 Aug 2026
Local Ai

Hardware-Software Co-Design for Float16 On-Device Training on RISC-V Single-Core

DGX agent

arXiv:2607.21130v1 Announce Type: cross Abstract: By leveraging standard RISC-V extensions, namely Zfh (scalar float16) and Zvfh (vector float16), this work proposes an open-source framework to enable

local-aiarxiv-cs-ai
24 Jul 2026
Research

Classical Hardware Acceleration of Quantum Autoencoders for Real-Time Anomaly Detection in Collider Experiments

DGX agent

arXiv:2607.20302v1 Announce Type: new Abstract: Quantum machine learning (QML) algorithms in high energy physics (HEP) can efficiently represent and leverage long-range, high-order correlations in hig

researcharxiv-cs-lg
23 Jul 2026
Research

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations

DGX agent

arXiv:2607.13770v1 Announce Type: cross Abstract: Video diffusion transformers (vDiTs) generate high quality video but introduce extremely high compute cost due to the long diffusion timesteps and sel

researcharxiv-cs-ai
16 Jul 2026
Local Ai

Rethinking Small VLM Quantization: From Component-Wise Analysis to Hardware-Aware Edge Deployment

DGX agent

arXiv:2607.08029v1 Announce Type: new Abstract: The emergence of vision language models with fewer than 3 billion parameters has accelerated the implementation of on-device multimodal intelligence. Ho

local-aiarxiv-cs-lg
10 Jul 2026
Local Ai

MedMambaLite: Hardware-Aware Mamba for Medical Image Classification

DGX agent

arXiv:2508.05049v1 Announce Type: cross Abstract: AI-powered medical devices have driven the need for real-time, on-device inference such as biomedical image classification. Deployment of deep learnin

local-aiarxiv-cs-cv
7 Jul 2026
Model Releases

AFFMAE: Scalable Vision Pre-Training for High-Resolution Microscopy Segmentation on Desktop Hardware

DGX agent

arXiv:2602.16249v2 Announce Type: replace Abstract: Self-supervised pretraining has transformed computer vision by enabling data-efficient fine-tuning, yet high-resolution pretraining typically requir

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

DGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Benchmarking Local LLMs for Natural-Language-to-SQL Querying in Biopharmaceutical Manufacturing: An Empirical Benchmark on Consumer-Grade Hardware

DGX agent

arXiv:2606.01338v1 Announce Type: new Abstract: Biopharmaceutical manufacturing organizations operate under regulatory frameworks such as FDA guidance, EU Good Manufacturing Practice (GMP), and the EU

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

DGX agent

arXiv:2606.01117v1 Announce Type: cross Abstract: Extreme multi-label classification (XMC) involves learning models over large output spaces with millions of labels, making the output layer a memory-c

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

EventShiftFlow: Towards Hardware-efficient FPGA-based Flow Estimation

DGX agent

arXiv:2605.28312v1 Announce Type: cross Abstract: Event-based vision sensors offer asynchronous, high-temporal-resolution measurements that are attractive for low-latency robotic perception, but many

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval

DGX agent

arXiv:2605.20815v1 Announce Type: new Abstract: Graph-based Retrieval Augmented Generation (GraphRAG) extends retrieval-augmented generation to support structured reasoning over complex corpora, but i

model-releasesarxiv-cs-cl
21 May 2026
Agents

Neuromorphic Control of a Flapping-Wing Robot on Resource-Constrained Hardware

DGX agent

arXiv:2605.19430v1 Announce Type: new Abstract: Flapping-Wing Micro Aerial Vehicles (FWMAVs) provide exceptional maneuverability and aerodynamic efficiency but pose significant challenges for onboard

agentsarxiv-cs-ro
20 May 2026
Model Releases

GQLA: Group-Query Latent Attention for Hardware-Adaptive Large Language Model Decoding

DGX agent

arXiv:2605.15250v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), the attention used in DeepSeek-V2/V3, jointly compresses keys and values into a low-rank latent and matches the H10

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

G-SHARP: Gaussian Surgical Hardware Accelerated Real-time Pipeline

DGX agent

arXiv:2512.02482v2 Announce Type: replace Abstract: We propose G-SHARP, a commercially compatible, real-time surgical scene reconstruction framework designed for minimally invasive procedures that req

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Real-Time Frame- and Event-based Object Detection with Spiking Neural Networks on Edge Neuromorphic Hardware: Design, Deployment and Benchmark

DGX agent

arXiv:2605.00146v1 Announce Type: new Abstract: Real-time object detection on energy-constrained platforms is critical for applications such as UAV-based inspection, autonomous navigation, and mobile

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Quantum Feature Selection with Higher-Order Binary Optimization on Trapped-Ion Hardware

DGX agent

arXiv:2604.26834v1 Announce Type: cross Abstract: We present a quantum feature-selection framework based on a higher-order unconstrained binary optimization (HUBO) formulation that explicitly incorpor

model-releasesarxiv-cs-lg
30 Apr 2026
Research

Quantum Gatekeeper: Multi-Factor Context-Bound Image Steganography with VQC Based Key Derivation on Quantum Hardware

DGX agent

arXiv:2604.26413v1 Announce Type: cross Abstract: This paper presents Quantum Gatekeeper, a context-bound image steganography framework where successful payload recovery depends on both cryptographic

researcharxiv-cs-ai
30 Apr 2026
Model Releases

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining

DGX agent

arXiv:2602.10718v3 Announce Type: replace-cross Abstract: While FP8 attention has shown substantial promise in innovations like FlashAttention-3, its integration into the decoding phase of the DeepSee

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Ternary Memristive Logic: Hardware for Reasoning Realized via Domain Algebra

DGX agent

arXiv:2604.20891v1 Announce Type: cross Abstract: Memristive crossbars store numerical weights needing aggregation and decoding; a single junction means nothing alone. This paper presents a fundamenta

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Benchmarking Quantum Kernel Support Vector Machines Against Classical Baselines on Tabular Data: A Rigorous Empirical Study with Hardware Validation

DGX agent

arXiv:2604.18837v1 Announce Type: cross Abstract: Quantum kernel methods have been proposed as a promising approach for leveraging near-term quantum computers for supervised learning, yet rigorous ben

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

EdgeCIM: A Hardware-Software Co-Design for CIM-Based Acceleration of Small Language Models

DGX agent

arXiv:2604.11512v1 Announce Type: cross Abstract: The growing demand for deploying Small Language Models (SLMs) on edge devices, including laptops, smartphones, and embedded platforms, has exposed fun

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning

DGX agent

arXiv:2604.08780v1 Announce Type: cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the underlying physics of its environment once to enable efficient planning a

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

DGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

When Do Fewer Visual Tokens Accelerate Multimodal Inference? A Break-Even Study Across Decision Locations and Hardware

DGX agent

arXiv:2608.03649v1 Announce Type: new Abstract: Fewer visual tokens do not guarantee lower end-to-end latency. We evaluate break-even with a reproducible protocol that accounts for decision overhead,

model-releasesarxiv-cs-cv
5 Aug 2026
Hardware

AgenticCANN: Automated Ascend C Operator Generation via Knowledge-Augmented Agentic Evolution

DGX agent

arXiv:2607.26661v1 Announce Type: new Abstract: Ascend C operator optimization is critical for NPU (Neural Processing Unit) inference performance but requires deep hardware expertise.While large langu

hardwarearxiv-cs-ai
31 Jul 2026
Model Releases

Low-cost Embedded Breathing Rate Determination Using 802.15.4z IR-UWB Hardware for Remote Healthcare

DGX agent

arXiv:2504.03772v3 Announce Type: replace-cross Abstract: Respiratory diseases account for a significant portion of global mortality. Affordable and early detection is an effective way of addressing t

model-releasesarxiv-cs-lg
30 Jul 2026
Local Ai

Self-organizing Architecture of Receptron Units: a Hardware-Aware Framework for Edge Intelligence

DGX agent

arXiv:2607.20162v1 Announce Type: new Abstract: The growing demand for intelligent processing at the edge of IoT networks is constrained by the severe computational and memory limitations of microcont

local-aiarxiv-cs-lg
23 Jul 2026
Model Releases

Industrial Dexterity Benchmark: A Hardware-Software Benchmarking Platform for Industrial Dexterous Manipulation

DGX agent

arXiv:2607.14021v1 Announce Type: new Abstract: Dexterous manipulation remains a critical bottleneck in industrial automation; tasks such as cable routing, connector insertion, and precision assembly

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

RF Spectrogram Anomaly Detection with Quantum Kitchen Sinks: Architecture, Representation, and Hardware Validation

DGX agent

arXiv:2607.13897v1 Announce Type: new Abstract: The broadcast nature of wireless channels exposes radio-frequency (RF) networks to anomalous and malicious transmissions, making anomaly detection a fun

model-releasesarxiv-cs-lg
16 Jul 2026
Local Ai

Smart Scissor: Coupling Spatial Redundancy Reduction and CNN Compression for Embedded Hardware

DGX agent

arXiv:2607.06915v1 Announce Type: new Abstract: Scaling down the resolution of input images can greatly reduce the computational overhead of convolutional neural networks (CNNs), which is promising fo

local-aiarxiv-cs-cv
9 Jul 2026
Research

A Spiking Sequence Generator for Polar Trajectories on Neuromorphic Hardware

DGX agent

arXiv:2607.02753v1 Announce Type: cross Abstract: Neuromorphic controllers for size, weight, and power-constrained systems require neural architectures that are both energy-efficient and interpretable

researcharxiv-cs-ro
7 Jul 2026
← Previous
12345…55
Next →