AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlog
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,446 results
Local Ai

CLANE: Continual Learning of Actions on Neuromorphic Hardware from Event Cameras

DGX agent

arXiv:2605.28387v1 Announce Type: cross Abstract: Recognizing and continuously learning novel human actions without forgetting prior classes is a requirement for emerging AR/VR and robotics applicatio

local-aiarxiv-cs-ai
28 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

The Energy Blind Spot: NVIDIA's Flagship Edge AI Hardware Cannot Support Process-Level Energy Attribution

DGX agent

arXiv:2605.27599v1 Announce Type: cross Abstract: Agentic AI workloads - where a single user goal triggers multi-step orchestration, tool calls, retries, and failure recovery - are being targeted for

local-aiarxiv-cs-ai
28 May 2026
Research

Hardware-Aware Federated Learning for Speech Emotion Recognition

DGX agent

arXiv:2605.24712v1 Announce Type: new Abstract: Federated learning (FL) enables privacy-preserving collaborative training across distributed edge devices, but real deployments involve heterogeneous cl

researcharxiv-cs-lg
26 May 2026
Model Releases

InnerQ: Hardware-Aware Tuning-Free Quantization of KV Cache for Large Language Models

DGX agent

arXiv:2602.23200v2 Announce Type: replace-cross Abstract: When transformer-based language models are deployed for text generation, most of the inference time is spent in the decoding stage, where outp

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

DGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

model-releasesarxiv-cs-lg
18 May 2026
Research

A Hardware-Aware, Per-Layer Methodology for Post-Training Quantization of Large Language Models

DGX agent

arXiv:2605.14929v1 Announce Type: new Abstract: Scaled Outer Product (SOP) is a post-training quantization methodology for large language model weights, designed to deliver near-lossless fidelity at 4

researcharxiv-cs-lg
15 May 2026
Safety

Evolving Layer-Specific Scalar Functions for Hardware-Aware Transformer Adaptation

DGX agent

arXiv:2605.14047v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve state-of-the-art performance on challenging vision tasks, but their deployment on edge devices is severely hindered b

safetyarxiv-cs-cv
15 May 2026
Research

MobileEgo Anywhere: Open Infrastructure for long horizon egocentric data on commodity hardware

DGX agent

arXiv:2605.05945v2 Announce Type: replace-cross Abstract: The recent advancement of Vision Language Action (VLA) models has driven a critical demand for large scale egocentric datasets. However, exist

researcharxiv-cs-cl
13 May 2026
Model Releases

CUDAHercules: Benchmarking Hardware-Aware Expert-level CUDA Optimization for LLMs

DGX agent

arXiv:2605.08467v1 Announce Type: new Abstract: Large language models show promise for automated CUDA programming, however even the strongest coding models (e.g., Claude-Opus-4.6) may still fall short

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

FastOmniTMAE: Parallel Clause Learning for Scalable and Hardware-Efficient Tsetlin Embeddings

DGX agent

arXiv:2605.06982v1 Announce Type: new Abstract: Embedding models in natural language processing (NLP) increasingly rely on deep architectures such as BERT, while simpler models such as Word2Vec provid

model-releasesarxiv-cs-lg
11 May 2026
Local Ai

MambaCSP: Hybrid-Attention State Space Models for Hardware-Efficient Channel State Prediction

DGX agent

arXiv:2604.21957v1 Announce Type: cross Abstract: Recent works have demonstrated that attention-based transformer and large language model (LLM) architectures can achieve strong channel state predicti

local-aiarxiv-cs-ai
27 Apr 2026
Safety

LogosKG: Hardware-Optimized Scalable and Interpretable Knowledge Graph Retrieval

DGX agent

arXiv:2604.18913v1 Announce Type: new Abstract: Knowledge graphs (KGs) are increasingly integrated with large language models (LLMs) to provide structured, verifiable reasoning. A core operation in th

safetyarxiv-cs-cl
22 Apr 2026
Research

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation

DGX agent

arXiv:2512.03053v2 Announce Type: replace-cross Abstract: We show for invertible problems that transform data from a source domain (for example, Logic Condition Tables (LCTs)) to a destination domain

researcharxiv-cs-ai
20 Apr 2026
Local Ai

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

DGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

local-aiarxiv-cs-ai
17 Apr 2026
Safety

Hardware-Efficient Neuro-Symbolic Networks with the Exp-Minus-Log Operator

DGX agent

arXiv:2604.13871v1 Announce Type: new Abstract: Deep neural networks (DNNs) deliver state-of-the-art accuracy on regression and classification tasks, yet two structural deficits persistently obstruct

safetyarxiv-cs-lg
16 Apr 2026
Model Releases

AHCQ-SAM: Toward Accurate and Hardware-Compatible Post-Training Segment Anything Model Quantization

DGX agent

arXiv:2503.03088v4 Announce Type: replace-cross Abstract: The Segment Anything Model (SAM) has revolutionized image and video segmentation with its powerful zero-shot capabilities. However, its massiv

model-releasesarxiv-cs-lg
10 Apr 2026
Research

Deployment Is Not Destiny: Robot Recomposition in the Field with Unseen Software, Hardware, and Compute Payloads

DGX agent

arXiv:2608.11063v1 Announce Type: new Abstract: The tight coupling of subsystems in most robots, though a natural consequence of their complexity, leads to monolithic designs that are time-consuming a

researcharxiv-cs-ro
12 Aug 2026
Local Ai

Open-weight video gen that actually delivers. Five days with MiniMax H3 on local hardware.

DGX agent

H3 weights went live on HuggingFace August 3rd and I started pulling them immediately. An omni-modal video model with native stereo audio in the same forward pass, where audio can actually drive the v

local-air-localllama
9 Aug 2026
Local Ai

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

DGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

local-aiarxiv-cs-ai
5 Aug 2026
Agents

Throwback to 2k7 vibez when I opened my first Iphone. Start of the AI Hardware era? Cool 🛳️ @OpenAI

DGX agent

On July 31, 2026 at 01:40 AM, Twitter user Murtaza Khomusi posted a “throwback” tweet reminiscing about opening his first iPhone in 2007 (“2k7 vibez”) and questioned whether this marked the start of t

agentsjerry-liu--x
31 Jul 2026
Local Ai

Hardware-Software Co-Design for Float16 On-Device Training on RISC-V Single-Core

DGX agent

arXiv:2607.21130v1 Announce Type: cross Abstract: By leveraging standard RISC-V extensions, namely Zfh (scalar float16) and Zvfh (vector float16), this work proposes an open-source framework to enable

local-aiarxiv-cs-ai
24 Jul 2026
Research

Classical Hardware Acceleration of Quantum Autoencoders for Real-Time Anomaly Detection in Collider Experiments

DGX agent

arXiv:2607.20302v1 Announce Type: new Abstract: Quantum machine learning (QML) algorithms in high energy physics (HEP) can efficiently represent and leverage long-range, high-order correlations in hig

researcharxiv-cs-lg
23 Jul 2026
Research

Kaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space Correlations

DGX agent

arXiv:2607.13770v1 Announce Type: cross Abstract: Video diffusion transformers (vDiTs) generate high quality video but introduce extremely high compute cost due to the long diffusion timesteps and sel

researcharxiv-cs-ai
16 Jul 2026
Local Ai

Rethinking Small VLM Quantization: From Component-Wise Analysis to Hardware-Aware Edge Deployment

DGX agent

arXiv:2607.08029v1 Announce Type: new Abstract: The emergence of vision language models with fewer than 3 billion parameters has accelerated the implementation of on-device multimodal intelligence. Ho

local-aiarxiv-cs-lg
10 Jul 2026
Local Ai

MedMambaLite: Hardware-Aware Mamba for Medical Image Classification

DGX agent

arXiv:2508.05049v1 Announce Type: cross Abstract: AI-powered medical devices have driven the need for real-time, on-device inference such as biomedical image classification. Deployment of deep learnin

local-aiarxiv-cs-cv
7 Jul 2026
Model Releases

AFFMAE: Scalable Vision Pre-Training for High-Resolution Microscopy Segmentation on Desktop Hardware

DGX agent

arXiv:2602.16249v2 Announce Type: replace Abstract: Self-supervised pretraining has transformed computer vision by enabling data-efficient fine-tuning, yet high-resolution pretraining typically requir

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Assistax: A Multi-Agent Hardware-Accelerated Reinforcement Learning Benchmark for Assistive Robotics

DGX agent

arXiv:2507.21638v2 Announce Type: replace Abstract: The development of reinforcement learning (RL) algorithms has been largely driven by ambitious challenge tasks and benchmarks. Games have dominated

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Benchmarking Local LLMs for Natural-Language-to-SQL Querying in Biopharmaceutical Manufacturing: An Empirical Benchmark on Consumer-Grade Hardware

DGX agent

arXiv:2606.01338v1 Announce Type: new Abstract: Biopharmaceutical manufacturing organizations operate under regulatory frameworks such as FDA guidance, EU Good Manufacturing Practice (GMP), and the EU

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

HASTE: Hardware-Aware Dynamic Sparse Training for Large Output Spaces

DGX agent

arXiv:2606.01117v1 Announce Type: cross Abstract: Extreme multi-label classification (XMC) involves learning models over large output spaces with millions of labels, making the output layer a memory-c

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

EventShiftFlow: Towards Hardware-efficient FPGA-based Flow Estimation

DGX agent

arXiv:2605.28312v1 Announce Type: cross Abstract: Event-based vision sensors offer asynchronous, high-temporal-resolution measurements that are attractive for low-latency robotic perception, but many

model-releasesarxiv-cs-cv
28 May 2026
Tutorials

[Guide] How to securely run ComfyUI on Windows (Docker>WSL2) [RTX 3090, logic can be applied to other hardware]

DGX agent

ComfyUI is a node-based interface for Stable Diffusion that can be run on Windows using Docker and WSL2 (Windows Subsystem for Linux 2) for improved security and performance. This guide provides instr

tutorialsr-stablediffusion
28 May 2026
Model Releases

GraphRAG on Consumer Hardware: Benchmarking Local LLMs for Healthcare EHR Schema Retrieval

DGX agent

arXiv:2605.20815v1 Announce Type: new Abstract: Graph-based Retrieval Augmented Generation (GraphRAG) extends retrieval-augmented generation to support structured reasoning over complex corpora, but i

model-releasesarxiv-cs-cl
21 May 2026
Agents

Neuromorphic Control of a Flapping-Wing Robot on Resource-Constrained Hardware

DGX agent

arXiv:2605.19430v1 Announce Type: new Abstract: Flapping-Wing Micro Aerial Vehicles (FWMAVs) provide exceptional maneuverability and aerodynamic efficiency but pose significant challenges for onboard

agentsarxiv-cs-ro
20 May 2026
Model Releases

GQLA: Group-Query Latent Attention for Hardware-Adaptive Large Language Model Decoding

DGX agent

arXiv:2605.15250v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), the attention used in DeepSeek-V2/V3, jointly compresses keys and values into a low-rank latent and matches the H10

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

G-SHARP: Gaussian Surgical Hardware Accelerated Real-time Pipeline

DGX agent

arXiv:2512.02482v2 Announce Type: replace Abstract: We propose G-SHARP, a commercially compatible, real-time surgical scene reconstruction framework designed for minimally invasive procedures that req

model-releasesarxiv-cs-cv
15 May 2026
Local Ai

Use Case: Invoice processing with local LLM - Which LLM and hardware requirements?

DGX agent

This discussion explores using Ollama to run large language models locally for invoice processing while maintaining control over data. The thread likely addresses selecting appropriate smaller models

local-air-ollama
12 May 2026
Model Releases

Real-Time Frame- and Event-based Object Detection with Spiking Neural Networks on Edge Neuromorphic Hardware: Design, Deployment and Benchmark

DGX agent

arXiv:2605.00146v1 Announce Type: new Abstract: Real-time object detection on energy-constrained platforms is critical for applications such as UAV-based inspection, autonomous navigation, and mobile

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Quantum Feature Selection with Higher-Order Binary Optimization on Trapped-Ion Hardware

DGX agent

arXiv:2604.26834v1 Announce Type: cross Abstract: We present a quantum feature-selection framework based on a higher-order unconstrained binary optimization (HUBO) formulation that explicitly incorpor

model-releasesarxiv-cs-lg
30 Apr 2026
Research

Quantum Gatekeeper: Multi-Factor Context-Bound Image Steganography with VQC Based Key Derivation on Quantum Hardware

DGX agent

arXiv:2604.26413v1 Announce Type: cross Abstract: This paper presents Quantum Gatekeeper, a context-bound image steganography framework where successful payload recovery depends on both cryptographic

researcharxiv-cs-ai
30 Apr 2026
Model Releases

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining

DGX agent

arXiv:2602.10718v3 Announce Type: replace-cross Abstract: While FP8 attention has shown substantial promise in innovations like FlashAttention-3, its integration into the decoding phase of the DeepSee

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Ternary Memristive Logic: Hardware for Reasoning Realized via Domain Algebra

DGX agent

arXiv:2604.20891v1 Announce Type: cross Abstract: Memristive crossbars store numerical weights needing aggregation and decoding; a single junction means nothing alone. This paper presents a fundamenta

researcharxiv-cs-ai
24 Apr 2026
Model Releases

Benchmarking Quantum Kernel Support Vector Machines Against Classical Baselines on Tabular Data: A Rigorous Empirical Study with Hardware Validation

DGX agent

arXiv:2604.18837v1 Announce Type: cross Abstract: Quantum kernel methods have been proposed as a promising approach for leveraging near-term quantum computers for supervised learning, yet rigorous ben

model-releasesarxiv-cs-lg
22 Apr 2026
Local Ai

Best Ollama model for n8n workflows (RAG, file handling, reasoning) + hardware requirements?

DGX agent

Models like Qwen3 and Llama 3.2 are commonly used for n8n RAG workflows , with selection depending on use case requirements. Running Ollama models locally requires at least 16 GB of RAM on your device

local-air-ollama
17 Apr 2026
Model Releases

EdgeCIM: A Hardware-Software Co-Design for CIM-Based Acceleration of Small Language Models

DGX agent

arXiv:2604.11512v1 Announce Type: cross Abstract: The growing demand for deploying Small Language Models (SLMs) on edge devices, including laptops, smartphones, and embedded platforms, has exposed fun

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

IC LoRAs for LTX2 have so much potential - you can train SOTA control video capabilities on potato hardware - 4 examples w/ links below

DGX agent

IC-LoRA (In-Context LoRA) enables conditioning video generation on reference video frames at inference time, allowing fine-grained video-to-video control on top of a text-to-video base model. Unlike t

local-air-stablediffusion
14 Apr 2026
Model Releases

Toward Hardware-Agnostic Quadrupedal World Models via Morphology Conditioning

DGX agent

arXiv:2604.08780v1 Announce Type: cross Abstract: World models promise a paradigm shift in robotics, where an agent learns the underlying physics of its environment once to enable efficient planning a

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

DGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

model-releasesarxiv-cs-lg
11 Aug 2026
Local Ai

Quick survey (2 min) on trust in hardware specs for open-source models

DGX agent

Hi everyone, I'm a systems analysis student researching a problem a lot of you probably know well: how much you actually trust the published VRAM/RAM requirements for open-source models before trying

local-air-ollama
8 Aug 2026
← Previous
1…34567…93
Next →