AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
18 May 2026

Searching on a Budget: HW-NAS with 10 Latency Probes

Model ReleasesDGX agent

arXiv:2504.00663v2 Announce Type: replace Abstract: Existing hardware-aware NAS (HW-NAS) methods typically assume access to precise information circa the target device, either via analytical approxima

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

Local AiDGX agent

arXiv:2605.16138v1 Announce Type: cross Abstract: Neural architecture search (NAS) is a powerful approach for automating model design, but existing methods often optimize for accuracy alone or rely on

17 May 2026

b9197

Local AiDGX agent

b9197 is a build release of llama.cpp , an open-source C/C++ implementation that enables efficient large language model inference on various hardware platforms. The release includes cross-platform bin

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
15 May 2026

AQKA: Active Quantum Kernel Acquisition Under a Shot Budget

SafetyDGX agent

arXiv:2605.14672v1 Announce Type: new Abstract: Estimating an N imes N quantum kernel from circuit fidelities requires Theta(N^2 S) measurement shots, the dominant bottleneck for deployment on near-te

b9159

Local AiDGX agent

b9159 is a release of llama.cpp published on May 14, 2026 . llama.cpp is a C/C++ implementation of large language model inference that enables efficient LLM execution on consumer hardware with minimal

b9172

Local AiDGX agent

b9172 is a release of llama.cpp that includes binaries for macOS, Linux, Android, Windows, and openEuler platforms with support for various hardware configurations including CPU, Vulkan, CUDA, ROCm, O

Natural Synthesis: Outperforming Reactive Synthesis Tools with Large Reasoning Models

ResearchDGX agent

arXiv:2605.15131v1 Announce Type: new Abstract: Reactive synthesis, the problem of automatically constructing a hardware circuit from a logical specification, is a long-standing challenge in formal ve

Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2511.16964v2 Announce Type: replace-cross Abstract: Maximizing performance on available GPU hardware is an ongoing challenge for modern AI inference systems. Traditional approaches include writi

Safe Bayesian Optimization for Complex Control Systems via Additive Gaussian Processes

Model ReleasesDGX agent

arXiv:2408.16307v3 Announce Type: replace-cross Abstract: Automatic controller tuning is attractive for robotics and mechatronic systems whose dynamics are difficult to model accurately, but direct bl

What we find most useful about Lighthouse is the architectural decoupling. The selection logic is fully separated from the attention kernel,…

ResearchDGX agent

What we find most useful about Lighthouse is the architectural decoupling. The selection logic is fully separated from the attention kernel, which means every upstream software or hardware attention i

14 May 2026

b9142

Local AiDGX agent

Release b9142 is an intermediate build of llama.cpp, a C/C++ library for running large language model inference on consumer hardware. llama.cpp performs inference on various large language models and

Controllable Quantum Memory Capacity in Quantum Reservoir Networks with Tunable partial-SWAPs

Model ReleasesDGX agent

arXiv:2605.12713v1 Announce Type: cross Abstract: In the field of quantum reservoir computing (QRC), many different computational models and architectures have been proposed. From these models, we ide

Data center cooling tech startup Iceotope aims to scale after raising $26M

IndustryDGX agent

Liquid cooling system startup Iceotope Technologies Ltd. said today it has closed on a 26 million Series B round of funding as it strives to accelerate the adoption of its data center-focused hardware

Energy Scaling Laws for Diffusion Models: Quantifying Compute in Image Generation

Model ReleasesDGX agent

arXiv:2511.17031v2 Announce Type: replace-cross Abstract: The rapidly growing computational demands of diffusion models for image generation have raised significant concerns about energy consumption a

HLS-Seek: QoR-Aware Code Generation for High-Level Synthesis via Proxy Comparative Reward Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.13536v1 Announce Type: cross Abstract: High-Level Synthesis (HLS) compiles algorithmic C/C++ descriptions into hardware, with Quality of Results (QoR) -- latency and resource utilization --

PG-LRF: Physiology-Guided Latent Rectified Flow for Electro-Hemodynamic PPG-to-ECG Generation

SafetyDGX agent

arXiv:2605.12541v1 Announce Type: cross Abstract: Electrocardiography (ECG) is the clinical standard for cardiac assessment but requires dedicated hardware that does not scale to daily-life monitoring

Physics Guided Generative Optimization for Trotter Suzuki Decomposition

ResearchDGX agent

arXiv:2605.13268v1 Announce Type: cross Abstract: Product formulas for Trotter Suzuki simulation remain a practical route to Hamiltonian evolution on noisy intermediate scale quantum (NISQ) hardware,

13 May 2026

Hermes Agent now runs natively on NVIDIA RTX PCs and DGX Spark. Hermes is designed for exactly the kind of always-on workload that NVIDIA's …

Local AiDGX agent

Hermes Agent now runs natively on NVIDIA RTX PCs and DGX Spark. Hermes is designed for exactly the kind of always-on workload that NVIDIA's hardware is built for, and their blog explains in depth why

Reconsidering the energy efficiency of spiking neural networks

Model ReleasesDGX agent

arXiv:2409.08290v4 Announce Type: replace-cross Abstract: Spiking Neural Networks (SNNs) promise higher energy efficiency over conventional Quantized Artificial Neural Networks (QNNs) due to their eve

RIO: Flexible Real-Time Robot I/O for Cross-Embodiment Robot Learning

SafetyDGX agent

arXiv:2605.11564v1 Announce Type: new Abstract: Despite recent efforts to collect multi-task, multi-embodiment datasets, to design recipes for training Vision-Language-Action models (VLAs), and to sho

SEVO: Semantic-Enhanced Virtual Observation for Robust VLA Manipulation via Active Illumination and Data-Centric Collection

SafetyDGX agent

arXiv:2605.11114v1 Announce Type: new Abstract: Vision-Language-Action (VLA) and imitation-learning policies trained via community toolchains on low-cost hardware frequently fail when deployed outside

12 May 2026

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

Model ReleasesDGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

Block-Wise Differentiable Sinkhorn Attention: Tail-Refinement Gradients with a Gap-Aware Dustbin Bridge

SafetyDGX agent

arXiv:2605.08123v1 Announce Type: cross Abstract: We study long-context balanced entropic optimal transport (OT) attention on TPU hardware through a stopped-base, fixed-depth tail-refinement surrogate

Constraint-Aware Diffusion Priors for High-Fidelity and Versatile Quadruped Locomotion

SafetyDGX agent

arXiv:2605.08804v1 Announce Type: new Abstract: Reinforcement learning combined with imitation learning has significantly advanced biomimetic quadrupedal locomotion. However, scaling these frameworks

CUDABeaver: Benchmarking LLM-Based Automated CUDA Debugging

Model ReleasesDGX agent

arXiv:2605.08455v1 Announce Type: new Abstract: Debugging CUDA programs has long been challenging because failures often arise from subtle interactions among hardware behavior, compiler decisions, mem

DexWrist: A Robotic Wrist for Constrained and Dynamic Manipulation

SafetyDGX agent

arXiv:2507.01008v3 Announce Type: replace Abstract: Development of dexterous manipulation hardware has primarily focused on hands and grippers. However, these end-effectors are often paired with bulky

ExecuTorch -- A Unified PyTorch Solution to Run AI Models On-Device

Local AiDGX agent

arXiv:2605.08195v1 Announce Type: new Abstract: Local execution of AI on edge devices is important for low latency and offline operation. However, deploying models on diverse hardware remains fragment

From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs

Local AiDGX agent

arXiv:2605.09370v1 Announce Type: cross Abstract: Large-scale AI training is now fundamentally a distributed systems problem, and hardware failures have become routine operating conditions rather than

Position: Life-Logging Video Streams Make the Privacy-Utility Trade-off Inevitable

ResearchDGX agent

arXiv:2605.10404v1 Announce Type: new Abstract: With the growing prevalence of always-on hardware such as smart glasses, body cameras, and home security systems, life-logging visual sensing is becomin

Rennala MVR: Improved Time Complexity for Parallel Stochastic Optimization via Momentum-Based Variance Reduction

Model ReleasesDGX agent

arXiv:2605.08871v1 Announce Type: cross Abstract: Large-scale machine learning models are trained on clusters of machines that exhibit heterogeneous performance due to hardware variability, network de

11 May 2026

A Cost-Effective and Climate-Resilient Air Pressure System for Rain Effect Reduction on Automated Vehicle Cameras

ResearchDGX agent

arXiv:2602.17472v2 Announce Type: replace Abstract: Recent advances in automated vehicles have focused on improving perception performance under adverse weather conditions; however, research on physic

Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation

SafetyDGX agent

arXiv:2605.07381v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models offer broad general capabilities, deploying them on specific hardware requires real-world adaptation to brid

Exploring CoCo Challenges in ML Engineering Teams: Insights From the Semiconductor Industry

ApplicationsDGX agent

arXiv:2605.07389v1 Announce Type: cross Abstract: The integration of machine learning (ML) into complex software systems has increased challenges in collaboration and communication (CoCo) of the teams

GTIG AI Threat Tracker: Adversaries Leverage AI for Vulnerability Exploitation, Augmented Operations, and Initial Access

Model ReleasesDGX agent

Executive Summary Since our February 2026 report on AI-related threat activity, Google Threat Intelligence Group (GTIG) has continued to track a maturing transition from nascent AI-enabled operations

9 May 2026

b9084

Local AiDGX agent

b9084 is a release build of llama.cpp, an open-source C/C++ library for LLM inference that enables running large language models on consumer hardware. llama.cpp is a software library that performs inf

8 May 2026

b9079

Local AiDGX agent

llama.cpp is an LLM inference implementation in C/C++ that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud . Build b9079

CyberSecQwen-4B: Why Defensive Cyber Needs Small, Specialized, Locally-Runnable Models

ToolsDGX agent

CyberSecQwen-4B is a specialized 4-billion parameter language model designed for cybersecurity defense tasks that can run locally on standard hardware. The model addresses the need for small, efficien

7 May 2026

b9055

Local AiDGX agent

B9055 is the latest version of llama.cpp, released on May 7, 2026. Llama.cpp is an LLM inference framework implemented in C/C++ that enables efficient large language model execution with broad hardwar

b9056

Local AiDGX agent

b9056 is a release of llama.cpp, a C/C++ implementation for LLM inference . The project enables users to run LLaMA models on consumer hardware without expensive GPUs or cloud infrastructure . This rel

Constraint-Aware Execution Planning for Hybrid Space-Ground Compute Workloads

TutorialsDGX agent

arXiv:2605.04052v1 Announce Type: cross Abstract: Low Earth orbit (LEO) satellites increasingly carry compute hardware capable of on-board processing, yet each satellite generates roughly two orders o

KernelBench-X: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernels

Model ReleasesDGX agent

arXiv:2605.04956v1 Announce Type: new Abstract: LLM-based Triton kernel generation has attracted significant interest, yet a fundamental empirical question remains unanswered: where does this capabili

6 May 2026

A Target-Free Harmonization Method for MRI

ResearchDGX agent

arXiv:2605.01282v1 Announce Type: cross Abstract: In MRI, variations in scan parameters, sequence, or hardware can lead to discrepancies in image appearance, even for the same subject. These inconsist

AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI

ApplicationsDGX agent

arXiv:2503.10070v2 Announce Type: replace-cross Abstract: Scaling Vision-Language-Action models for embodied manipulation demands large volumes of diverse manipulation data, yet the high cost of comme

Optimization of CV-QKD Under Practical Constraints

ResearchDGX agent

arXiv:2605.02045v1 Announce Type: cross Abstract: Using reinforcement learning, we optimize for practical hardware constraints, including limited FIR filter taps at the transmitter and receiver, mean

RFPrompt: Prompt-Based Expert Adaptation of the Large Wireless Model for Modulation Classification

Model ReleasesDGX agent

arXiv:2605.03279v1 Announce Type: new Abstract: Automatic modulation classification (AMC) in real-world deployments demands robustness to distribution shifts arising from hardware impairments, unseen

SpecMD: A Comprehensive Study on Speculative Expert Prefetching

ResearchDGX agent

Mixture-of-Experts (MoE) models enable sparse expert activation, meaning that only a subset of the model’s parameters is used during each inference. However, to translate this sparsity into practical

What is sample-based quantum diagonalization?

ResearchDGX agent

Sample-based quantum diagonalization is a quantum algorithm technique that estimates the eigenvalues and eigenvectors of matrices using measurements from quantum hardware, rather than requiring full q

5 May 2026

OpenAI is reportedly launching a phone for ChatGPT

ApplicationsDGX agent

OpenAI's first hardware product might be a phone instead of a mysterious Jony Ive gadget. As reported by MacRumors, supply chain analyst Ming-Chi Kuo shared details about the rumored phone, claiming O

QuantWare raises $178M to become the TSMC of quantum computing

IndustryDGX agent

Netherlands-based quantum computing hardware company QuantWare B.V. said today it has raised 178 million in funding to accelerate its plans of creating an “open architecture” for the industry. The rou

4 May 2026

b9018

Local AiDGX agent

B9018 is the latest version of llama.cpp released on May 4, 2026. Llama.cpp is a project for LLM inference in C/C++ , providing efficient large language model execution with broad hardware support inc

Scalable Learning in Structured Recurrent Spiking Neural Networks without Backpropagation

Model ReleasesDGX agent

arXiv:2605.00402v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) provide a promising framework for energy-efficient and biologically grounded computation; however, scalable learning in

2 May 2026

b9002

Local AiDGX agent

B9002 is a build release of llama.cpp, a C/C++ implementation for LLM inference. The release includes compiled binaries and artifacts for multiple platforms and hardware configurations, as part of the

1 May 2026

b8995

Local AiDGX agent

llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Release b8995 is a build/versio

RuC: HDL-Agnostic Rule Completion Benchmark Generation

Model ReleasesDGX agent

arXiv:2604.27780v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RT

30 Apr 2026

Hermes Agent + LM Studio 👾🪽

AgentsDGX agent

Hermes Agent + LM Studio 👾🪽 LM Studio is the most popular way to run open-source LLMs on your own hardware. Your Hermes Agent now runs natively on @lmstudio: auto-discovering your models, loading them

New laptop for running ollama locally

Local AiDGX agent

This Reddit thread likely discusses hardware recommendations and considerations for purchasing a new laptop capable of running Ollama, a tool that enables running large language models locally. The di

Super-resolution Multi-signal Direction-of-Arrival Estimation by Hankel-structured Sensing and Decomposition

AgentsDGX agent

arXiv:2604.26793v1 Announce Type: new Abstract: Motivated by sensing modalities in modern autonomous systems that involve hardware-constrained spatial sampling over large arrays with limited coherence

29 Apr 2026

b8978

Local AiDGX agent

b8978 is a release of llama.cpp, a project designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. The llama.cpp project uses sequential build

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

Model ReleasesDGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

TEACar: An Open-Source Autonomous Driving Platform

AgentsDGX agent

arXiv:2604.24934v1 Announce Type: new Abstract: Intelligent Transportation Systems (ITS) increasingly rely on vision-based perception and learning-based control, necessitating experimental platforms t

← Previous
1…3839404142…75
Next →