AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
28 Apr 2026

An Affordable,Wearable Stereo-Eye-Tracking Platform

ResearchDGX agent

arXiv:2604.24331v1 Announce Type: new Abstract: Research on video-based eye-tracking has long explored stereo and glint-based methods, yet existing wearable eye trackers - both commercial and open-sou

b8960

Local AiDGX agent

B8960 is a build release of llama.cpp, an open-source C/C++ project for running large language models locally with minimal setup and optimized performance across various hardware platforms. It is one

b8963

Local AiDGX agent

B8963 is a build release of llama.cpp, the C/C++ implementation of LLM inference. Llama.cpp is an open-source project that enables efficient language model execution on consumer hardware with minimal

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ResearchDGX agent

arXiv:2604.24447v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under t

ECoLAD: Deployment-Oriented Evaluation for Automotive Time-Series Anomaly Detection

ResearchDGX agent

arXiv:2603.10926v1 Announce Type: cross Abstract: Time-series anomaly detectors are commonly compared on workstation-class hardware under unconstrained execution. In-vehicle monitoring, however, requi

Self-Organising Memristive Networks as Physical Learning Systems

AgentsDGX agent

arXiv:2509.00747v2 Announce Type: replace-cross Abstract: Learning with physical systems is an emerging paradigm that seeks to harness the intrinsic nonlinear dynamics of physical substrates for learn

Tessera: Secure, Near-Line-Rate Weight Streaming for UMA Edge Accelerators

ResearchDGX agent

arXiv:2604.23205v1 Announce Type: cross Abstract: Deploying proprietary Deep Neural Networks (DNNs) on commodity edge devices demands hardware-backed Digital Rights Management (DRM) capable of withsta

27 Apr 2026

Running Qwen3.5-397B-A17B (4bit quants, 177 GB) on two DGX Sparks using llama.cpp with RPC and RDMA:

Model ReleasesDGX agent

This post documents a technical demonstration of running the large Qwen3.5-397B-A17B model across distributed hardware using llama.cpp with advanced networking protocols. The approach leverages 4-bit

26 Apr 2026

b8934

Local AiDGX agent

Release b8934 of llama.cpp was released on April 26, 2026, and includes improvements to hexagon hardware support, specifically guarding HMX clock requests for v75+ platforms. The release provides bina

25 Apr 2026

Anyone got DeepSeek-V4-Flash running on a Mac yet? 512GB or 256GB or 128GB or smaller?

Model ReleasesDGX agent

Simon Willison inquires about running DeepSeek-V4-Flash on Mac hardware, specifically asking about feasibility across different RAM configurations from 512GB down to smaller amounts. This reflects dis

b8931

Local AiDGX agent

b8931 is a release version of llama.cpp published on April 25, 2026 . llama.cpp is an LLM inference framework in C/C++ that enables running large language models efficiently on various hardware. The r

http://reddit.com/r/LocalLLaMA

ResearchDGX agent

r/LocalLLaMA is a subreddit community dedicated to discussing and sharing resources about running large language models locally on personal computers, covering topics like model optimization, hardware

Why there is no cloud version for Qwen 3.6 27/35B?

Model ReleasesDGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

24 Apr 2026

b8924

Local AiDGX agent

B8924 is a release build number from the llama.cpp project, a C/C++ framework for efficient large language model inference on consumer hardware. The llama.cpp project publishes multiple releases in a

Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots o…

Model ReleasesDGX agent

Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots of models to buy/use, but very few hardware options upon whic

ImageHD: Energy-Efficient On-Device Continual Learning of Visual Representations via Hyperdimensional Computing

Local AiDGX agent

arXiv:2604.21280v1 Announce Type: new Abstract: On-device continual learning (CL) is critical for edge AI systems operating on non-stationary data streams, but most existing methods rely on backpropag

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

Model ReleasesDGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference

Local AiDGX agent

arXiv:2604.21231v1 Announce Type: cross Abstract: Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill

23 Apr 2026

b8902

Local AiDGX agent

B8902 is a release build of llama.cpp, an open-source C/C++ framework for running large language model inference on consumer hardware with minimal dependencies. As an intermediate build release from t

FlashNorm: Fast Normalization for Transformers

Model ReleasesDGX agent

arXiv:2407.09577v4 Announce Type: replace Abstract: Normalization layers are ubiquitous in large language models (LLMs) yet represent a compute bottleneck: on hardware with distinct vector and matrix

Option Pricing on Noisy Intermediate-Scale Quantum Computers: A Quantum Neural Network Approach

Model ReleasesDGX agent

arXiv:2604.19832v1 Announce Type: cross Abstract: In a global derivatives market with notional values in the hundreds of trillions of dollars, the accuracy and efficiency of pricing models are of fund

we've quantized kimi-k2.6 to mxfp4 on amd! download and use today! @AIatAMD

IndustryDGX agent

Hugging Face and AMD have collaborated to quantize the Kimi-K2.6 model to MXFP4 format, enabling efficient inference on AMD hardware. MXFP4 is a low-precision floating-point quantization technique tha

22 Apr 2026

b8882

Local AiDGX agent

b8882 is a release of llama.cpp, a C/C++ implementation for LLM inference . The project follows a rapid release cycle with frequent updates to support multiple hardware architectures and platforms. Th

b8884

Local AiDGX agent

llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Release b8884 is a build/versio

CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmark

Model ReleasesDGX agent

arXiv:2505.16968v4 Announce Type: replace-cross Abstract: Cross-architecture GPU code transpilation is essential for unlocking low-level hardware portability, yet no scalable solution exists. We intro

Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence

ResearchDGX agent

arXiv:2602.12851v3 Announce Type: replace-cross Abstract: Deploying expressive learning models directly on programmable dataplanes promises line-rate, low-latency traffic analysis but remains hindered

Flux 2-Klein-9B NVFP4 works well on my RTX 3050, but it takes 55sec to generate 1024 resolution.

Local AiDGX agent

Flux 2-Klein-9B NVFP4 is a quantized image generation model that runs on mid-range GPUs like the RTX 3050. On this hardware, the model produces 1024-resolution images but with relatively slow inferenc

IoT in Manufacturing: Strategy, Components, Use Cases, and Challenges

ApplicationsDGX agent

This article explores the implementation of Internet of Things technology in manufacturing environments, covering strategic approaches, key hardware and software components, practical applications acr

New innovations in Google Distributed Cloud

Model ReleasesDGX agent

Today at Google Cloud Next, we’re announcing new capabilities in Google Distributed Cloud (GDC) that bring Gemini and our advanced AI stack to wherever your data is, so you don’t need to compromise be

21 Apr 2026

Adaptive Quantized Planetary Crater Detection System for Autonomous Space Exploration

AgentsDGX agent

arXiv:2508.18025v4 Announce Type: replace-cross Abstract: Autonomous planetary exploration demands real-time, high-fidelity environmental perception. Standard deep learning models require massive comp

Fully Analog Resonant Recurrent Neural Network via Metacircuit

Local AiDGX agent

arXiv:2604.17277v1 Announce Type: new Abstract: Physical neural networks offer a transformative route to edge intelligence, providing superior inference speed and energy efficiency compared to convent

M100: An Orchestrated Dataflow Architecture Powering General AI Computing

Model ReleasesDGX agent

arXiv:2604.17862v1 Announce Type: new Abstract: As deep learning-based AI technologies gain momentum, the demand for general-purpose AI computing architectures continues to grow. While GPGPU-based arc

Trustworthy Endoscopic Super-Resolution

Local AiDGX agent

arXiv:2604.18001v1 Announce Type: new Abstract: Super-resolution (SR) models are attracting growing interest for enhancing minimally invasive surgery and diagnostic videos under hardware constraints.

20 Apr 2026

b8859

Local AiDGX agent

Release b8859 is a version of llama.cpp, an open-source project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. The llama.cpp project uses a

Contact-Aware Planning and Control of Continuum Robots in Highly Constrained Environments

SafetyDGX agent

arXiv:2604.15638v1 Announce Type: new Abstract: Continuum robots are well suited for navigating confined and fragile environments, such as vascular or endoluminal anatomy, where contact with surroundi

Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation

Model ReleasesDGX agent

arXiv:2604.15388v1 Announce Type: cross Abstract: Recent advances in large language models have improved code generation, but their use in hardware description languages is still limited. Moreover, tr

What's new at IBM Quantum Q1 2026

ResearchDGX agent

IBM Quantum's Q1 2026 updates likely cover recent advances in quantum computing hardware, software, and applications, including announcements about processor improvements, new quantum algorithms, or e

18 Apr 2026

b8840

Local AiDGX agent

b8840 is a release version of llama.cpp, a C/C++ implementation for large language model inference. This release build includes compiled binaries and updates for various platforms and hardware acceler

@openclaw And of course @Ollama for the local model-serving engine. 🦙

Local AiDGX agent

Ollama is a local model-serving engine that enables users to run large language models on their own hardware without relying on cloud services. The post appears to highlight Ollama's integration with

17 Apr 2026

After a saga of broken promises, a European rover finally has a ride to Mars

IndustryDGX agent

The European Space Agency's Rosalind Franklin Mars rover finally secured its launch mission when NASA approved the ROSA project in April 2026, providing launch vehicle and hardware support to the ESA-

b8832

Local AiDGX agent

b8832 is a recent build release of llama.cpp from April 17, 2026 . Llama.cpp is a C/C++ implementation for running large language model inference efficiently on consumer hardware with minimal setup. T

Prism: Symbolic Superoptimization of Tensor Programs

Model ReleasesDGX agent

arXiv:2604.15272v1 Announce Type: cross Abstract: This paper presents Prism, the first symbolic superoptimizer for tensor programs. The key idea is sGraph, a symbolic, hierarchical representation that

16 Apr 2026

Analog Optical Inference on Million-Record Mortgage Data

Model ReleasesDGX agent

arXiv:2604.13251v1 Announce Type: new Abstract: Analog optical computers promise large efficiency gains for machine learning inference, yet no demonstration has moved beyond small-scale image benchmar

b8814

Local AiDGX agent

llama.cpp is a C/C++ library for LLM inference that enables running large language models on consumer hardware. Release b8814 is a specific version in the project's continuous release cycle, which fol

15 Apr 2026

A Dataset and Evaluation for Complex 4D Markerless Human Motion Capture

SafetyDGX agent

arXiv:2604.12765v1 Announce Type: new Abstract: Marker-based motion capture (MoCap) systems have long been the gold standard for accurate 4D human modeling, yet their reliance on specialized hardware

ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

Model ReleasesDGX agent

arXiv:2602.11236v2 Announce Type: replace-cross Abstract: Building general-purpose embodied agents across diverse hardware remains a central challenge in robotics, often framed as the ''one-brain, man

Accelerating decode-heavy LLM inference with speculative decoding on AWS Trainium and vLLM

TutorialsDGX agent

Speculative decoding is a technique used to accelerate the slow, sequential token generation (decode stage) in LLM inference. This method significantly reduces latency and improves hardware utilizatio

Best realism model under 16GB VRAM

Local AiDGX agent

This r/StableDiffusion Reddit thread discusses community recommendations for photorealistic image generation models that can run within a 16GB VRAM constraint, a common hardware limit for consumer GPU

14 Apr 2026

b8790

Local AiDGX agent

Build b8790 is an incremental automated release of llama.cpp, the open-source C/C++ library for efficient LLM inference on local hardware. Like other builds in the project's continuous release cycle,

Problem Reductions at Scale: Agentic Integration of Computationally Hard Problems

AgentsDGX agent

arXiv:2604.11535v1 Announce Type: new Abstract: Solving an NP-hard optimization problem often requires reformulating it for a specific solver -- quantum hardware, a commercial optimizer, or a domain h

Quantum computing moves into real-world workflows

ApplicationsDGX agent

As the fifth annual World Quantum Day gets underway on April 14, researchers say the focus of their efforts is shifting from hardware experimentation to integrating quantum systems into broader comput

Three Roles, One Model: Role Orchestration at Inference Time to Close the Performance Gap Between Small and Large Agents

Model ReleasesDGX agent

arXiv:2604.11465v1 Announce Type: new Abstract: Large language model (LLM) agents show promise on realistic tool-use tasks, but deploying capable agents on modest hardware remains challenging. We stud

13 Apr 2026

ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning

SafetyDGX agent

arXiv:2507.04736v2 Announce Type: replace Abstract: Large Language Models have emerged as powerful tools for automating Register-Transfer Level (RTL) code generation, yet they face critical limitation

LogLens with local AI “Ollama”

Local AiDGX agent

LogLens with local AI 'Ollama' is a Reddit post on r/ollama discussing the integration of LogLens — a log analysis tool — with Ollama to perform AI-powered log inspection entirely on local hardware, w

Starlink Mini enables reliable high-speed internet on the go 🛰️🛣️

IndustryDGX agent

Starlink Mini is a compact, portable version of SpaceX's Starlink satellite internet hardware designed for mobile and on-the-go connectivity. It offers high-speed internet access in a smaller, more tr

Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures

Model ReleasesDGX agent

arXiv:2604.09048v1 Announce Type: cross Abstract: While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient

12 Apr 2026

Has anyone actually gotten a reliable local AI system running?

Local AiDGX agent

This r/ollama Reddit thread addresses a common question among the local AI community about whether reliable, self-hosted AI systems are genuinely achievable. Community members in this space typically

KIV: 1M token context window on a RTX 4070 (12GB VRAM), no retraining, drop-in HuggingFace cache replacement - Works with any model that uses DynamicCache [P]

Model ReleasesDGX agent

KIV is a project shared on r/MachineLearning presenting a drop-in replacement for HuggingFace's `DynamicCache` that enables up to 1 million token context windows on consumer hardware with only 12GB of

11 Apr 2026

ace step 1.5 xl sft terrible results

Local AiDGX agent

A Reddit thread on r/StableDiffusion where a user reports poor output quality when using the ACE-Step 1.5 XL SFT model variant for AI music generation. The SFT (Supervised Fine-Tuning) variant of ACE-

b8752

Local AiDGX agent

llama.cpp release **b8752** is a tagged build of the [ggml-org/llama.cpp](https://github.com/ggml-org/llama.cpp) project — a C/C++ framework for efficient LLM inference on a wide range of hardware....

← Previous
1…3940414243…75
Next →