AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,479 results
Local Ai

b8978

DGX agent

b8978 is a release of llama.cpp, a project designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. The llama.cpp project uses sequential build

local-aillama-cpp-releases
29 Apr 2026
Model Releases

From Local to Global: Revisiting Structured Pruning Paradigms for Large Language Models

DGX agent

arXiv:2510.18030v2 Announce Type: replace Abstract: Structured pruning is a practical approach to deploying large language models (LLMs) efficiently, as it yields compact, hardware-friendly architectu

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
model-releasesarxiv-cs-cl
29 Apr 2026
Agents

TEACar: An Open-Source Autonomous Driving Platform

DGX agent

arXiv:2604.24934v1 Announce Type: new Abstract: Intelligent Transportation Systems (ITS) increasingly rely on vision-based perception and learning-based control, necessitating experimental platforms t

agentsarxiv-cs-ro
29 Apr 2026
Research

An Affordable,Wearable Stereo-Eye-Tracking Platform

DGX agent

arXiv:2604.24331v1 Announce Type: new Abstract: Research on video-based eye-tracking has long explored stereo and glint-based methods, yet existing wearable eye trackers - both commercial and open-sou

researcharxiv-cs-cv
28 Apr 2026
Local Ai

b8960

DGX agent

B8960 is a build release of llama.cpp, an open-source C/C++ project for running large language models locally with minimal setup and optimized performance across various hardware platforms. It is one

local-aillama-cpp-releases
28 Apr 2026
Local Ai

b8963

DGX agent

B8963 is a build release of llama.cpp, the C/C++ implementation of LLM inference. Llama.cpp is an open-source project that enables efficient language model execution on consumer hardware with minimal

local-aillama-cpp-releases
28 Apr 2026
Research

Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment

DGX agent

arXiv:2604.24447v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are promising for generalist robot control, but on-robot deployment is bottlenecked by real-time inference under t

researcharxiv-cs-ai
28 Apr 2026
Research

ECoLAD: Deployment-Oriented Evaluation for Automotive Time-Series Anomaly Detection

DGX agent

arXiv:2603.10926v1 Announce Type: cross Abstract: Time-series anomaly detectors are commonly compared on workstation-class hardware under unconstrained execution. In-vehicle monitoring, however, requi

researcharxiv-cs-ai
28 Apr 2026
Agents

Self-Organising Memristive Networks as Physical Learning Systems

DGX agent

arXiv:2509.00747v2 Announce Type: replace-cross Abstract: Learning with physical systems is an emerging paradigm that seeks to harness the intrinsic nonlinear dynamics of physical substrates for learn

agentsarxiv-cs-lg
28 Apr 2026
Research

Tessera: Secure, Near-Line-Rate Weight Streaming for UMA Edge Accelerators

DGX agent

arXiv:2604.23205v1 Announce Type: cross Abstract: Deploying proprietary Deep Neural Networks (DNNs) on commodity edge devices demands hardware-backed Digital Rights Management (DRM) capable of withsta

researcharxiv-cs-lg
28 Apr 2026
Model Releases

Running Qwen3.5-397B-A17B (4bit quants, 177 GB) on two DGX Sparks using llama.cpp with RPC and RDMA:

DGX agent

This post documents a technical demonstration of running the large Qwen3.5-397B-A17B model across distributed hardware using llama.cpp with advanced networking protocols. The approach leverages 4-bit

model-releasesgeorgi-gerganov--x
27 Apr 2026
Local Ai

b8934

DGX agent

Release b8934 of llama.cpp was released on April 26, 2026, and includes improvements to hexagon hardware support, specifically guarding HMX clock requests for v75+ platforms. The release provides bina

local-aillama-cpp-releases
26 Apr 2026
Model Releases

Anyone got DeepSeek-V4-Flash running on a Mac yet? 512GB or 256GB or 128GB or smaller?

DGX agent

Simon Willison inquires about running DeepSeek-V4-Flash on Mac hardware, specifically asking about feasibility across different RAM configurations from 512GB down to smaller amounts. This reflects dis

model-releasessimon-willison--x
25 Apr 2026
Local Ai

b8931

DGX agent

b8931 is a release version of llama.cpp published on April 25, 2026 . llama.cpp is an LLM inference framework in C/C++ that enables running large language models efficiently on various hardware. The r

local-aillama-cpp-releases
25 Apr 2026
Research

http://reddit.com/r/LocalLLaMA

DGX agent

r/LocalLLaMA is a subreddit community dedicated to discussing and sharing resources about running large language models locally on personal computers, covering topics like model optimization, hardware

researchnous-research--x
25 Apr 2026
Model Releases

Why there is no cloud version for Qwen 3.6 27/35B?

DGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

model-releasesr-ollama
25 Apr 2026
Local Ai

b8924

DGX agent

B8924 is a release build number from the llama.cpp project, a C/C++ framework for efficient large language model inference on consumer hardware. The llama.cpp project publishes multiple releases in a

local-aillama-cpp-releases
24 Apr 2026
Model Releases

Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots o…

DGX agent

Cloud-hosted models will become more niche/seldom-used as much cheaper local models fulfill that which 90% of consumers need from AI. Lots of models to buy/use, but very few hardware options upon whic

model-releasesclem-delangue--x
24 Apr 2026
Local Ai

ImageHD: Energy-Efficient On-Device Continual Learning of Visual Representations via Hyperdimensional Computing

DGX agent

arXiv:2604.21280v1 Announce Type: new Abstract: On-device continual learning (CL) is critical for edge AI systems operating on non-stationary data streams, but most existing methods rely on backpropag

local-aiarxiv-cs-cv
24 Apr 2026
Model Releases

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

DGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

model-releasesarxiv-cs-lg
24 Apr 2026
Local Ai

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference

DGX agent

arXiv:2604.21231v1 Announce Type: cross Abstract: Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill

local-aiarxiv-cs-ai
24 Apr 2026
Local Ai

b8902

DGX agent

B8902 is a release build of llama.cpp, an open-source C/C++ framework for running large language model inference on consumer hardware with minimal dependencies. As an intermediate build release from t

local-aillama-cpp-releases
23 Apr 2026
Model Releases

FlashNorm: Fast Normalization for Transformers

DGX agent

arXiv:2407.09577v4 Announce Type: replace Abstract: Normalization layers are ubiquitous in large language models (LLMs) yet represent a compute bottleneck: on hardware with distinct vector and matrix

model-releasesarxiv-cs-lg
23 Apr 2026
Model Releases

Option Pricing on Noisy Intermediate-Scale Quantum Computers: A Quantum Neural Network Approach

DGX agent

arXiv:2604.19832v1 Announce Type: cross Abstract: In a global derivatives market with notional values in the hundreds of trillions of dollars, the accuracy and efficiency of pricing models are of fund

model-releasesarxiv-cs-lg
23 Apr 2026
Industry

we've quantized kimi-k2.6 to mxfp4 on amd! download and use today! @AIatAMD

DGX agent

Hugging Face and AMD have collaborated to quantize the Kimi-K2.6 model to MXFP4 format, enabling efficient inference on AMD hardware. MXFP4 is a low-precision floating-point quantization technique tha

industryclem-delangue--x
23 Apr 2026
Local Ai

b8882

DGX agent

b8882 is a release of llama.cpp, a C/C++ implementation for LLM inference . The project follows a rapid release cycle with frequent updates to support multiple hardware architectures and platforms. Th

local-aillama-cpp-releases
22 Apr 2026
Local Ai

b8884

DGX agent

llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Release b8884 is a build/versio

local-aillama-cpp-releases
22 Apr 2026
Model Releases

CASS: Nvidia to AMD Transpilation with Data, Models, and Benchmark

DGX agent

arXiv:2505.16968v4 Announce Type: replace-cross Abstract: Cross-architecture GPU code transpilation is essential for unlocking low-level hardware portability, yet no scalable solution exists. We intro

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Chimera: Neuro-Symbolic Attention Primitives for Trustworthy Dataplane Intelligence

DGX agent

arXiv:2602.12851v3 Announce Type: replace-cross Abstract: Deploying expressive learning models directly on programmable dataplanes promises line-rate, low-latency traffic analysis but remains hindered

researcharxiv-cs-ai
22 Apr 2026
Local Ai

Flux 2-Klein-9B NVFP4 works well on my RTX 3050, but it takes 55sec to generate 1024 resolution.

DGX agent

Flux 2-Klein-9B NVFP4 is a quantized image generation model that runs on mid-range GPUs like the RTX 3050. On this hardware, the model produces 1024-resolution images but with relatively slow inferenc

local-air-stablediffusion
22 Apr 2026
Applications

IoT in Manufacturing: Strategy, Components, Use Cases, and Challenges

DGX agent

This article explores the implementation of Internet of Things technology in manufacturing environments, covering strategic approaches, key hardware and software components, practical applications acr

applicationsdatabricks
22 Apr 2026
Model Releases

New innovations in Google Distributed Cloud

DGX agent

Today at Google Cloud Next, we’re announcing new capabilities in Google Distributed Cloud (GDC) that bring Gemini and our advanced AI stack to wherever your data is, so you don’t need to compromise be

model-releasesgoogle-cloud-ai
22 Apr 2026
Agents

Adaptive Quantized Planetary Crater Detection System for Autonomous Space Exploration

DGX agent

arXiv:2508.18025v4 Announce Type: replace-cross Abstract: Autonomous planetary exploration demands real-time, high-fidelity environmental perception. Standard deep learning models require massive comp

agentsarxiv-cs-cv
21 Apr 2026
Local Ai

Fully Analog Resonant Recurrent Neural Network via Metacircuit

DGX agent

arXiv:2604.17277v1 Announce Type: new Abstract: Physical neural networks offer a transformative route to edge intelligence, providing superior inference speed and energy efficiency compared to convent

local-aiarxiv-cs-lg
21 Apr 2026
Model Releases

M100: An Orchestrated Dataflow Architecture Powering General AI Computing

DGX agent

arXiv:2604.17862v1 Announce Type: new Abstract: As deep learning-based AI technologies gain momentum, the demand for general-purpose AI computing architectures continues to grow. While GPGPU-based arc

model-releasesarxiv-cs-lg
21 Apr 2026
Local Ai

Trustworthy Endoscopic Super-Resolution

DGX agent

arXiv:2604.18001v1 Announce Type: new Abstract: Super-resolution (SR) models are attracting growing interest for enhancing minimally invasive surgery and diagnostic videos under hardware constraints.

local-aiarxiv-cs-cv
21 Apr 2026
Local Ai

b8859

DGX agent

Release b8859 is a version of llama.cpp, an open-source project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. The llama.cpp project uses a

local-aillama-cpp-releases
20 Apr 2026
Safety

Contact-Aware Planning and Control of Continuum Robots in Highly Constrained Environments

DGX agent

arXiv:2604.15638v1 Announce Type: new Abstract: Continuum robots are well suited for navigating confined and fragile environments, such as vascular or endoluminal anatomy, where contact with surroundi

safetyarxiv-cs-ro
20 Apr 2026
Model Releases

Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation

DGX agent

arXiv:2604.15388v1 Announce Type: cross Abstract: Recent advances in large language models have improved code generation, but their use in hardware description languages is still limited. Moreover, tr

model-releasesarxiv-cs-ai
20 Apr 2026
Research

What's new at IBM Quantum Q1 2026

DGX agent

IBM Quantum's Q1 2026 updates likely cover recent advances in quantum computing hardware, software, and applications, including announcements about processor improvements, new quantum algorithms, or e

researchibm-research
20 Apr 2026
Local Ai

b8840

DGX agent

b8840 is a release version of llama.cpp, a C/C++ implementation for large language model inference. This release build includes compiled binaries and updates for various platforms and hardware acceler

local-aillama-cpp-releases
18 Apr 2026
Local Ai

@openclaw And of course @Ollama for the local model-serving engine. 🦙

DGX agent

Ollama is a local model-serving engine that enables users to run large language models on their own hardware without relying on cloud services. The post appears to highlight Ollama's integration with

local-aiollama--x
18 Apr 2026
Industry

After a saga of broken promises, a European rover finally has a ride to Mars

DGX agent

The European Space Agency's Rosalind Franklin Mars rover finally secured its launch mission when NASA approved the ROSA project in April 2026, providing launch vehicle and hardware support to the ESA-

industryars-technica
17 Apr 2026
Local Ai

b8832

DGX agent

b8832 is a recent build release of llama.cpp from April 17, 2026 . Llama.cpp is a C/C++ implementation for running large language model inference efficiently on consumer hardware with minimal setup. T

local-aillama-cpp-releases
17 Apr 2026
Model Releases

Prism: Symbolic Superoptimization of Tensor Programs

DGX agent

arXiv:2604.15272v1 Announce Type: cross Abstract: This paper presents Prism, the first symbolic superoptimizer for tensor programs. The key idea is sGraph, a symbolic, hierarchical representation that

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Analog Optical Inference on Million-Record Mortgage Data

DGX agent

arXiv:2604.13251v1 Announce Type: new Abstract: Analog optical computers promise large efficiency gains for machine learning inference, yet no demonstration has moved beyond small-scale image benchmar

model-releasesarxiv-cs-lg
16 Apr 2026
Local Ai

b8814

DGX agent

llama.cpp is a C/C++ library for LLM inference that enables running large language models on consumer hardware. Release b8814 is a specific version in the project's continuous release cycle, which fol

local-aillama-cpp-releases
16 Apr 2026
Safety

A Dataset and Evaluation for Complex 4D Markerless Human Motion Capture

DGX agent

arXiv:2604.12765v1 Announce Type: new Abstract: Marker-based motion capture (MoCap) systems have long been the gold standard for accurate 4D human modeling, yet their reliance on specialized hardware

safetyarxiv-cs-cv
15 Apr 2026
← Previous
1…4950515253…94
Next →