AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,479 results
Model Releases

Apertus LLM Family Expansion via Distillation and Quantization

DGX agent

arXiv:2605.29128v1 Announce Type: new Abstract: The wide adoption of LLMs has led to their use in great variety of applications and scenarios, such as chatbot assistants and data annotation, creating

model-releasesarxiv-cs-lg
29 May 2026
Local Ai

b9410

DGX agent

b9410 is a release of llama.cpp, a C/C++ project that enables LLM inference with minimal setup and state-of-the-art performance on various hardware platforms. This build identifier represents a specif

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
local-aillama-cpp-releases
29 May 2026
Research

EMAG: Differentiable 4D Gaussian Mixture Splatting for EEG Spatial Super-Resolution

DGX agent

arXiv:2605.29731v1 Announce Type: new Abstract: High-density electroencephalography (HD-EEG) enables fine-grained measurement of cortical activity but requires expensive hardware and lengthy setup tim

researcharxiv-cs-lg
29 May 2026
Research

Fingerprinting Inference Systems of Large Language Models

DGX agent

arXiv:2605.29979v1 Announce Type: cross Abstract: The behavior of LLMs does not depend solely on the model itself. Components of the inference system, such as the inference engine, attention backend,

researcharxiv-cs-lg
29 May 2026
Agents

Human-in-the-Loop Swarms: A Bionic Swarm Approach to Real-World Soil Mapping

DGX agent

arXiv:2605.29091v1 Announce Type: new Abstract: Swarm and field robotics face significant barriers to real-world validation due to the high cost and development time to deploy hardware. This paper int

agentsarxiv-cs-ro
29 May 2026
Model Releases

Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives

DGX agent

arXiv:2505.21627v4 Announce Type: replace-cross Abstract: State-of-the-art large language models require specialized hardware and substantial energy to operate. As a consequence, cloud-based services

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Just dropped 🧑‍🍳 Fixed version of DeepSeek-V4-Pro-NVFP4 by @NVIDIAAI https://huggingface.co/nvidia/DeepSeek-V4-Pro-NVFP4

DGX agent

NVIDIA has released a corrected version of DeepSeek-V4-Pro-NVFP4, a quantized model variant optimized for NVIDIA hardware using NV-FP4 (NVIDIA's 4-bit floating-point format). The model is available on

model-releasesclem-delangue--x
29 May 2026
Model Releases

Ariel-ML: Computing Parallelization with Embedded Rust for Neural Networks on Heterogeneous Multi-core Microcontrollers

DGX agent

arXiv:2512.09800v2 Announce Type: replace Abstract: Low-power microcontroller (MCU) hardware is currently evolving from single-core architectures to predominantly multi-core architectures. In parallel

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

Bayesian Optimization Parameter Tuning Framework for a Lyapunov Based Path Following Controller

DGX agent

arXiv:2512.12649v2 Announce Type: replace Abstract: Parameter tuning in real-world experiments is constrained by the limited evaluation budget available on hardware. The path-following controller stud

model-releasesarxiv-cs-ro
28 May 2026
Research

Can Hallucinations Be Useful? Solving Multi-Hop Questions With SLMs By Chaining System-I/II Reasoning

DGX agent

arXiv:2605.27596v1 Announce Type: new Abstract: Recently, there has been increased interest in Small Language Models (SLMs), which are fast, show good performance, and have lower hardware demands than

researcharxiv-cs-cl
28 May 2026
Model Releases

Mitigating Staleness in Asynchronous Pipeline Parallelism via Basis Rotation

DGX agent

arXiv:2602.03515v2 Announce Type: replace-cross Abstract: Asynchronous pipeline parallelism maximizes hardware utilization by eliminating the pipeline bubbles inherent in synchronous execution, offeri

model-releasesarxiv-cs-ai
28 May 2026
Research

TinyDejaVu: Smaller RAM and Faster Inference with Neural Networks on MCUs for Sensor Data Streams

DGX agent

arXiv:2512.09786v2 Announce Type: replace Abstract: Examples of embedded intelligence include a wide variety of tiny neural networks used on-board wireless sensors and actuators, which are expected to

researcharxiv-cs-lg
28 May 2026
Local Ai

b9368

DGX agent

b9368 is an intermediate build release of llama.cpp, a C/C++ implementation for efficient LLM inference on consumer hardware. The release follows the project's rapid development cycle where multiple b

local-aillama-cpp-releases
27 May 2026
Safety

Cross-Receiver Generalization for RF Fingerprint Identification via Feature Disentanglement and Adversarial Training

DGX agent

arXiv:2510.09405v2 Announce Type: replace Abstract: Radio frequency fingerprint identification (RFFI) is a key technique for wireless network security, leveraging intrinsic hardware imperfections to e

safetyarxiv-cs-lg
27 May 2026
Industry

Motorola's 2026 Razrs are almost worth buying just for their stunning looks… almost

DGX agent

Motorola's 2026 Razr lineup features durable and visually refined hardware design , but critics note the devices struggle to establish clear leadership in any single category . Motorola raised prices

industryars-technica
27 May 2026
Research

Private analytics via zero-trust aggregation

DGX agent

Google Research presents a zero-trust approach to private analytics that combines cryptographic and hardware protection mechanisms to reduce trust in any single entity. The solution uses a cryptograph

researchgoogle-research
27 May 2026
Safety

The Rescue Effect: Spatio-Semantic Early Exit Bypasses Quantization Collapse in CLIP

DGX agent

arXiv:2605.26415v1 Announce Type: cross Abstract: Deploying Vision-Language Models on resource-constrained hardware typically requires INT8 quantization, but in joint-embedding architectures such as C

safetyarxiv-cs-ai
27 May 2026
Local Ai

b9320

DGX agent

B9320 is a release of llama.cpp, a C/C++ implementation for enabling LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. The project

local-aillama-cpp-releases
26 May 2026
Model Releases

ChainLearn: A Blockchain-Based Capacity-Aware Framework for Federated Ensemble Learning

DGX agent

arXiv:2605.24418v1 Announce Type: new Abstract: Federated learning is used in medical imaging where privacy prohibits centralizing data. Standard federated algorithms assume homogeneous hardware, iden

model-releasesarxiv-cs-lg
26 May 2026
Model Releases

Fine-Tuning and Serving Gemma 4 31B on Google Cloud TPU: A Technical Comparison with GPU Baselines

DGX agent

arXiv:2605.25645v1 Announce Type: cross Abstract: We present the first end-to-end demonstration of fine-tuning and serving Google's Gemma 4 31B model on TPU hardware, providing an empirical comparison

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

OrpQuant: Geometric Orthogonal Residual Projection for Multiplier-Free Power-of-Two Transformer Quantization

DGX agent

arXiv:2605.26092v1 Announce Type: cross Abstract: The deployment of Large Language Models (LLMs) and Vision Transformers (ViTs) on edge devices is significantly constrained by memory limitations and t

model-releasesarxiv-cs-ai
26 May 2026
Research

DCGAN inference on a microcontroller: 12.6M parameters, 512KB SRAM, 26-second generation, pure C [P]

DGX agent

This post describes implementing DCGAN (Deep Convolutional Generative Adversarial Network) inference on resource-constrained microcontroller hardware, achieving image generation with a 12.6 million pa

researchr-machinelearning
25 May 2026
Safety

ReCoVer: Resilient LLM Pre-Training System via Fault-Tolerant Collective and Versatile Workload

DGX agent

arXiv:2605.11215v2 Announce Type: replace-cross Abstract: Pre-training large language models on massive GPU clusters has made hardware faults routine rather than rare, driving the need for resilient t

safetyarxiv-cs-ai
25 May 2026
Research

Efficient training for compact compression models via sequential distillation

DGX agent

arXiv:2601.05639v2 Announce Type: replace Abstract: Deep learning models for image compression often face practical limitations in hardware-constrained applications. Although these models achieve high

researcharxiv-cs-cv
21 May 2026
Model Releases

Hugging Face just released LeRobot Humanoid An open-source, low-cost (~$2.5k), 3D-printed humanoid built for robot learning and not just dem…

DGX agent

Hugging Face just released LeRobot Humanoid An open-source, low-cost (~$2.5k), 3D-printed humanoid built for robot learning and not just demos. What’s cool is it’s a full stack release: • hardware + C

model-releasesclem-delangue--x
21 May 2026
Local Ai

SpineContextResUNet: A Computationally Efficient Residual UNet for Spine CT Segmentation

DGX agent

arXiv:2605.20760v1 Announce Type: new Abstract: Automated segmentation of the vertebral column in Computed Tomography (CT) scans is a prerequisite for pathological assessment and surgical planning. Ho

local-aiarxiv-cs-cv
21 May 2026
Model Releases

VDFP: Video Deflickering with Flicker-banding Priors

DGX agent

arXiv:2605.21079v1 Announce Type: new Abstract: Capturing digital screens with smartphones frequently induces severe banding due to hardware synchronization mismatches. Existing video restoration meth

model-releasesarxiv-cs-cv
21 May 2026
Research

When Does Adaptation Win? Scaling Laws for Meta-Learning in Quantum Control

DGX agent

arXiv:2601.18973v4 Announce Type: replace Abstract: Quantum hardware suffers from intrinsic device heterogeneity and environmental drift, forcing practitioners to choose between suboptimal non-adaptiv

researcharxiv-cs-lg
21 May 2026
Local Ai

b9251

DGX agent

B9251 is a build identifier for a release in the llama.cpp project, which enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware - locally and in the clo

local-aillama-cpp-releases
20 May 2026
Local Ai

DAG-Based QoS-Aware Dynamic Task Placement for Networked Multi-Stage Control Pipelines

DGX agent

arXiv:2605.19887v1 Announce Type: cross Abstract: Current Physical AI (PAI) relies heavily on closed-loop visual-servoing pipelines, whose perception and planning stages may become computationally int

local-aiarxiv-cs-ro
20 May 2026
Model Releases

Information Processing Capacity of Stationary Physical Systems: Theory, Data-efficient Estimation Methods, and Photonic Demonstration

DGX agent

arXiv:2605.19152v1 Announce Type: cross Abstract: Physical computing systems provide a promising route toward hardware-native machine learning, but their computational capabilities remain difficult to

model-releasesarxiv-cs-lg
20 May 2026
Research

A Theory of Training Profit-Optimal LLMs

DGX agent

arXiv:2605.16430v1 Announce Type: cross Abstract: Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure

researcharxiv-cs-ai
19 May 2026
Research

AMS-HD: Hyperdimensional Computing for Real-Time and Energy-Efficient Acute Mountain Sickness Detection

DGX agent

arXiv:2602.08916v2 Announce Type: replace-cross Abstract: Objective: Acute mountain sickness (AMS) is the most prevalent altitude illness, affecting unacclimatized individuals ascending above 2,500 m

researcharxiv-cs-lg
19 May 2026
Safety

Building Reliable Arithmetic Multipliers Under NBTI Aging and Process Variations

DGX agent

arXiv:2605.18444v1 Announce Type: cross Abstract: Hardware aging poses a significant challenge for integrated circuits (ICs), leading to performance degradation and eventual failure. In this work, we

safetyarxiv-cs-ai
19 May 2026
Model Releases

CarbonScaling: Extending Neural Scaling Laws for Carbon Footprint in Large Language Models

DGX agent

arXiv:2508.06524v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly follow neural scaling laws that tie performance gains to rapidly expanding computational budgets, ra

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

DexHoldem: Playing Texas Hold'em with Dexterous Embodied System

DGX agent

arXiv:2605.18727v1 Announce Type: cross Abstract: Evaluating embodied systems on real dexterous hardware requires more than isolated primitive skills: an agent must perceive a changing tabletop scene,

model-releasesarxiv-cs-ai
19 May 2026
Safety

Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems

DGX agent

arXiv:2605.17909v1 Announce Type: new Abstract: As autonomous agentic systems scale across regulated critical infrastructures, the lack of mechanistic, hardware-rooted enforcement for high-frequency p

safetyarxiv-cs-ai
19 May 2026
Research

NeuroLiDAR: Adaptive Frame Rate Depth Sensing via Neuromorphic Event-LiDAR Fusion

DGX agent

arXiv:2605.16805v1 Announce Type: new Abstract: LiDARs are widely used for 3D depth reconstruction, but their performance is often limited by inherent hardware constraints that impose trade-offs betwe

researcharxiv-cs-cv
19 May 2026
Research

Layer Selection in Feature-Based Losses Affects Image Quality and Microstructural Consistency in Deep Learning Super-Resolution of Brain Diffusion MRI

DGX agent

arXiv:2605.15895v1 Announce Type: cross Abstract: Clinical application of high-resolution diffusion MRI is hindered by hardware limitations and prohibitive scan times, motivating computational super-r

researcharxiv-cs-cv
18 May 2026
Model Releases

Searching on a Budget: HW-NAS with 10 Latency Probes

DGX agent

arXiv:2504.00663v2 Announce Type: replace Abstract: Existing hardware-aware NAS (HW-NAS) methods typically assume access to precise information circa the target device, either via analytical approxima

model-releasesarxiv-cs-lg
18 May 2026
Local Ai

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

DGX agent

arXiv:2605.16138v1 Announce Type: cross Abstract: Neural architecture search (NAS) is a powerful approach for automating model design, but existing methods often optimize for accuracy alone or rely on

local-aiarxiv-cs-ai
18 May 2026
Local Ai

b9197

DGX agent

b9197 is a build release of llama.cpp , an open-source C/C++ implementation that enables efficient large language model inference on various hardware platforms. The release includes cross-platform bin

local-aillama-cpp-releases
17 May 2026
Safety

AQKA: Active Quantum Kernel Acquisition Under a Shot Budget

DGX agent

arXiv:2605.14672v1 Announce Type: new Abstract: Estimating an N imes N quantum kernel from circuit fidelities requires Theta(N^2 S) measurement shots, the dominant bottleneck for deployment on near-te

safetyarxiv-cs-lg
15 May 2026
Local Ai

b9159

DGX agent

b9159 is a release of llama.cpp published on May 14, 2026 . llama.cpp is a C/C++ implementation of large language model inference that enables efficient LLM execution on consumer hardware with minimal

local-aillama-cpp-releases
15 May 2026
Local Ai

b9172

DGX agent

b9172 is a release of llama.cpp that includes binaries for macOS, Linux, Android, Windows, and openEuler platforms with support for various hardware configurations including CPU, Vulkan, CUDA, ROCm, O

local-aillama-cpp-releases
15 May 2026
Research

Natural Synthesis: Outperforming Reactive Synthesis Tools with Large Reasoning Models

DGX agent

arXiv:2605.15131v1 Announce Type: new Abstract: Reactive synthesis, the problem of automatically constructing a hardware circuit from a logical specification, is a long-standing challenge in formal ve

researcharxiv-cs-lg
15 May 2026
Model Releases

Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems

DGX agent

arXiv:2511.16964v2 Announce Type: replace-cross Abstract: Maximizing performance on available GPU hardware is an ongoing challenge for modern AI inference systems. Traditional approaches include writi

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Safe Bayesian Optimization for Complex Control Systems via Additive Gaussian Processes

DGX agent

arXiv:2408.16307v3 Announce Type: replace-cross Abstract: Automatic controller tuning is attractive for robotics and mechatronic systems whose dynamics are difficult to model accurately, but direct bl

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…4748495051…94
Next →