AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,479 results
Model Releases

ABot-M0: VLA Foundation Model for Robotic Manipulation with Action Manifold Learning

DGX agent

arXiv:2602.11236v2 Announce Type: replace-cross Abstract: Building general-purpose embodied agents across diverse hardware remains a central challenge in robotics, often framed as the ''one-brain, man

model-releasesarxiv-cs-cl
15 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials

Accelerating decode-heavy LLM inference with speculative decoding on AWS Trainium and vLLM

DGX agent

Speculative decoding is a technique used to accelerate the slow, sequential token generation (decode stage) in LLM inference. This method significantly reduces latency and improves hardware utilizatio

tutorialsaws-ml-blog
15 Apr 2026
Local Ai

Best realism model under 16GB VRAM

DGX agent

This r/StableDiffusion Reddit thread discusses community recommendations for photorealistic image generation models that can run within a 16GB VRAM constraint, a common hardware limit for consumer GPU

local-air-stablediffusion
15 Apr 2026
Local Ai

b8790

DGX agent

Build b8790 is an incremental automated release of llama.cpp, the open-source C/C++ library for efficient LLM inference on local hardware. Like other builds in the project's continuous release cycle,

local-aillama-cpp-releases
14 Apr 2026
Agents

Problem Reductions at Scale: Agentic Integration of Computationally Hard Problems

DGX agent

arXiv:2604.11535v1 Announce Type: new Abstract: Solving an NP-hard optimization problem often requires reformulating it for a specific solver -- quantum hardware, a commercial optimizer, or a domain h

agentsarxiv-cs-ai
14 Apr 2026
Applications

Quantum computing moves into real-world workflows

DGX agent

As the fifth annual World Quantum Day gets underway on April 14, researchers say the focus of their efforts is shifting from hardware experimentation to integrating quantum systems into broader comput

applicationssiliconangle
14 Apr 2026
Model Releases

Three Roles, One Model: Role Orchestration at Inference Time to Close the Performance Gap Between Small and Large Agents

DGX agent

arXiv:2604.11465v1 Announce Type: new Abstract: Large language model (LLM) agents show promise on realistic tool-use tasks, but deploying capable agents on modest hardware remains challenging. We stud

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning

DGX agent

arXiv:2507.04736v2 Announce Type: replace Abstract: Large Language Models have emerged as powerful tools for automating Register-Transfer Level (RTL) code generation, yet they face critical limitation

safetyarxiv-cs-ai
13 Apr 2026
Local Ai

LogLens with local AI “Ollama”

DGX agent

LogLens with local AI 'Ollama' is a Reddit post on r/ollama discussing the integration of LogLens — a log analysis tool — with Ollama to perform AI-powered log inspection entirely on local hardware, w

local-air-ollama
13 Apr 2026
Industry

Starlink Mini enables reliable high-speed internet on the go 🛰️🛣️

DGX agent

Starlink Mini is a compact, portable version of SpaceX's Starlink satellite internet hardware designed for mobile and on-the-go connectivity. It offers high-speed internet access in a smaller, more tr

industryelon-musk--x
13 Apr 2026
Model Releases

Watt Counts: Energy-Aware Benchmark for Sustainable LLM Inference on Heterogeneous GPU Architectures

DGX agent

arXiv:2604.09048v1 Announce Type: cross Abstract: While the large energy consumption of Large Language Models (LLMs) is recognized by the community, system operators lack guidance for energy-efficient

model-releasesarxiv-cs-ai
13 Apr 2026
Local Ai

Has anyone actually gotten a reliable local AI system running?

DGX agent

This r/ollama Reddit thread addresses a common question among the local AI community about whether reliable, self-hosted AI systems are genuinely achievable. Community members in this space typically

local-air-ollama
12 Apr 2026
Model Releases

KIV: 1M token context window on a RTX 4070 (12GB VRAM), no retraining, drop-in HuggingFace cache replacement - Works with any model that uses DynamicCache [P]

DGX agent

KIV is a project shared on r/MachineLearning presenting a drop-in replacement for HuggingFace's `DynamicCache` that enables up to 1 million token context windows on consumer hardware with only 12GB of

model-releasesr-machinelearning
12 Apr 2026
Local Ai

ace step 1.5 xl sft terrible results

DGX agent

A Reddit thread on r/StableDiffusion where a user reports poor output quality when using the ACE-Step 1.5 XL SFT model variant for AI music generation. The SFT (Supervised Fine-Tuning) variant of ACE-

local-air-stablediffusion
11 Apr 2026
Local Ai

b8752

DGX agent

llama.cpp release **b8752** is a tagged build of the [ggml-org/llama.cpp](https://github.com/ggml-org/llama.cpp) project — a C/C++ framework for efficient LLM inference on a wide range of hardware....

local-aillama-cpp-releases
11 Apr 2026
Safety

AI-Driven Research for Databases

DGX agent

arXiv:2604.06566v1 Announce Type: cross Abstract: As the complexity of modern workloads and hardware increasingly outpaces human research and engineering capacity, existing methods for database perfor

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

CADENCE: Context-Adaptive Depth Estimation for Navigation and Computational Efficiency

DGX agent

arXiv:2604.07286v1 Announce Type: cross Abstract: Autonomous vehicles deployed in remote environments typically rely on embedded processors, compact batteries, and lightweight sensors. These hardware

model-releasesarxiv-cs-ai
10 Apr 2026
Local Ai

Does ollama cloud pro will generate token faster than free?

DGX agent

Ollama Cloud offers Free, Pro ($20/month), and Max ($100/month) subscription tiers for cloud-hosted inference, but token generation speed depends on model size, architecture, and hardware optimiza...

local-air-ollama
10 Apr 2026
Research

Extraction of linearized models from pre-trained networks via knowledge distillation

DGX agent

arXiv:2604.06732v1 Announce Type: new Abstract: Recent developments in hardware, such as photonic integrated circuits and optical devices, are driving demand for research on constructing machine learn

researcharxiv-cs-lg
10 Apr 2026
Tutorials

Migrating to Google Cloud’s Application Load Balancer: A practical guide

DGX agent

Migrating your existing application load balancer infrastructure from an on-premises hardware solution to Cloud Load Balancing offers substantial advantages in scalability, cost-efficiency, and tight

tutorialsgoogle-cloud-ai
10 Apr 2026
Local Ai

MLX creator @awnihannun sharing the story of MLX. Apple management called him right after the launch: “why didn’t you tell us this was going…

DGX agent

MLX is an open-source machine learning framework developed by Apple ML researcher Awni Hannun and colleagues, announced on December 5, 2023, and designed specifically for Apple Silicon hardware. Th...

local-aiollama--x
10 Apr 2026
Model Releases

The Geometry of Forgetting

DGX agent

arXiv:2604.06222v1 Announce Type: cross Abstract: Why do we forget? Why do we remember things that never happened? The conventional answer points to biological hardware. We propose a different one: ge

model-releasesarxiv-cs-ai
10 Apr 2026
Local Ai

b8739

DGX agent

llama.cpp release **b8739** is a build of the open-source C/C++ LLM inference engine that introduces HIP backend support for the CDNA4 (gfx950) GPU architecture, enabling hardware acceleration on A...

local-aillama-cpp-releases
9 Apr 2026
Model Releases

EXPERIMENT: Qwen3.8-2.4T-A95B running locally on an RTX 5090 + RTX 5060 Ti at ~0.80 tok/s

DGX agent

I managed to get Qwen3.8-2.4T-A95B running locally with llama.cpp on mu PC just for fun, cause why not. I was using the Unsloth Qwen3.8-2.4T-A95B-UD-Q1_0 GGUF quantization. The full GGUF is about 397

model-releasesr-localllama
13 Aug 2026
Safety

Generative Learning for Quantum Measurement Design

DGX agent

arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting

safetyarxiv-cs-lg
13 Aug 2026
Safety

A Neural Network Based Teleoperation for Remote Controlled Vehicles

DGX agent

arXiv:2608.10367v1 Announce Type: new Abstract: Direct teleoperation of vehicles faces critical technical bottlenecks: communication latency and the operator's inability to physically perceive unmodel

safetyarxiv-cs-ro
12 Aug 2026
Local Ai

Conversational Orchestration for Organic 6G

DGX agent

arXiv:2608.10714v1 Announce Type: cross Abstract: The Organic 6G vision of a network of networks spanning an edge-cloud continuum complemented by non-terrestrial resources requires, to realize its pro

local-aiarxiv-cs-ai
12 Aug 2026
Agents

Seeing above the waves: A modular sensing framework for data acquisition at sea

DGX agent

arXiv:2608.10997v1 Announce Type: new Abstract: Advancing autonomy for surface vessels requires systematic evaluation of their sensing and perception subsystems. Yet, maritime environments impose uniq

agentsarxiv-cs-ro
12 Aug 2026
Model Releases

Static in Frames, Dynamic in Events: Rethinking Features in Event Cameras as Motion Cues

DGX agent

arXiv:2608.11075v1 Announce Type: new Abstract: Event cameras capture intensity changes asynchronously with high temporal resolution, requiring novel preprocessing methods for downstream tasks. Unlike

model-releasesarxiv-cs-cv
12 Aug 2026
Safety

ArchAgent v2: A Case Study with the Data Prefetching Championship

DGX agent

arXiv:2608.09874v1 Announce Type: new Abstract: Agentic artificial intelligence has shown great promise in automating algorithm design, but scaling similar techniques to computer microarchitecture dis

safetyarxiv-cs-ai
11 Aug 2026
Research

Data collection from highways: a geometric, class-agnostic approach to embedded vehicle counting

DGX agent

arXiv:2608.07643v1 Announce Type: new Abstract: Traffic data collection is dominated today by deep object detectors followed by tracking-by-detection, a pipeline that presupposes what is often missing

researcharxiv-cs-cv
11 Aug 2026
Research

Transformer-Based Neural Quantum Digital Twins for Many-Body Spectral Reconstruction and Adaptive Quantum-Annealing Schedule Design

DGX agent

arXiv:2505.15662v3 Announce Type: replace-cross Abstract: We introduce Transformer-based Neural Quantum Digital Twins (Tx-NQDTs) to reconstruct the low-energy spectral evolution of many-body quantum s

researcharxiv-cs-ai
11 Aug 2026
Research

Ultra-Low-Impedance Robotic Gripper for High-Bandwidth and Transparent Physical Interaction

DGX agent

arXiv:2608.09198v1 Announce Type: new Abstract: Conventional robotic grippers often use high-ratio transmissions to generate grasping torque and external force sensors to measure physical interaction.

researcharxiv-cs-ro
11 Aug 2026
Model Releases

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

DGX agent

Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apa

model-releasesr-localllama
10 Aug 2026
Local Ai

MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model

DGX agent

arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy

local-aiarxiv-cs-lg
10 Aug 2026
Model Releases

Multi-Level Modeling of Large Language Model Inference Latency and Energy via Hybrid Analytical--Machine-Learning Predictors

DGX agent

arXiv:2608.06723v1 Announce Type: cross Abstract: The rapid scaling of Large Language Models (LLMs) has significantly increased computational cost, energy consumption, and inference latency, making ac

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Multiple Hypothesis Flow Estimation for Video Frame Interpolation under Matching Ambiguity

DGX agent

arXiv:2608.07120v1 Announce Type: new Abstract: Many flow-based video frame interpolation (VFI) methods synthesize an intermediate frame by estimating optical flow fields, warping the two input frames

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Quantization Damage Is Multiplicative, Not Additive

DGX agent

arXiv:2608.06564v1 Announce Type: cross Abstract: Quantization is how large language models are actually deployed, and below four bits it is known to hurt. What nobody can say is which of the model's

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

A llama.cpp PR makes Q2_0 3.0–3.6x faster on x86 CPUs, 8B decode goes 2.39 → 8.20 tok/s

DGX agent

I was going through the current llama.cpp CPU PRs and #26348 stood out because this isn't the usual +5% kernel optimization. It adds an x86 VNNI implementation for the Q2_0 × Q8_0 dot product, and the

model-releasesr-localllama
7 Aug 2026
Model Releases

Advancing brain tumor research with privacy-first AI

DGX agent

The intersection of medicine and AI has led to remarkable innovations. However, developers now face the thorny challenge of building robust medical AI tools that have been tested and evaluated on dive

model-releasesgoogle-cloud-ai
6 Aug 2026
Tutorials

Dense Metric Depth Completion from Sparse Direct Time-of-Flight Sensors

DGX agent

arXiv:2608.04737v1 Announce Type: new Abstract: Direct Time-of-Flight (dToF) sensors provide highly accurate metric depth and are more robust than indirect ToF systems in challenging real-world condit

tutorialsarxiv-cs-cv
6 Aug 2026
Model Releases

I ported vLLM's serving stack to C++20: 66 MiB binary, no Python at inference, output checked token-for-token against vLLM

DGX agent

I'm the author, so discount the enthusiasm accordingly. This is an unaffiliated community port, not endorsed by the vLLM project, which it uses to verify its correctness. What started it: I love vLLM,

model-releasesr-localllama
6 Aug 2026
Industry

OpenAI says Apple’s trade secrets lawsuit is ‘rotten to its core’

DGX agent

OpenAI has asked a federal judge to toss out Apple's landmark lawsuit accusing the ChatGPT maker of stealing trade secrets, describing the allegations as 'meritless.' In a motion filed yesterday to di

industrythe-verge-ai
6 Aug 2026
Model Releases

They almost catched up on Frontier performance, so now catching up on prices

DGX agent

Users also report that the free version was significantly downgraded after the release of the new models this is very important for us when considering local hosting. A lot of people decided not to bu

model-releasesr-localllama
6 Aug 2026
Research

NANQ: Noise-Floor-Aware Mixed-Precision Non-Uniform Quantization for Analog Compute-in-Memory

DGX agent

arXiv:2608.02700v1 Announce Type: cross Abstract: Analog compute-in-memory (CIM) enables energy-efficient neural network inference, but device variation and read noise can severely degrade low-bit qua

researcharxiv-cs-ai
5 Aug 2026
Safety

Privacy-Preserving AI Verification via Minimal Information Disclosure

DGX agent

arXiv:2608.02774v1 Announce Type: cross Abstract: AI verification crosses a trust boundary: a verifier must learn enough to establish an authorized claim, yet the same evidence can reveal sensitive de

safetyarxiv-cs-ai
5 Aug 2026
Applications

Semantic Haptic Feedback Enhances Dexterous Robotic Teleoperation

DGX agent

arXiv:2608.02780v1 Announce Type: new Abstract: In robot teleoperation, haptic feedback can be used to help human operators accomplish dexterous manipulation tasks. However, existing haptic feedback m

applicationsarxiv-cs-ro
5 Aug 2026
Safety

CoLI: A Reproducible Platform for Continuum Robot Learning via Monolithic 3D Printing and Isomorphic Teleoperation

DGX agent

arXiv:2606.20389v2 Announce Type: replace Abstract: Continuum robots offer strong potential for manipulation tasks due to their high degrees of freedom, compliant structures, and operational safety. H

safetyarxiv-cs-ro
4 Aug 2026
← Previous
1…5051525354…94
Next →