AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
10 Apr 2026

AI-Driven Research for Databases

SafetyDGX agent

arXiv:2604.06566v1 Announce Type: cross Abstract: As the complexity of modern workloads and hardware increasingly outpaces human research and engineering capacity, existing methods for database perfor

CADENCE: Context-Adaptive Depth Estimation for Navigation and Computational Efficiency

Model ReleasesDGX agent

arXiv:2604.07286v1 Announce Type: cross Abstract: Autonomous vehicles deployed in remote environments typically rely on embedded processors, compact batteries, and lightweight sensors. These hardware

Does ollama cloud pro will generate token faster than free?

Local AiDGX agent

Ollama Cloud offers Free, Pro ($20/month), and Max ($100/month) subscription tiers for cloud-hosted inference, but token generation speed depends on model size, architecture, and hardware optimiza...

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Extraction of linearized models from pre-trained networks via knowledge distillation

ResearchDGX agent

arXiv:2604.06732v1 Announce Type: new Abstract: Recent developments in hardware, such as photonic integrated circuits and optical devices, are driving demand for research on constructing machine learn

Migrating to Google Cloud’s Application Load Balancer: A practical guide

TutorialsDGX agent

Migrating your existing application load balancer infrastructure from an on-premises hardware solution to Cloud Load Balancing offers substantial advantages in scalability, cost-efficiency, and tight

MLX creator @awnihannun sharing the story of MLX. Apple management called him right after the launch: “why didn’t you tell us this was going…

Local AiDGX agent

MLX is an open-source machine learning framework developed by Apple ML researcher Awni Hannun and colleagues, announced on December 5, 2023, and designed specifically for Apple Silicon hardware. Th...

The Geometry of Forgetting

Model ReleasesDGX agent

arXiv:2604.06222v1 Announce Type: cross Abstract: Why do we forget? Why do we remember things that never happened? The conventional answer points to biological hardware. We propose a different one: ge

9 Apr 2026

b8739

Local AiDGX agent

llama.cpp release **b8739** is a build of the open-source C/C++ LLM inference engine that introduces HIP backend support for the CDNA4 (gfx950) GPU architecture, enabling hardware acceleration on A...

13 Aug 2026

Generative Learning for Quantum Measurement Design

SafetyDGX agent

arXiv:2608.11396v1 Announce Type: cross Abstract: Extracting quantum information from a quantum state is a fundamental task of quantum computation, often requiring the estimation of many non-commuting

12 Aug 2026

A Neural Network Based Teleoperation for Remote Controlled Vehicles

SafetyDGX agent

arXiv:2608.10367v1 Announce Type: new Abstract: Direct teleoperation of vehicles faces critical technical bottlenecks: communication latency and the operator's inability to physically perceive unmodel

Conversational Orchestration for Organic 6G

Local AiDGX agent

arXiv:2608.10714v1 Announce Type: cross Abstract: The Organic 6G vision of a network of networks spanning an edge-cloud continuum complemented by non-terrestrial resources requires, to realize its pro

Seeing above the waves: A modular sensing framework for data acquisition at sea

AgentsDGX agent

arXiv:2608.10997v1 Announce Type: new Abstract: Advancing autonomy for surface vessels requires systematic evaluation of their sensing and perception subsystems. Yet, maritime environments impose uniq

Static in Frames, Dynamic in Events: Rethinking Features in Event Cameras as Motion Cues

Model ReleasesDGX agent

arXiv:2608.11075v1 Announce Type: new Abstract: Event cameras capture intensity changes asynchronously with high temporal resolution, requiring novel preprocessing methods for downstream tasks. Unlike

11 Aug 2026

ArchAgent v2: A Case Study with the Data Prefetching Championship

SafetyDGX agent

arXiv:2608.09874v1 Announce Type: new Abstract: Agentic artificial intelligence has shown great promise in automating algorithm design, but scaling similar techniques to computer microarchitecture dis

Data collection from highways: a geometric, class-agnostic approach to embedded vehicle counting

ResearchDGX agent

arXiv:2608.07643v1 Announce Type: new Abstract: Traffic data collection is dominated today by deep object detectors followed by tracking-by-detection, a pipeline that presupposes what is often missing

Transformer-Based Neural Quantum Digital Twins for Many-Body Spectral Reconstruction and Adaptive Quantum-Annealing Schedule Design

ResearchDGX agent

arXiv:2505.15662v3 Announce Type: replace-cross Abstract: We introduce Transformer-based Neural Quantum Digital Twins (Tx-NQDTs) to reconstruct the low-energy spectral evolution of many-body quantum s

Ultra-Low-Impedance Robotic Gripper for High-Bandwidth and Transparent Physical Interaction

ResearchDGX agent

arXiv:2608.09198v1 Announce Type: new Abstract: Conventional robotic grippers often use high-ratio transmissions to generate grasping torque and external force sensors to measure physical interaction.

10 Aug 2026

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

Model ReleasesDGX agent

Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apa

MiCoPro: End-to-End Mixed Precision HW/SW Co-design with HW-aware Proxy Model

Local AiDGX agent

arXiv:2608.06916v1 Announce Type: new Abstract: Quantized Neural Networks~(QNN) with low-bitwidth data have proven promising in efficient storage and computation on edge devices. To mitigate accuracy

Multi-Level Modeling of Large Language Model Inference Latency and Energy via Hybrid Analytical--Machine-Learning Predictors

Model ReleasesDGX agent

arXiv:2608.06723v1 Announce Type: cross Abstract: The rapid scaling of Large Language Models (LLMs) has significantly increased computational cost, energy consumption, and inference latency, making ac

Multiple Hypothesis Flow Estimation for Video Frame Interpolation under Matching Ambiguity

Model ReleasesDGX agent

arXiv:2608.07120v1 Announce Type: new Abstract: Many flow-based video frame interpolation (VFI) methods synthesize an intermediate frame by estimating optical flow fields, warping the two input frames

Quantization Damage Is Multiplicative, Not Additive

Model ReleasesDGX agent

arXiv:2608.06564v1 Announce Type: cross Abstract: Quantization is how large language models are actually deployed, and below four bits it is known to hurt. What nobody can say is which of the model's

7 Aug 2026

A llama.cpp PR makes Q2_0 3.0–3.6x faster on x86 CPUs, 8B decode goes 2.39 → 8.20 tok/s

Model ReleasesDGX agent

I was going through the current llama.cpp CPU PRs and #26348 stood out because this isn't the usual +5% kernel optimization. It adds an x86 VNNI implementation for the Q2_0 × Q8_0 dot product, and the

6 Aug 2026

Advancing brain tumor research with privacy-first AI

Model ReleasesDGX agent

The intersection of medicine and AI has led to remarkable innovations. However, developers now face the thorny challenge of building robust medical AI tools that have been tested and evaluated on dive

Dense Metric Depth Completion from Sparse Direct Time-of-Flight Sensors

TutorialsDGX agent

arXiv:2608.04737v1 Announce Type: new Abstract: Direct Time-of-Flight (dToF) sensors provide highly accurate metric depth and are more robust than indirect ToF systems in challenging real-world condit

I ported vLLM's serving stack to C++20: 66 MiB binary, no Python at inference, output checked token-for-token against vLLM

Model ReleasesDGX agent

I'm the author, so discount the enthusiasm accordingly. This is an unaffiliated community port, not endorsed by the vLLM project, which it uses to verify its correctness. What started it: I love vLLM,

OpenAI says Apple’s trade secrets lawsuit is ‘rotten to its core’

IndustryDGX agent

OpenAI has asked a federal judge to toss out Apple's landmark lawsuit accusing the ChatGPT maker of stealing trade secrets, describing the allegations as 'meritless.' In a motion filed yesterday to di

They almost catched up on Frontier performance, so now catching up on prices

Model ReleasesDGX agent

Users also report that the free version was significantly downgraded after the release of the new models this is very important for us when considering local hosting. A lot of people decided not to bu

5 Aug 2026

NANQ: Noise-Floor-Aware Mixed-Precision Non-Uniform Quantization for Analog Compute-in-Memory

ResearchDGX agent

arXiv:2608.02700v1 Announce Type: cross Abstract: Analog compute-in-memory (CIM) enables energy-efficient neural network inference, but device variation and read noise can severely degrade low-bit qua

Privacy-Preserving AI Verification via Minimal Information Disclosure

SafetyDGX agent

arXiv:2608.02774v1 Announce Type: cross Abstract: AI verification crosses a trust boundary: a verifier must learn enough to establish an authorized claim, yet the same evidence can reveal sensitive de

Semantic Haptic Feedback Enhances Dexterous Robotic Teleoperation

ApplicationsDGX agent

arXiv:2608.02780v1 Announce Type: new Abstract: In robot teleoperation, haptic feedback can be used to help human operators accomplish dexterous manipulation tasks. However, existing haptic feedback m

4 Aug 2026

CoLI: A Reproducible Platform for Continuum Robot Learning via Monolithic 3D Printing and Isomorphic Teleoperation

SafetyDGX agent

arXiv:2606.20389v2 Announce Type: replace Abstract: Continuum robots offer strong potential for manipulation tasks due to their high degrees of freedom, compliant structures, and operational safety. H

First Deployable Dynamic-CoM: A Unified Policy and Method-Agnostic Benchmark for Humanoid Single-Leg Balance

Model ReleasesDGX agent

arXiv:2608.00500v1 Announce Type: new Abstract: Unified humanoid policies handle agile whole-body motion, yet stumble on a simple demand: staying balanced on one leg. On our single-leg-balance benchma

Gram-Space: Structure-Preserving Codebook Compression for Memory-Efficient Neuro-Symbolic AI

Model ReleasesDGX agent

arXiv:2608.01528v1 Announce Type: new Abstract: Vector symbolic architectures (VSA) are widely used for reasoning in neuro-symbolic (NeSy) AI, yet high-dimensional codebooks often create severe memory

I built a DwarfStar-inspired Vulkan/Metal inference engine for Qwen3.6-35B-A3B on 16 GB machines

Model ReleasesDGX agent

Disclosure: I’m the author and maintainer of QuarkStar. I built QuarkStar, a small native inference engine inspired by Antirez’s DwarfStar. QuarkStar currently supports: Qwen3.6-35B-A3B, using the sam

LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation

SafetyDGX agent

arXiv:2608.01804v1 Announce Type: new Abstract: Post-training large language models (LLMs) via reinforcement learning (RL) has significantly advanced code generation capabilities. To bypass the heavy

Real-Time Visual Obstruction Detection in Surgical Augmented Reality

Model ReleasesDGX agent

arXiv:2608.00232v1 Announce Type: new Abstract: Surgical augmented reality (AR) can provide contextual guidance by overlaying virtual annotations, tool cues, and procedural information onto the surgic

3 Aug 2026

Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning

TutorialsDGX agent

arXiv:2603.25464v2 Announce Type: replace-cross Abstract: Zero-shot reinforcement learning (RL) algorithms aim to learn a family of policies from a reward-free dataset, and recover optimal policies fo

Versatile On-device Adaptation at the Edge by Unifying Few-shot, Zero-shot, Continual, and In-context Learning

Local AiDGX agent

arXiv:2607.29353v1 Announce Type: cross Abstract: With the ever-increasing pervasiveness of smart edge devices, the demand is growing for applications that can be tailored to users (e.g., custom keywo

31 Jul 2026

Belief-Guided Decision Making with Uncertainty Gating in the Game of Go

SafetyDGX agent

arXiv:2607.26946v1 Announce Type: new Abstract: Recent advancements in Computer Go, driven by AlphaZero and MuZero, rely heavily on Monte Carlo Tree Search (MCTS) to correct the errors of the neural n

LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model Inference

Local AiDGX agent

arXiv:2607.27704v1 Announce Type: cross Abstract: As large language models (LLMs) continue to demonstrate exceptional capabilities across various domains, the challenge of achieving energy-efficient a

MiniMax H3: Open-weight multimodel video model

Model ReleasesDGX agent

Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post: Today, we're launching MiniMax H3, a

Minimax-H3 video model released, open weights coming in the next few days

Model ReleasesDGX agent

https://x.com/MiniMax_AI/status/2083006198828417501?s=20 Quote from their article: Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context acro

mmRadarTwin: A Measurement-Calibrated Signal-Level Digital Twin Platform for Indoor mmWave Radar

ResearchDGX agent

arXiv:2607.28108v1 Announce Type: new Abstract: Indoor mmWave radar perception is difficult to reproduce because measured range-angle responses depend on scene geometry, material response, multipath,

Towards Real-Time PixOOD: Efficient Anomaly Segmentation for Autonomous Vehicles

SafetyDGX agent

arXiv:2607.28483v1 Announce Type: new Abstract: Real-time anomaly segmentation is essential for the safety of autonomous systems. Although recent approaches offer high accuracy, their computational co

30 Jul 2026

From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for Embodied Intelligence

TutorialsDGX agent

arXiv:2607.26903v1 Announce Type: cross Abstract: The key bottleneck in embodied AI is not model architecture but data. Although billions of human manipulation videos exist online, robots cannot direc

From Tokens to Watt-hours: Analytical Energy Estimation for LLM Inference on Modern GPUs

Model ReleasesDGX agent

arXiv:2607.26571v1 Announce Type: new Abstract: The operational energy consumption of large language model (LLM) inference is becoming an increasingly important component of the environmental footprin

The Sparsity Ceiling: Where Spiking Networks Can and Cannot Trade Activity for Energy

ResearchDGX agent

arXiv:2607.26648v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are promoted as an energy-efficient substrate because sparse, event-driven activity replaces dense multiply-accumulates

29 Jul 2026

5060ti Chads, vllm updates and nvfp4

Model ReleasesDGX agent

Hey y'all! How is it going. Today this will be a short posting for posterity, mostly so the future llm/scraping overlords catch it since they like reddit and also for anyone out there trying this shit

Automated Modernization of Machine Learning Engineering Notebooks for Reproducibility

Model ReleasesDGX agent

arXiv:2602.07195v2 Announce Type: replace-cross Abstract: Interactive computational notebooks (e.g., Jupyter notebooks) are widely used in machine learning engineering (MLE) to program and share end-t

I tried running a 1.56TB MoE model on a 6GB RTX 4050 Laptop, Here’s the result

Model ReleasesDGX agent

The Test Bench Setup I tested running a massive 1.56TB Mixture-of-Experts (MoE) checkpoint (96 shards, 93 layers, 896 experts/layer, ~4.46 bits/param MXFP4) on a budget gaming laptop. Laptop: HP Victu

OmniQEC: discovering practical quantum error-correcting codes by an AI scientist

ResearchDGX agent

arXiv:2607.25865v1 Announce Type: cross Abstract: Quantum error correction (QEC) is indispensable for scalable fault-tolerant quantum computing. However, discovering QEC codes that remain effective is

28 Jul 2026

A Replay-Constrained Simulation Framework for Personalization of Powered Knee--Ankle Prosthesis Controllers

Model ReleasesDGX agent

arXiv:2607.22858v1 Announce Type: new Abstract: Personalization of impedance controllers for powered prosthetic legs is critical to accommodating individual gait biomechanics but remains challenging.

OrchNAS: Orchestrated Neural Architecture Search Service for Personalised Federated Edge Intelligence

Model ReleasesDGX agent

arXiv:2607.22805v1 Announce Type: cross Abstract: We propose OrchNAS, an energy-aware, personalised, federated edge intelligence framework that leverages a Neural Architecture Search Service to automa

Quotient Tree Arithmetic: Deferred-Division Computation with Bounded Symbolic Depth and Cross-Subtree Cancellation

ResearchDGX agent

arXiv:2607.22612v1 Announce Type: cross Abstract: We introduce Quotient Tree Arithmetic (QTA), a computational substrate in which values are represented as deferred quotient pairs (N, D) whose ratio i

Superpixel-Based QUBO for Scalable Quantum-Enhanced Medical Image Segmentation

ResearchDGX agent

arXiv:2607.24288v1 Announce Type: new Abstract: Quadratic unconstrained binary optimization (QUBO) has emerged as a powerful framework for medical computing problems. Binary decision variables natural

Through the Bottleneck: How Multi-head Latent Attention Separates Content from Position in Language Models

Model ReleasesDGX agent

arXiv:2607.23054v1 Announce Type: cross Abstract: Multi-head Latent Attention (MLA), introduced in DeepSeek-V2, compresses key-value pairs through a shared low-rank bottleneck (cKV), achieving 81% KV-

Unequal Trips, Unequal Places: Diagnosing and Mitigating Delay Inequity in Autonomous Vehicle Fleet Coordination

SafetyDGX agent

arXiv:2607.24336v1 Announce Type: new Abstract: City-scale autonomous vehicle fleet coordinators are typically optimized for aggregate travel time, yet fleet averages conceal how delay is distributed

27 Jul 2026

btw anthropic's internal document on this literally said 'we don't want it to be known that we are working on this.” it was called project p…

Model ReleasesDGX agent

btw anthropic's internal document on this literally said 'we don't want it to be known that we are working on this.” it was called project panama. here's exactly what happened: 1: anthropic concluded

Efficient Domain-Adaptive Policy Learning via Kernel Representation with Application to Quadrotor Control under Non-Stationary Disturbances

SafetyDGX agent

arXiv:2606.13842v2 Announce Type: replace Abstract: We present an algorithm for efficient domain-adaptive policy learning via kernel representations. Learning domain-adaptive policies is challenging s

← Previous
1…4041424344…75
Next →