AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Hardware

CPR: Chained Perceptual Refinement for Coarse-to-Fine Medical Image Classification

DGX agent

arXiv:2607.02591v1 Announce Type: new Abstract: High resolution medical images contain fine grained, spatially sparse cues that are critical for diagnosis, yet preserving full resolution incurs substa

hardwarearxiv-cs-cv
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

DGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Fast Asymptotically Optimal Kinodynamic Planning via Vectorization

DGX agent

arXiv:2607.03987v1 Announce Type: new Abstract: Sampling-based motion planners have been shown to be effective for systems with complex kinodynamic constraints and high dimensionality. However, these

hardwarearxiv-cs-ro
7 Jul 2026
Model Releases

From Arithmetic to Logic: The Resilience of Logic and Lookup-Based Neural Networks Under Parameter Bit-Flips

DGX agent

arXiv:2603.22770v2 Announce Type: replace-cross Abstract: The deployment of deep neural networks (DNNs) in safety-critical edge environments necessitates robustness against hardware-induced bit-flip e

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

Graph Sparse Sampling: Breaking the Curse of the Horizon in Continuous MDP Planning

DGX agent

arXiv:2607.05359v1 Announce Type: new Abstract: Planning under uncertainty in continuous domains is essential for autonomous systems, yet computationally demanding. Tree-based search methods such as M

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Lights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy Consumption

DGX agent

arXiv:2607.04553v1 Announce Type: cross Abstract: We present a bidirectional framework for estimating the energy consumption of text-to-video (T2V) and text-to-video-audio (T2VA) models from architect

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

DGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization

DGX agent

arXiv:2607.02834v1 Announce Type: new Abstract: Molecular optimization often starts from a pretrained generative model that captures a broad prior over valid molecular structures. At test time, howeve

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

DGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

DGX agent

arXiv:2504.08791v3 Announce Type: replace-cross Abstract: On-device inference offers privacy, offline use, and instant response, but consumer hardware restricts large language models (LLMs) to low thr

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

RoMa v2: Harder Better Faster Denser Feature Matching

DGX agent

arXiv:2511.15706v3 Announce Type: replace Abstract: Dense feature matching aims to estimate all correspondences between two images of a 3D scene and has recently been established as the gold standard

hardwarearxiv-cs-cv
7 Jul 2026
Hardware

SILO: Simulation-in-the-Loop Sim-to-Real Transfer for Multi-Stage Cable Routing

DGX agent

arXiv:2607.04616v1 Announce Type: cross Abstract: Linear-deformable manipulation remains challenging due to the complex deformations of objects such as cables and ropes. Prior data-driven approaches,

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

DGX agent

arXiv:2607.03333v1 Announce Type: cross Abstract: LLM agents are becoming a common interface for research, coding, and question answering, yet their Thought-Action-Observation loop is often serial: th

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

TabPack: Efficient Hyperparameter Ensembles for Tabular Deep Learning

DGX agent

arXiv:2607.05380v1 Announce Type: new Abstract: In deep learning for tabular data, efficient ensembles of multilayer perceptrons (MLPs) have recently emerged as effective and practical architectures.

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

DGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

DGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

EHHN: An Event-driven Heterogeneous Hypergraph Network for Object-Centric Next Activity Prediction

DGX agent

arXiv:2607.01785v1 Announce Type: new Abstract: Next activity prediction helps service-oriented processes anticipate upcoming steps before delays, exceptions, or service-level risks occur. Most existi

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation

DGX agent

arXiv:2607.01803v1 Announce Type: cross Abstract: Recent advances in 3D content generation from text or images have achieved impressive results, yet view inconsistency from 2D generators and the scarc

hardwarearxiv-cs-ro
3 Jul 2026
Hardware

Stable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory Gates

DGX agent

arXiv:2607.02363v1 Announce Type: cross Abstract: Quantum Fast-Weight Programmers (QFWPs) store temporal information in dynamically programmed variational-circuit parameters rather than in nonlinear r

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

DGX agent

arXiv:2607.00773v1 Announce Type: new Abstract: Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the ac

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

Condensing Large-Scale Datasets Directly with Minimal Information Loss

DGX agent

arXiv:2607.00916v1 Announce Type: new Abstract: Recent advancements in scaling dataset distillation rely heavily on decoupled information extraction pipelines, comprising SQUEEZE, RECOVER, and RELABEL

hardwarearxiv-cs-cv
2 Jul 2026
Hardware

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

DGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics

DGX agent

arXiv:2607.01203v1 Announce Type: cross Abstract: This paper studies real-time robust optimal control for uncertain nonlinear systems, where linear time-varying (LTV) approximations make planning trac

hardwarearxiv-cs-ai
2 Jul 2026
Hardware

MosaicKV: Serving Long-Context LLM with Dynamic Two-D KV Cache Compression

DGX agent

arXiv:2607.00760v1 Announce Type: new Abstract: Long-context LLM services now sustain prompts with hundreds of thousands to millions of tokens, making the key-value (KV) cache a first-order serving co

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

ROSA: A Robotics Foundation Model Serving System for Robot Factories

DGX agent

arXiv:2607.01088v1 Announce Type: new Abstract: Robotics foundation models (RFMs) are making general-purpose robots increasingly practical for factory deployments. While RFM serving systems are centra

hardwarearxiv-cs-ro
2 Jul 2026
Hardware

SNAP-FM: Sparse Nonlinear Accelerated Projection for Physics-Constrained Generative Modeling

DGX agent

arXiv:2607.00095v1 Announce Type: cross Abstract: Generative models have emerged as scalable surrogates for physical simulation, yet they offer no guarantee that their outputs respect the conservation

hardwarearxiv-cs-ai
2 Jul 2026
Hardware

An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU

DGX agent

arXiv:2603.16428v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) has become essential for domain adaptation, but its memory-intensive property exceeds the capabilitie

hardwarearxiv-cs-ai
1 Jul 2026
Hardware

Learning Video Dynamics with Predictive Differentiable Rendering

DGX agent

arXiv:2606.31050v1 Announce Type: cross Abstract: How to accurately predict a high-fidelity future world? While the visual world is inherently continuous, existing deterministic video prediction model

hardwarearxiv-cs-ai
1 Jul 2026
Hardware

SeKV: Resolution-Adaptive KV Cache with Hierarchical Semantic Memory for Long-Context LLM Inference

DGX agent

arXiv:2606.31145v1 Announce Type: new Abstract: Large language models increasingly operate over long contexts, where the KV cache becomes a dominant memory bottleneck: its size grows linearly with seq

hardwarearxiv-cs-cl
1 Jul 2026
Hardware

A Trainable-by-Parts Operator Learning Framework: Bridging DeepONet and Karhunen-Loeve Expansions for Large-Scale Applications

DGX agent

arXiv:2606.28519v1 Announce Type: new Abstract: Training operator-learning models for large-scale problems governed by partial differential equations (PDEs) is challenging due to the curse of dimensio

hardwarearxiv-cs-lg
30 Jun 2026
Hardware

CLEAR-MoE: Shared-Basis Expert Extraction from Frozen Vision Transformers via Calibration-Driven Layer Selection

DGX agent

arXiv:2606.28516v1 Announce Type: new Abstract: We present CLEAR-MoE, a four-phase post-training pipeline that converts a frozen pretrained Vision Transformer (ViT) into a sparse Mixture-of-Experts (M

hardwarearxiv-cs-cv
30 Jun 2026
Hardware

ConCise: Training-Free Conclusion-Chain State Compression for Cost-Efficient Multi-Step RAG Services

DGX agent

arXiv:2606.28361v1 Announce Type: cross Abstract: Multi-step retrieval-augmented generation (RAG) has been widely deployed as LLM-powered web services for complex question answering, where iterative r

hardwarearxiv-cs-ai
30 Jun 2026
Hardware

DreamForge-World 0.1 Preview: A Low-Compute Real-Time Controllable World Model

DGX agent

arXiv:2606.30292v1 Announce Type: cross Abstract: We present DreamForge-World 0.1 Preview, a preview foundational world model for real-time interactive world simulation. The system adapts the LongLive

hardwarearxiv-cs-cv
30 Jun 2026
Hardware

Enhanced Diffusion Sampling: Efficient Rare Event Sampling and Free Energy Calculation with Diffusion Models

DGX agent

arXiv:2602.16634v2 Announce Type: replace-cross Abstract: The rare-event sampling problem has long been the central limiting factor in molecular dynamics (MD), especially in biomolecular simulation. R

hardwarearxiv-cs-ai
30 Jun 2026
Hardware

Evolution Fine-Tuning: Learning to Discover Across 371 Optimization Tasks

DGX agent

arXiv:2606.29082v1 Announce Type: new Abstract: Would experience designing faster GPU kernels also help close in on a long-standing open mathematical conjecture? Large Language Models (LLMs) integrate

hardwarearxiv-cs-cl
30 Jun 2026
Hardware

GPU-Accelerated Inverse Structural Anastylosis from Block Collapse Dynamics

DGX agent

arXiv:2606.28394v1 Announce Type: new Abstract: The physical anastylosis of collapsed architectural monuments -- the meticulous reassembly of fallen stone elements into their original structural confi

hardwarearxiv-cs-cv
30 Jun 2026
Hardware

GPU Parallelization Strategies for Forward and Backward Propagation in Shallow Neural Networks: A CUDA-Based Comparative Study

DGX agent

arXiv:2606.30497v1 Announce Type: cross Abstract: We present a comparative study of CUDA optimization strategies applied to forward and backward propagation in a shallow neural network. Three stacked

hardwarearxiv-cs-lg
30 Jun 2026
Hardware

HARD-KV: Head-Adaptive Regularization for Decoding-time KV Compression

DGX agent

arXiv:2606.28831v1 Announce Type: cross Abstract: Long-context LLM inference faces a fundamental conflict: head-adaptive compression algorithms (e.g., Top-p nucleus sampling) offer superior accuracy b

hardwarearxiv-cs-ai
30 Jun 2026
Local Ai

KernelSight-LM: A Kernel-Level LLM Inference Simulator

DGX agent

arXiv:2606.28565v1 Announce Type: cross Abstract: As large language models (LLMs) move into production serving, practitioners must rapidly evaluate inference performance across diverse hardware, model

local-aiarxiv-cs-ai
30 Jun 2026
Hardware

Knowing in Advance When an Evolutionary Outer Loop Will Not Help: A Pre-Registered Cheap-Baseline Screening Rule

DGX agent

arXiv:2606.29119v1 Announce Type: cross Abstract: We introduce a pre-registered screening rule that decides, before any implementation, whether an evolutionary / population / lifecycle outer loop over

hardwarearxiv-cs-ai
30 Jun 2026
Hardware

RAGA: Real Time Ray Traced Gaussian Shadow Casting for 3DGS Avatar-Scene Interaction

DGX agent

arXiv:2606.29329v1 Announce Type: new Abstract: We study the problem of physically plausible shadow casting when animating 3D Gaussian Splatting (3DGS) avatars, either individually or in multi-avatar

hardwarearxiv-cs-cv
30 Jun 2026
Hardware

Robust and Efficient Monocular 3D Gaussian SLAM for Kilometer-Scale Outdoor Scenes

DGX agent

arXiv:2606.30436v1 Announce Type: new Abstract: Scaling monocular 3D Gaussian Splatting (3DGS) SLAM to kilometer-level outdoor environments poses two tightly coupled challenges: fragile long-term pose

hardwarearxiv-cs-cv
30 Jun 2026
Hardware

Spanning the Visual Analogy Space with a Weight Basis of LoRAs

DGX agent

arXiv:2602.15727v2 Announce Type: replace-cross Abstract: Visual analogy learning enables image editing via demonstration rather than textual description, allowing users to specify complex transformat

hardwarearxiv-cs-ai
30 Jun 2026
Hardware

SSM Meets Video Diffusion Models: Efficient Long-Term Video Generation with Structured State Spaces

DGX agent

arXiv:2403.07711v5 Announce Type: replace-cross Abstract: Given the remarkable achievements in image generation through diffusion models, the research community has shown increasing interest in extend

hardwarearxiv-cs-ai
30 Jun 2026
Hardware

SwarmX: Agentic Scheduling for Low-Latency Agentic Systems

DGX agent

arXiv:2606.21401v2 Announce Type: replace-cross Abstract: Agentic AI applications compose multiple model calls and tool executions, creating new scheduling challenges for GPU-CPU clusters. Their infer

hardwarearxiv-cs-ai
30 Jun 2026
Hardware

Transolver-3: Scaling Up Transformer Solvers to Industrial-Scale Geometries

DGX agent

arXiv:2602.04940v2 Announce Type: replace Abstract: Deep learning has emerged as a transformative tool for the neural surrogate modeling of partial differential equations (PDEs), known as neural PDE s

hardwarearxiv-cs-lg
30 Jun 2026
Hardware

DataStates-LLM: Scalable Checkpointing for Transformer Models Using Composable State Providers

DGX agent

arXiv:2601.16956v1 Announce Type: cross Abstract: The rapid growth of Large Transformer-based models, specifically Large Language Models (LLMs), now scaling to trillions of parameters, has necessitate

hardwarearxiv-cs-ai
29 Jun 2026
Hardware

Test-Input Generation for Tensor Programs: What Actually Finds Kernel Bugs

DGX agent

arXiv:2606.27396v1 Announce Type: cross Abstract: Test-input generation for tensor kernels is folkloric. Most projects pick a representative shape and dtype, run a fixed-shape allclose-style check, an

hardwarearxiv-cs-lg
29 Jun 2026
← Previous
1…1011121314…55
Next →