AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Hardware

Token-Space Mask Prediction for Efficient Vision Transformer Segmentation

DGX agent

arXiv:2605.18177v1 Announce Type: new Abstract: Query-based Vision Transformer segmentation models typically reconstruct dense spatial feature maps to predict masks, inheriting design patterns from co

hardwarearxiv-cs-cv
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks

DGX agent

arXiv:2605.17170v1 Announce Type: new Abstract: Agentic workloads have emerged as a major workload for LLM inference. They differ significantly from chat-only workloads, requiring long-context process

hardwarearxiv-cs-lg
19 May 2026
Hardware

Truthful Calibration Errors for Multi-Class Prediction

DGX agent

arXiv:2510.06388v2 Announce Type: replace Abstract: Calibrated predictions are useful because their numerical values can be interpreted as probabilities. Calibration errors are therefore widely used t

hardwarearxiv-cs-lg
19 May 2026
Hardware

VeriCache: Turning Lossy KV Cache into Lossless LLM Inference

DGX agent

arXiv:2605.17613v1 Announce Type: cross Abstract: The large size of the KV cache has become a major bottleneck for serving LLMs with increasing context lengths. In response, many KV cache compression

hardwarearxiv-cs-lg
19 May 2026
Hardware

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

DGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

hardwarearxiv-cs-ai
18 May 2026
Hardware

A Unified Non-Parametric and Interpretable Point Cloud Analysis via t-FCW Graph Representation

DGX agent

arXiv:2605.15475v1 Announce Type: new Abstract: We introduce an empowered transposed Fully Connected Weighted (t-FCW) graph representation to embed point clouds into a metric space. While original t-F

hardwarearxiv-cs-cv
18 May 2026
Hardware

AgriMind: An Ensemble Deep Learning Framework for Multi-Class Plant Disease Classification

DGX agent

arXiv:2605.16076v1 Announce Type: cross Abstract: Plant disease detection is still largely manual in Bangladesh, where extension workers eyeball leaf samples across millions of smallholdings. We built

hardwarearxiv-cs-ai
18 May 2026
Hardware

Bridging Silicon and the Hippocampus: Algebro-Deterministic Memory 'VaCoAl' as a Substrate for Vector-HaSH and TEM

DGX agent

arXiv:2605.15652v1 Announce Type: cross Abstract: Vector-HaSH and the Tolman-Eichenbaum Machine (TEM) propose that the hippocampal-entorhinal circuit factorizes content from a prestructured grid-cell

hardwarearxiv-cs-ai
18 May 2026
Hardware

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization

DGX agent

arXiv:2605.15824v1 Announce Type: new Abstract: Human-centric video customization, particularly at the garment level, has shown significant commercial value. However, existing approaches cannot suppor

hardwarearxiv-cs-cv
18 May 2026
Hardware

frax: Fast Robot Kinematics and Dynamics in JAX

DGX agent

arXiv:2604.04310v2 Announce Type: replace Abstract: In robot control, planning, and learning, there is a need for rigid-body dynamics libraries that are highly performant, easy to use, and compatible

hardwarearxiv-cs-ro
18 May 2026
Hardware

BEAM: Binary Expert Activation Masking for Dynamic Routing in MoE

DGX agent

arXiv:2605.14438v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enhance the efficiency of large language models by activating only a subset of experts per token. However, standa

hardwarearxiv-cs-ai
15 May 2026
Hardware

DiffPhD: A Unified Differentiable Solver for Projective Heterogeneous Materials in Elastodynamics with Contact-Rich GPU-Acceleration

DGX agent

arXiv:2605.14526v1 Announce Type: cross Abstract: Differentiable simulation of soft bodies is a foundation for system identification, trajectory optimization, and Real2Sim transfer. Yet, existing meth

hardwarearxiv-cs-ro
15 May 2026
Hardware

EMA: Efficient Model Adaptation for Learning-based Systems

DGX agent

arXiv:2605.13942v1 Announce Type: new Abstract: Machine learning (ML) is increasingly applied to optimize system performance in tasks such as resource management and network simulation. Unlike traditi

hardwarearxiv-cs-lg
15 May 2026
Hardware

EnergyLens: Predictive Energy-Aware Exploration for Multi-GPU LLM Inference Optimization

DGX agent

arXiv:2605.14249v1 Announce Type: new Abstract: We present EnergyLens, an end-to-end framework for energy-aware large language model (LLM) inference optimization. As LLMs scale, predicting and reducin

hardwarearxiv-cs-lg
15 May 2026
Hardware

Neural Field Thermal Tomography: A Differentiable Physics Framework for Non-Destructive Evaluation

DGX agent

arXiv:2603.11045v2 Announce Type: replace-cross Abstract: Inverse problems for stiff parabolic partial differential equations (PDEs), such as the inverse heat conduction problem (IHCP), are severely i

hardwarearxiv-cs-ai
15 May 2026
Hardware

Parallelizing Counterfactual Regret Minimization

DGX agent

arXiv:2605.14277v1 Announce Type: new Abstract: Parallelization has played an instrumental role in the field of artificial intelligence (AI), drastically reducing the time taken to train and evaluate

hardwarearxiv-cs-ai
15 May 2026
Model Releases

QOuLiPo: What a quantum computer sees when it reads a book

DGX agent

arXiv:2605.14188v1 Announce Type: cross Abstract: What does a book look like to a quantum computer? This paper takes eight classical works of the Renaissance and its late-antique inheritance -- from A

model-releasesarxiv-cs-cl
15 May 2026
Hardware

SurgicalMamba: Dual-Path SSD with State Regramming for Online Surgical Phase Recognition

DGX agent

arXiv:2605.14889v1 Announce Type: cross Abstract: Online surgical phase recognition (SPR) underpins context-aware operating-room systems and requires committing to a prediction at every frame from pas

hardwarearxiv-cs-ai
15 May 2026
Hardware

Synthetic Sociality: How Generative Models Privatize the Social Fabric

DGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

hardwarearxiv-cs-lg
15 May 2026
Hardware

Towards Resource-Efficient LLMs: End-to-End Energy Accounting of Distillation Pipelines

DGX agent

arXiv:2605.13981v1 Announce Type: cross Abstract: The rise in deployment of large language models has driven a surge in GPU demand and datacenter scaling, raising concerns about electricity use, grid

hardwarearxiv-cs-ai
15 May 2026
Hardware

Woodelf++: A Fast and Unified Partial Dependence Plot Algorithm for Decision Tree Ensembles

DGX agent

arXiv:2605.14578v1 Announce Type: new Abstract: Partial Dependence Plots (PDPs) visualize how changes in a single feature affect the average model prediction. They are widely used in practice to inter

hardwarearxiv-cs-lg
15 May 2026
Hardware

DIVER:Diving Deeper into Distilled Data via Expressive Semantic Recovery

DGX agent

arXiv:2605.12649v1 Announce Type: new Abstract: Dataset distillation aims to synthesize a compact proxy dataset that is unreadable or non-raw from the original dataset for privacy protection and highl

hardwarearxiv-cs-cv
14 May 2026
Hardware

EMO: Frustratingly Easy Progressive Training of Extendable MoE

DGX agent

arXiv:2605.13247v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) models offer a powerful way to scale model size without increasing compute, as per-token FLOPs depend only on k active e

hardwarearxiv-cs-lg
14 May 2026
Hardware

FlashSampling: Fast and Memory-Efficient Exact Sampling

DGX agent

arXiv:2603.15854v2 Announce Type: replace-cross Abstract: Sampling from a categorical distribution is mathematically simple, but in large-vocabulary decoding, it often triggers extra memory traffic an

hardwarearxiv-cs-ai
14 May 2026
Hardware

Flow Augmentation and Knowledge Distillation for Lightweight Face Presentation Attack Detection

DGX agent

arXiv:2605.13108v1 Announce Type: new Abstract: Face presentation attack detection (FacePAD) remains challenging under diverse spoofing representation, including 2D print and replay, 3D mask-based spo

hardwarearxiv-cs-cv
14 May 2026
Hardware

OP4KSR: One-Step Patch-Free 4K Super-Resolution with Periodic Artifact Suppression

DGX agent

arXiv:2605.13457v1 Announce Type: new Abstract: Diffusion-based real-world image super-resolution (Real-ISR) has achieved remarkable perceptual quality; however, directly super-resolving images to 4K

hardwarearxiv-cs-cv
14 May 2026
Hardware

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

DGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

hardwarearxiv-cs-cl
13 May 2026
Hardware

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

DGX agent

arXiv:2512.12131v2 Announce Type: replace Abstract: The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures o

hardwarearxiv-cs-lg
13 May 2026
Hardware

ChunkFlow: Communication-Aware Chunked Prefetching for Layerwise Offloading in Distributed Diffusion Transformer Inference

DGX agent

arXiv:2605.11335v1 Announce Type: cross Abstract: Layerwise offloading reduces the GPU memory footprint of large diffusion transformer (DiT) inference by prefetching upcoming layers from host memory,

hardwarearxiv-cs-lg
13 May 2026
Hardware

Efficient Remote KV Cache Reuse with GPU-native Video Codec

DGX agent

arXiv:2602.09725v3 Announce Type: replace-cross Abstract: Remote KV cache reuse fetches KV cache for identical contexts from remote storage, avoiding recomputation, accelerating LLM inference. While i

hardwarearxiv-cs-lg
13 May 2026
Hardware

Fast MoE Inference via Predictive Prefetching and Expert Replication

DGX agent

arXiv:2605.11537v1 Announce Type: new Abstract: The Mixture of Experts (MoE) architecture has become a fundamental building block in state-of-the-art large language models (LLMs), improving domain-spe

hardwarearxiv-cs-lg
13 May 2026
Hardware

The Illusion of Power Capping in LLM Decode: A Phase-Aware Energy Characterisation Across Attention Architectures

DGX agent

arXiv:2605.11999v1 Announce Type: cross Abstract: Power capping is the standard GPU energy lever in LLM serving, and it appears to work: throughput drops, power readings fall, and energy budgets are m

hardwarearxiv-cs-lg
13 May 2026
Hardware

To Err Is Human; To Annotate, SILICON? Toward Robust Reproducibility in LLM Annotation

DGX agent

arXiv:2412.14461v4 Announce Type: replace Abstract: Unstructured text data annotation is foundational to management research. LLMs offer a cost-effective and scalable alternative to human annotation,

hardwarearxiv-cs-cl
13 May 2026
Hardware

TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion

DGX agent

arXiv:2605.12220v1 Announce Type: new Abstract: Safe autonomous agents and mobile robots need fast real time 3D perception, especially for vulnerable road users (VRUs) such as pedestrians. We introduc

hardwarearxiv-cs-cv
13 May 2026
Hardware

AAAC: Activation-Aware Adaptive Codebooks for 4-bit LLM Weight Quantization

DGX agent

arXiv:2605.08692v1 Announce Type: cross Abstract: Post-training weight-only quantization to 4 bits is widely used to reduce the memory and compute costs of large language model inference. Existing PTQ

hardwarearxiv-cs-cl
12 May 2026
Hardware

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

DGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

hardwarearxiv-cs-cv
12 May 2026
Hardware

Efficient Dataset Distillation for Pre-Trained Self-Supervised Models via Statistical Flow Matching

DGX agent

arXiv:2602.05391v2 Announce Type: replace Abstract: Dataset distillation seeks to synthesize a highly compact dataset that achieves performance comparable to the original dataset on downstream tasks.

hardwarearxiv-cs-cv
12 May 2026
Hardware

Energy Consumption of Dataframe Libraries for End-to-End Deep Learning Pipelines:A Comparative Analysis

DGX agent

arXiv:2511.08644v3 Announce Type: replace-cross Abstract: This paper presents a detailed comparative analysis of the performance of three major Python data manipulation libraries - Pandas, Polars, and

hardwarearxiv-cs-ai
12 May 2026
Hardware

FlashSVD v1.5: Making Low-Rank Transformers Inference Actually Fast

DGX agent

arXiv:2605.08314v1 Announce Type: cross Abstract: SVD-based Low-rank compression reduces transformer parameters and nominal FLOPs, but these savings often translate poorly into real LLM serving speedu

hardwarearxiv-cs-ai
12 May 2026
Hardware

Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models

DGX agent

arXiv:2605.09681v1 Announce Type: new Abstract: Autoregressive (AR) video diffusion models adopt a streaming generation framework, enabling long-horizon video generation with real-time responsiveness,

hardwarearxiv-cs-cv
12 May 2026
Hardware

Geometric 4D Stitching for Grounded 4D Generation

DGX agent

arXiv:2605.09984v1 Announce Type: cross Abstract: Recent 4D generation methods complete scene-level missing information using generative models and reconstruct the scene into radiance-based representa

hardwarearxiv-cs-ai
12 May 2026
Hardware

GPU-Accelerated Synthesis of Mixed-Boolean Arithmetic: Beyond Caching

DGX agent

arXiv:2605.08243v1 Announce Type: cross Abstract: Synthesizing Mixed-Boolean Arithmetic (MBA) expressions from input-output examples is central to program deobfuscation and also useful for compiler op

hardwarearxiv-cs-lg
12 May 2026
Hardware

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

DGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

hardwarearxiv-cs-ai
12 May 2026
Hardware

mHC-SSM: Manifold-Constrained Hyper-Connections for State Space Language Models with Stream-Specialized Adapters

DGX agent

arXiv:2605.08300v1 Announce Type: cross Abstract: Manifold-Constrained Hyper-Connections (mHC) introduce a stability-motivated variant of multi stream residual mixing by constraining residual stream m

hardwarearxiv-cs-ai
12 May 2026
Hardware

Model-Aware Tokenizer Transfer

DGX agent

arXiv:2510.21954v2 Announce Type: replace Abstract: Large Language Models (LLMs) are trained to support an increasing number of languages, yet their predefined tokenizers remain a bottleneck for adapt

hardwarearxiv-cs-cl
12 May 2026
Hardware

Not All Thoughts Need HBM: Semantics-Aware Memory Hierarchy for LLM Reasoning

DGX agent

arXiv:2605.09490v1 Announce Type: new Abstract: Reasoning LLMs produce thousands of chain-of-thought tokens whose KV cache must reside in scarce GPU HBM. The dominant response -- permanently evicting

hardwarearxiv-cs-cl
12 May 2026
Hardware

Novel GPU Boruta algorithms for feature selection from high-dimensional data

DGX agent

arXiv:2605.09950v1 Announce Type: cross Abstract: Most feature selection algorithms, especially wrapper methods, run inefficiently on CPU based platforms because of their high computational complexity

hardwarearxiv-cs-ai
12 May 2026
Hardware

Optimal Transport-Guided Adversarial Attacks on Graph Neural Network-Based Bot Detection

DGX agent

arXiv:2602.00318v2 Announce Type: replace-cross Abstract: The rise of bot accounts on social media poses significant risks to public discourse. To address this threat, modern bot detectors increasingl

hardwarearxiv-cs-ai
12 May 2026
← Previous
1…1617181920…55
Next →