AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Hardware

AdaGScale: Viewpoint-Adaptive Gaussian Scaling in 3D Gaussian Splatting to Reduce Gaussian-Tile Pairs

DGX agent

arXiv:2604.18980v1 Announce Type: new Abstract: Reducing the number of Gaussian-tile pairs is one of the most promising approaches to improve 3D Gaussian Splatting (3D-GS) rendering speed on GPUs. How

hardwarearxiv-cs-cv
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs

DGX agent

arXiv:2604.18788v1 Announce Type: new Abstract: Apple Neural Engine (ANE) is a dedicated neural processing unit (NPU) present in every Apple Silicon chip. Mixture-of-Experts (MoE) LLMs improve inferen

hardwarearxiv-cs-lg
22 Apr 2026
Hardware

LBLLM: Lightweight Binarization of Large Language Models via Three-Stage Distillation

DGX agent

arXiv:2604.19167v1 Announce Type: cross Abstract: Deploying large language models (LLMs) in resource-constrained environments is hindered by heavy computational and memory requirements. We present LBL

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

Preserving Clusters in Error-Bounded Lossy Compression of Particle Data

DGX agent

arXiv:2604.18801v1 Announce Type: new Abstract: Lossy compression is widely used to reduce storage and I/O costs for large-scale particle datasets in scientific applications such as cosmology, molecul

hardwarearxiv-cs-lg
22 Apr 2026
Hardware

Silicon Aware Neural Networks

DGX agent

arXiv:2604.19334v1 Announce Type: new Abstract: Recent work in the machine learning literature has demonstrated that deep learning can train neural networks made of discrete logic gate functions to pe

hardwarearxiv-cs-cv
22 Apr 2026
Hardware

Text Slider: Efficient and Plug-and-Play Continuous Concept Control for Image/Video Synthesis via LoRA Adapters

DGX agent

arXiv:2509.18831v2 Announce Type: replace-cross Abstract: Recent advances in diffusion models have significantly improved image and video synthesis. In addition, several concept control methods have b

hardwarearxiv-cs-ai
22 Apr 2026
Hardware

AQPIM: Breaking the PIM Capacity Wall for LLMs with In-Memory Activation Quantization

DGX agent

arXiv:2604.18137v1 Announce Type: cross Abstract: Processing-in-Memory (PIM) architectures offer a promising solution to the memory bottlenecks in data-intensive machine learning, yet often overlook t

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Bit-Flip Vulnerability of Shared KV-Cache Blocks in LLM Serving Systems

DGX agent

arXiv:2604.17249v1 Announce Type: cross Abstract: Rowhammer on GPU DRAM has enabled adversarial bit flips in model weights; shared KV-cache blocks in LLM serving systems present an analogous but previ

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Copy-as-Decode: Grammar-Constrained Parallel Prefill for LLM Editing

DGX agent

arXiv:2604.18170v1 Announce Type: new Abstract: LLMs edit text and code by autoregressively regenerating the full output, even when most tokens appear verbatim in the input. We study Copy-as-Decode, a

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

FLASH: Fast Learning via GPU-Accelerated Simulation for High-Fidelity Deformable Manipulation in Minutes

DGX agent

arXiv:2604.17513v1 Announce Type: new Abstract: Simulation frameworks such as Isaac Sim have enabled scalable robot learning for locomotion and rigid-body manipulation; however, contact-rich simulatio

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

FlexiCache: Leveraging Temporal Stability of Attention Heads for Efficient KV Cache Management

DGX agent

arXiv:2511.00868v2 Announce Type: replace Abstract: Large Language Model (LLM) serving is increasingly constrained by the growing size of the key-value (KV) cache, which scales with both context lengt

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Earth System Prediction

DGX agent

arXiv:2604.16590v1 Announce Type: new Abstract: Accurate weather and climate prediction relies on data assimilation (DA), which estimates the Earth system state by integrating observations with models

hardwarearxiv-cs-lg
21 Apr 2026
Research

LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization

DGX agent

arXiv:2604.18117v1 Announce Type: new Abstract: Post-training quantization (PTQ) is essential for deploying large diffusion transformers on resource-constrained hardware, but aggressive 4-bit quantiza

researcharxiv-cs-lg
21 Apr 2026
Hardware

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

DGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

Neptune: Advanced ML Operator Fusion for Locality and Parallelism on GPUs

DGX agent

arXiv:2510.08726v2 Announce Type: replace-cross Abstract: Operator fusion has become a key optimization for deep learning, which combines multiple deep learning operators to improve data reuse and red

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning

DGX agent

arXiv:2509.18169v3 Announce Type: replace-cross Abstract: Tasks on complex systems require high-precision numerical computation to support decisions, but current large language models (LLMs) cannot in

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

POLAR: Online Learning for LoRA Adapter Caching and Routing in Edge LLM Serving

DGX agent

arXiv:2604.16583v1 Announce Type: new Abstract: Edge deployment of large language models (LLMs) increasingly relies on libraries of lightweight LoRA adapters, yet GPU/DRAM can keep only a small reside

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

Probabilistic Programs of Thought

DGX agent

arXiv:2604.17290v1 Announce Type: new Abstract: LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to rea

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

DGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

AutoFlows++: Hierarchical Message Flow Mining for System on Chip Designs

DGX agent

arXiv:2604.15359v1 Announce Type: cross Abstract: Understanding communication behavior in modern system-on-chip (SoC) designs is critical for functional verification, performance analysis, and post-si

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Lossless Compression via Chained Lightweight Neural Predictors with Information Inheritance

DGX agent

arXiv:2604.15472v1 Announce Type: cross Abstract: This paper is dedicated to lossless data compression with probability estimation using neural networks. First, we propose a probability estimation arc

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

PyLO: Towards Accessible Learned Optimizers in PyTorch

DGX agent

arXiv:2506.10315v3 Announce Type: replace Abstract: Learned optimizers have been an active research topic over the past decade, with increasing progress toward practical, general-purpose optimizers th

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

AdaSplash-2: Faster Differentiable Sparse Attention

DGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

DGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

cuRoboV2: Dynamics-Aware Motion Generation with Depth-Fused Distance Fields for High-DoF Robots

DGX agent

arXiv:2603.05493v2 Announce Type: replace Abstract: Effective robot autonomy requires motion generation that is safe, feasible, and reactive. Current methods are fragmented: fast planners output physi

hardwarearxiv-cs-ro
17 Apr 2026
Hardware

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

DGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

DEEP-GAP: Deep-learning Evaluation of Execution Parallelism in GPU Architectural Performance

DGX agent

arXiv:2604.14552v1 Announce Type: cross Abstract: Modern datacenters increasingly rely on low-power, single-slot inference accelerators to balance performance, energy efficiency, and rack density cons

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding

DGX agent

arXiv:2601.14724v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated significant improvement in offline video understanding. Howe

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

DGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

Nautilus: An Auto-Scheduling Tensor Compiler for Efficient Tiled GPU Kernels

DGX agent

arXiv:2604.14825v1 Announce Type: cross Abstract: We present Nautilus, a novel tensor compiler that moves toward fully automated math-to-kernel optimization. Nautilus compiles a high-level algebraic s

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

Thermodynamic Diffusion Inference with Minimal Digital Conditioning

DGX agent

arXiv:2604.14332v1 Announce Type: new Abstract: Diffusion-model inference and overdamped Langevin dynamics are formally identical. A physical substrate that encodes the score function therefore equili

hardwarearxiv-cs-lg
17 Apr 2026
Local Ai

A Lightweight, Transferable, and Self-Adaptive Framework for Intelligent DC Arc-Fault Detection in Photovoltaic Systems

DGX agent

arXiv:2603.25749v2 Announce Type: replace-cross Abstract: Arc-fault circuit interrupters (AFCIs) are essential for mitigating fire hazards in residential photovoltaic (PV) systems, yet achieving relia

local-aiarxiv-cs-lg
16 Apr 2026
Applications

Cross-Layer Co-Optimized LSTM Accelerator for Real-Time Gait Analysis

DGX agent

arXiv:2604.13543v1 Announce Type: cross Abstract: Long Short-Term Memory (LSTM) neural networks have penetrated healthcare applications where real-time requirements and edge computing capabilities are

applicationsarxiv-cs-lg
16 Apr 2026
Hardware

Event Tensor: A Unified Abstraction for Compiling Dynamic Megakernel

DGX agent

arXiv:2604.13327v1 Announce Type: cross Abstract: Modern GPU workloads, especially large language model (LLM) inference, suffer from kernel launch overheads and coarse synchronization that limit inter

hardwarearxiv-cs-lg
16 Apr 2026
Hardware

Fluids You Can Trust: Property-Preserving Operator Learning for Incompressible Flows

DGX agent

arXiv:2602.15472v4 Announce Type: replace-cross Abstract: We present a novel property-preserving kernel-based operator learning method for incompressible flows governed by the incompressible Navier--S

hardwarearxiv-cs-lg
16 Apr 2026
Hardware

Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model

DGX agent

arXiv:2603.28554v2 Announce Type: replace Abstract: Visual document understanding typically requires separate retrieval and generation models, doubling memory and system complexity. We present Hydra,

hardwarearxiv-cs-cv
16 Apr 2026
Hardware

MaMe & MaRe: Matrix-Based Token Merging and Restoration for Efficient Visual Perception and Synthesis

DGX agent

arXiv:2604.13432v1 Announce Type: new Abstract: Token compression is crucial for mitigating the quadratic complexity of self-attention mechanisms in Vision Transformers (ViTs), which often involve num

hardwarearxiv-cs-cv
16 Apr 2026
Hardware

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

DGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

hardwarearxiv-cs-lg
16 Apr 2026
Hardware

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

DGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

hardwarearxiv-cs-cv
16 Apr 2026
Hardware

DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos

DGX agent

arXiv:2604.12270v1 Announce Type: new Abstract: Stereo video inpainting, which aims to fill the occluded regions of warped videos with visually coherent content while maintaining temporal consistency,

hardwarearxiv-cs-cv
15 Apr 2026
Hardware

Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data

DGX agent

arXiv:2503.10676v2 Announce Type: replace-cross Abstract: We study the efficacy of fine-tuning Large Language Models (LLMs) for the specific task of report (government archives, news, intelligence rep

hardwarearxiv-cs-ai
15 Apr 2026
Hardware

PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving

DGX agent

arXiv:2604.12171v1 Announce Type: cross Abstract: Pipeline parallelism (PP) is widely used to partition layers of large language models (LLMs) across GPUs, enabling scalable inference for large models

hardwarearxiv-cs-lg
15 Apr 2026
Hardware

Scalable Trajectory Generation for Whole-Body Mobile Manipulation

DGX agent

arXiv:2604.12565v1 Announce Type: cross Abstract: Robots deployed in unstructured environments must coordinate whole-body motion -- simultaneously moving a mobile base and arm -- to interact with the

hardwarearxiv-cs-cv
15 Apr 2026
Model Releases

A Compact and Efficient 1.251 Million Parameter Machine Learning CNN Model PD36-C for Plant Disease Detection: A Case Study

DGX agent

arXiv:2604.11332v1 Announce Type: cross Abstract: Deep learning has markedly advanced image based plant disease diagnosis as improved hardware and dataset quality have enabled increasingly accurate ne

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

AIRA_2: Overcoming Bottlenecks in AI Research Agents

DGX agent

arXiv:2603.26499v2 Announce Type: replace Abstract: Existing research has identified three structural performance bottlenecks in AI research agents: (1) synchronous single-GPU execution constrains sam

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

DGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

CASK: Core-Aware Selective KV Compression for Reasoning Traces

DGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

hardwarearxiv-cs-ai
14 Apr 2026
← Previous
1…1920212223…55
Next →