AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Hardware

Cartridges at Scale: Training Modular KV Caches over Large Document Collections

DGX agent

arXiv:2606.04557v1 Announce Type: new Abstract: Large Language Models can reason over long contexts, yet prefilling millions of tokens is wasteful as much of the content remains static across queries.

hardwarearxiv-cs-cl
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

DSA: Dynamic Step Allocation for Fast Autoregressive Video Generation

DGX agent

arXiv:2606.04432v1 Announce Type: new Abstract: Video diffusion transformers have achieved state-of-the-art visual quality, but their high inference cost remains a major bottleneck for real-time appli

hardwarearxiv-cs-cv
4 Jun 2026
Hardware

Fitting scattered data with optional monotonicity constraints on GPU: LipFit package

DGX agent

arXiv:2606.04670v1 Announce Type: cross Abstract: This paper presents a method of multivariate scattered data interpolation and approximation that produces optimal Lipschitz-continuous approximation,

hardwarearxiv-cs-lg
4 Jun 2026
Hardware

LLM Compression with Jointly Optimizing Architectural and Quantization choices

DGX agent

arXiv:2606.04063v1 Announce Type: cross Abstract: Deploying large language models (LLMs) is challenging due to their significant memory and computational requirements. While some methods address this

hardwarearxiv-cs-ai
4 Jun 2026
Hardware

Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models

DGX agent

arXiv:2606.04287v1 Announce Type: cross Abstract: Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity,

hardwarearxiv-cs-ai
4 Jun 2026
Hardware

SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

DGX agent

arXiv:2606.04511v1 Announce Type: new Abstract: Sparse attention reduces compute and memory bandwidth for long-context LLM inference. However, two key challenges remain: (1) KV cache capacity still gr

hardwarearxiv-cs-cl
4 Jun 2026
Model Releases

StepPRM-RTL: Stepwise Process-Reward Guided LLM Fine-Tuning for Enhanced RTL Synthesis

DGX agent

arXiv:2606.04246v1 Announce Type: new Abstract: Automatic generation of RTL code for digital hardware designs remains challenging due to long-horizon reasoning, multi-step dependencies, and strict cor

model-releasesarxiv-cs-ai
4 Jun 2026
Hardware

TITAN-FedAnil+: Trust-Based Adaptive Blockchain Federated Learning for Resource-Constrained Intelligent Enterprises

DGX agent

arXiv:2606.04388v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as an effective paradigm for collaborative intelligence while preserving data privacy. However, data heterogeneity

hardwarearxiv-cs-ai
4 Jun 2026
Hardware

GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning

DGX agent

arXiv:2606.02857v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limit

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency

DGX agent

arXiv:2606.03014v1 Announce Type: new Abstract: Mixture-of-Agents (MoA) systems improve reasoning accuracy by routing each query to multiple expert LLMs and aggregating their outputs. Efficiently exec

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

NVIDIA Isaac Sim: Enabling Scalable, GPU-Accelerated Simulation for Robotics

DGX agent

arXiv:2606.03551v1 Announce Type: new Abstract: Simulation has become a core infrastructure for robotics research. Unlike previous simulators, NVIDIA Isaac Sim leverages GPU acceleration to enable lar

hardwarearxiv-cs-ro
3 Jun 2026
Hardware

Pruning Deep Neural Networks via the Marchenko--Pastur Distribution

DGX agent

arXiv:2606.02608v1 Announce Type: new Abstract: We study a Marchenko--Pastur (MP) random-matrix approach to pruning deep neural networks with very small post-pruning fine-tuning budgets. The main prac

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles

DGX agent

arXiv:2505.08222v3 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) offer a cost-effective solution for scientific missions such as underwater tracking. Reinforcement learning (RL) has

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

Speedrunning Tabular Foundation Model Pretraining

DGX agent

arXiv:2606.03681v1 Announce Type: new Abstract: Pretraining cost is a major bottleneck for research on tabular foundation models, slowing the iteration cycle for new architectures, priors, and optimiz

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

Towards Compact Autonomous Driving Perception with Balanced Learning and Multi-sensor Fusion

DGX agent

arXiv:2606.02979v1 Announce Type: cross Abstract: We present a novel compact deep multi-task learning model to handle various autonomous driving perception tasks in one forward pass. The model perform

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

APB-V: Accelerating Long-Video Understanding via Sequence-Parallelism-aware Approximate Attention

DGX agent

arXiv:2601.21444v2 Announce Type: replace-cross Abstract: The efficiency of long-video inference remains a critical bottleneck, mainly due to the dense computation in the prefill stage of Large Multim

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding

DGX agent

arXiv:2606.00144v1 Announce Type: cross Abstract: Speculative decoding speeds up autoregressive decoding by using a drafter to propose multiple tokens that a verifier validates in parallel. In resourc

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Cohort-Scale Neural Atlases of Ultrasound Video

DGX agent

arXiv:2606.00890v1 Announce Type: new Abstract: Ultrasound is the most widely used real-time imaging modality in clinical practice, yet per-frame video annotation remains a major bottleneck: expert la

hardwarearxiv-cs-cv
2 Jun 2026
Hardware

CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving

DGX agent

arXiv:2603.28768v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has recently emerged as the mainstream architecture for efficiently scaling large language models while maintaining n

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap

DGX agent

arXiv:2512.10236v2 Announce Type: replace-cross Abstract: Modern ML workloads demand distributing training and inference across multiple GPUs. However, these parallelization techniques often suffer fr

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing

DGX agent

arXiv:2511.04791v2 Announce Type: replace Abstract: Modern LLM serving systems must sustain high throughput while meeting strict latency SLOs across two distinct inference phases: compute-intensive pr

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

DGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

hardwarearxiv-cs-cl
2 Jun 2026
Hardware

LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models

DGX agent

arXiv:2606.01838v1 Announce Type: cross Abstract: Agentic language model systems alternate between two structurally distinct step types: structured tool calls (short, deterministic, low perplexity) an

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

LithoGRPO: Fast Inverse Lithography via GRPO Reinforced Flow Matching

DGX agent

arXiv:2606.00228v1 Announce Type: new Abstract: In semiconductor manufacturing, lithography projects circuit layouts onto silicon wafers through an optical mask. As circuit features shrink below the w

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Lodestar: An Online-Learning LLM Inference Router

DGX agent

arXiv:2606.00946v1 Announce Type: cross Abstract: Efficiently serving large language model (LLM) inference tasks is crucial both for user-perceived latency such as time-to-first-token (TTFT) and for G

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Observation, Not Prediction: Conversation-Level Disaggregated Scheduling for Agentic Serving

DGX agent

arXiv:2606.01839v1 Announce Type: cross Abstract: LLM-based agents resolve a user task through many turns of dependent inference and tool calls, producing a workload whose total cost is unknown when t

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

PortBERT: Navigating the Depths of Portuguese Language Models

DGX agent

arXiv:2606.02100v1 Announce Type: new Abstract: Transformer models dominate modern NLP, but efficient, language-specific models remain scarce. In Portuguese, most focus on scale or accuracy, often neg

hardwarearxiv-cs-cl
2 Jun 2026
Hardware

Practical Aspects on Solving Differential Equations Using Deep Learning: A Primer

DGX agent

arXiv:2408.11266v5 Announce Type: replace Abstract: Deep learning is now common across many scientific fields, including the study of partial differential equations. This article provides a brief, acc

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Predicting Future Utility: Global Combinatorial Optimization for Task-Agnostic KV Cache Eviction

DGX agent

arXiv:2602.08585v2 Announce Type: replace-cross Abstract: Given the quadratic complexity of attention, KV cache eviction is vital to accelerate model inference. Current KV cache eviction methods typic

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills

DGX agent

arXiv:2503.05641v4 Announce Type: replace-cross Abstract: Combining existing pre-trained LLMs is a promising approach for diverse reasoning tasks. However, task-level expert selection is often too coa

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

DGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

SUPREME: A Multi-GPU Framework for Reproducible Image Unlearning Method Evaluation

DGX agent

arXiv:2606.00380v1 Announce Type: cross Abstract: Machine unlearning removes the influence of specific training data from a trained model without retraining it from scratch. Evaluating an unlearning m

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Batched Differentiable Rigid Body Dynamics in PyTorch for GPU-Accelerated Robot Learning

DGX agent

arXiv:2605.31481v1 Announce Type: new Abstract: As robot control shifts toward large-scale reinforcement learning with in-loop dynamics computation, the community's reliance on CPU-bound libraries suc

hardwarearxiv-cs-ro
1 Jun 2026
Hardware

CoMo3R-SLAM: Collaborative Monocular Dense SLAM with Learned 3D Reconstruction Priors for Outdoor Multi-Agent Systems

DGX agent

arXiv:2605.30488v1 Announce Type: new Abstract: Collaborative dense SLAM is essential for multi-robot teams to achieve scalable and consistent 3D perception across large-scale outdoor environments. Ex

hardwarearxiv-cs-ro
1 Jun 2026
Hardware

Cross-Entropy Optimization of Physically Grounded Task and Motion Plans

DGX agent

arXiv:2512.11571v2 Announce Type: replace Abstract: Autonomously performing tasks often requires robots to plan high-level discrete actions and continuous low-level motions to realize them. Previous T

hardwarearxiv-cs-ro
1 Jun 2026
Hardware

Deep-learning-based low-energy trigger algorithms for the Hyper-Kamiokande experiment

DGX agent

arXiv:2605.31391v1 Announce Type: cross Abstract: Modern machine learning techniques have become increasingly important in particle physics because of their powerful pattern-recognition capabilities,

hardwarearxiv-cs-lg
1 Jun 2026
Hardware

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

DGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

hardwarearxiv-cs-cl
1 Jun 2026
Hardware

DSD-GS: Dynamic-Static Decomposition of Gaussian Splatting for Efficient and High-Fidelity Dynamic Scene Reconstruction

DGX agent

arXiv:2605.30863v1 Announce Type: new Abstract: Dynamic scene reconstruction and novel view synthesis are fundamental to next-generation visual intelligence applications such as virtual reality, robot

hardwarearxiv-cs-cv
1 Jun 2026
Hardware

Elastic ViTs from Pretrained Models without Retraining

DGX agent

arXiv:2510.17700v2 Announce Type: replace Abstract: Vision foundation models achieve remarkable performance but are only available in a limited set of pre-determined sizes, forcing sub-optimal deploym

hardwarearxiv-cs-cv
1 Jun 2026
Research

Gradient-Free Training of Spiking Neural Networks via Low-Rank Evolution Strategies

DGX agent

arXiv:2605.30361v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer compelling energy efficiency on neuromorphic hardware, yet their training remains challenging because the discret

researcharxiv-cs-ai
1 Jun 2026
Research

HetCCL: Enabling Collective Communication For Mixed-Vendor Heterogeneous Clusters

DGX agent

arXiv:2605.31000v1 Announce Type: cross Abstract: Training Large Language Models (LLMs) on heterogeneous clusters presents significant challenges for collective communication, as hardware from multipl

researcharxiv-cs-lg
1 Jun 2026
Hardware

Neurosim: A Fast Simulator for Neuromorphic Robot Perception

DGX agent

arXiv:2602.15018v2 Announce Type: replace-cross Abstract: Neurosim is a fast, real-time, high-performance library for simulating sensors such as dynamic vision sensors, RGB cameras, depth sensors, and

hardwarearxiv-cs-cv
1 Jun 2026
Hardware

ParisKV: Fast and Drift-Robust KV-Cache Retrieval for Long-Context LLMs

DGX agent

arXiv:2602.07721v3 Announce Type: replace-cross Abstract: KV-cache retrieval is essential for long-context LLM inference, yet existing methods struggle with distribution drift and high latency at scal

hardwarearxiv-cs-cl
1 Jun 2026
Hardware

PithTrain: A Compact and Agent-Native MoE Training System

DGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

hardwarearxiv-cs-ai
1 Jun 2026
Hardware

Reducing the GPU Memory Bottleneck with Lossless Compression for ML -- Extended

DGX agent

arXiv:2605.30728v1 Announce Type: new Abstract: Machine learning (ML) training and inference often process data sets far exceeding GPU memory capacity, forcing them to rely on PCIe for on-demand tenso

hardwarearxiv-cs-lg
1 Jun 2026
Hardware

SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer

DGX agent

arXiv:2605.30409v1 Announce Type: cross Abstract: Real-time streaming video-to-video editing (V2V) is critical for interactive applications such as live broadcasting and gaming, yet it remains a formi

hardwarearxiv-cs-ai
1 Jun 2026
Hardware

Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Generation

DGX agent

arXiv:2605.30716v1 Announce Type: cross Abstract: Generating clinically useful pathology reports for pathology cases from whole-slide images (WSIs) is challenging due to gigapixel resolution, long vis

hardwarearxiv-cs-ai
1 Jun 2026
Hardware

SubsurfaceGen: Procedural Generation of Field-Scale Earth Models and Seismic Data

DGX agent

arXiv:2605.30541v1 Announce Type: new Abstract: Full waveform inversion (FWI) is the gold standard for subsurface imaging, with applications from carbon sequestration to energy and mineral exploration

hardwarearxiv-cs-lg
1 Jun 2026
← Previous
1…1314151617…55
Next →