AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
2,627 results
Hardware

Litespark Inference For CPUs: Ultra-Fast SIMD Framework for Ternary (1.58-bit) Language Models

DGX agent

arXiv:2605.06485v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have transformed artificial intelligence, but their computational requirements remain prohibitive for most users.

hardwarearxiv-cs-ai
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Small Experiments, Cheaper Decisions: A Case Study in Staged Promotion for Micro-Pretraining

DGX agent

arXiv:2606.11387v1 Announce Type: cross Abstract: Short pretraining runs can reduce experimental cost, but they can also over-promote configurations that only look strong at tiny budgets. We study an

hardwarearxiv-cs-ai
11 Jun 2026
Research

A Spiking Neural Architecture for Coordinating Arm and Locomotor Control

DGX agent

arXiv:2606.11034v1 Announce Type: new Abstract: Spiking Neural Networks (SNNs) coupled with neuromorphic hardware offer energy-efficient solutions for humanoid robot control. However, existing SNN-bas

researcharxiv-cs-ro
10 Jun 2026
Hardware

Accounting for AI Inference in Corporate GHG Inventories: A Four-Tier Methodology for Scope 3 Category 1 Reporting

DGX agent

arXiv:2606.10660v1 Announce Type: cross Abstract: AI inference services -- API subscriptions, enterprise chat tools, and SaaS products with embedded AI features -- fall unambiguously within Scope 3 Ca

hardwarearxiv-cs-ai
10 Jun 2026
Hardware

ASTRA-sim 3.0: Next-Level Distributed Machine Learning Simulations via High-Fidelity GPU and Infrastructure Modeling

DGX agent

arXiv:2606.10440v1 Announce Type: cross Abstract: Distributed machine learning (ML) is a key paradigm for today's large-scale artificial intelligence applications. As model inference arises as an impo

hardwarearxiv-cs-lg
10 Jun 2026
Agents

EstRTL: Functional Estimation Guided RTL Code Generation

DGX agent

arXiv:2606.09867v1 Announce Type: cross Abstract: Optimizing register transfer level (RTL) code is of vital importance in hardware design. Large language models (LLMs) provide new methods for the auto

agentsarxiv-cs-ai
10 Jun 2026
Hardware

Flash-GMM: A Memory-Efficient Kernel for Scalable Soft Clustering

DGX agent

arXiv:2606.10896v1 Announce Type: new Abstract: We present extbf{Flash-GMM}, a fused Triton kernel for efficient computation of Gaussian Mixture Models (GMMs) over large-scale data in a single GPU pas

hardwarearxiv-cs-lg
10 Jun 2026
Hardware

One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation

DGX agent

arXiv:2503.13358v5 Announce Type: replace Abstract: Diffusion models for super-resolution (SR) produce high-quality visual results but require expensive computational costs. Despite the development of

hardwarearxiv-cs-cv
10 Jun 2026
Model Releases

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

DGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

model-releasesarxiv-cs-cl
10 Jun 2026
Hardware

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

DGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

hardwarearxiv-cs-lg
10 Jun 2026
Hardware

SpenseGPT: Practical One-shot Pruning Enabling Sparse and Dense GEMMs for LLM Inference

DGX agent

arXiv:2606.10445v1 Announce Type: cross Abstract: Semi-structured 2:4 sparsity is widely supported by modern accelerators, providing up to a 2x theoretical speedup. However, its strict 50% sparsity co

hardwarearxiv-cs-cl
10 Jun 2026
Hardware

torch-sla: Differentiable Sparse Linear Algebra with Adjoint Solvers and Sparse Tensor Parallelism for PyTorch

DGX agent

arXiv:2601.13994v3 Announce Type: replace-cross Abstract: Differentiable sparse linear algebra is foundational for scientific machine learning, yet PyTorch lacks a unified library for it: torch.sparse

hardwarearxiv-cs-ai
10 Jun 2026
Hardware

Unifying Data, Memory, and Compute Efficiency in LLM training: A Survey

DGX agent

arXiv:2606.10706v1 Announce Type: cross Abstract: Resource constraints increasingly determine what can be trained, fine-tuned, and deployed in large language models (LLMs), yet efficiency is often stu

hardwarearxiv-cs-ai
10 Jun 2026
Hardware

Accelerating Birkhoff Projection for Manifold-Constrained Hyper-Connections

DGX agent

arXiv:2606.07574v1 Announce Type: cross Abstract: Manifold-constrained hyper-connections (mHCs) have recently been proposed as a principled extension of hyper-connections, where the residual mixing ma

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

Aqua Boundary-Saliency Attention Module for Lightweight Underwater Salient Instance Segmentation Detection Transformer

DGX agent

arXiv:2606.08002v1 Announce Type: new Abstract: Underwater instance segmentation integrates pixel-level mask prediction and instance-level discrimination for marine resource exploration, ecological mo

hardwarearxiv-cs-cv
9 Jun 2026
Hardware

BRAIN: Bayesian Reasoning via Active Inference for Agentic and Embodied Intelligence in Mobile Networks

DGX agent

arXiv:2602.14033v1 Announce Type: cross Abstract: Future sixth-generation (6G) mobile networks will demand artificial intelligence (AI) agents that are not only autonomous and efficient, but also capa

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

FuseFSS: Efficient Secure LLM Inference with Function Secret Sharing

DGX agent

arXiv:2606.09551v1 Announce Type: cross Abstract: Two-server secure inference allows a client to query a hosted large language model (LLM) without revealing prompts or embeddings. Recent GPU systems b

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

Hybridizing Equilibrium Propagation with Ising Machines for Efficient Energy-Based Learning

DGX agent

arXiv:2606.09112v1 Announce Type: cross Abstract: The rapid evolution of artificial intelligence has led to substantial advances in deep neural networks. Nonetheless, conventional GPU-based training r

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

Kunlun: Establishing Scaling Laws for Massive-Scale Recommendation Systems through Unified Architecture Design

DGX agent

arXiv:2602.10016v3 Announce Type: replace-cross Abstract: Deriving predictable scaling laws that govern the relationship between model performance and computational investment is crucial for designing

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT

DGX agent

arXiv:2601.20408v2 Announce Type: replace-cross Abstract: Enterprise LLM deployment faces a critical scalability challenge: organizations must optimize models systematically to scale AI initiatives wi

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning

DGX agent

arXiv:2606.08039v1 Announce Type: new Abstract: Robotic simulators are a cornerstone of modern research in aerial robotics, serving both as a vehicle for the development of new control algorithms and

hardwarearxiv-cs-ro
9 Jun 2026
Hardware

Resource-aware Computation-Communication Overlap for multi-GPU ML Workloads

DGX agent

arXiv:2606.09200v1 Announce Type: cross Abstract: The rapid growth of large-scale machine learning (ML) has made distributed training across multiple GPUs a fundamental component of modern ML systems.

hardwarearxiv-cs-ai
9 Jun 2026
Model Releases

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

DGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

model-releasesarxiv-cs-ai
9 Jun 2026
Hardware

SoccerNet 2026 Player-Centric Ball-Action Spotting:Retraining and Post-Processing Extensions to the FOOTPASS Baselines

DGX agent

arXiv:2606.09679v1 Announce Type: new Abstract: We describe our system for the SoccerNet 2026 Player-Centric Ball-Action Spotting Challenge, which requires predicting who performs which action and whe

hardwarearxiv-cs-cv
9 Jun 2026
Hardware

STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank Control

DGX agent

arXiv:2606.08382v1 Announce Type: cross Abstract: Low-rank projection has emerged as a promising approach for compressing the KV cache by exploiting hidden-dimension redundancy. However, prior methods

hardwarearxiv-cs-ai
9 Jun 2026
Hardware

TUDSR: Twice Upsampling-Diffusion for Higher Super-Resolution

DGX agent

arXiv:2606.09608v1 Announce Type: new Abstract: Diffusion-based generative models have achieved remarkable success in real-world image super-resolution (SR). With tiled diffusion techniques, these mod

hardwarearxiv-cs-cv
9 Jun 2026
Hardware

Ultra Flash: Scaling Real-Time Streaming Video Generation to High Resolutions

DGX agent

arXiv:2606.09150v1 Announce Type: new Abstract: While recent autoregressive video diffusion models achieve remarkable streaming quality, they remain confined to low resolutions (e.g., 480P), leaving e

hardwarearxiv-cs-cv
9 Jun 2026
Research

Ablation Study of Block Size, Weight Precision, and Scale Precision in NVFP4 Inference for Low-Power Edge-Efficient Neural Networks

DGX agent

arXiv:2606.06527v1 Announce Type: cross Abstract: Energy-efficient edge inference requires reducing arithmetic cost, memory traffic, and hardware overhead. This paper presents an ablation-focused stud

researcharxiv-cs-lg
8 Jun 2026
Hardware

Accelerated Fourier SAT (AFSAT): Fully Realising a GPU-based Symmetric Pseudo-Boolean SAT Solver

DGX agent

arXiv:2606.06641v1 Announce Type: new Abstract: We present Accelerated Fourier SAT (AFSAT), a GPU-accelerated solver for pseudo-Boolean satisfiability based on continuous local search (CLS). AFSAT rea

hardwarearxiv-cs-ai
8 Jun 2026
Hardware

Broadband Hyperspectral 3D Imaging using Dispersed Structured Light

DGX agent

arXiv:2605.25757v2 Announce Type: replace Abstract: Hyperspectral 3D imaging enables the capture of dense spectral information and scene geometry but has traditionally been confined to narrow spectral

hardwarearxiv-cs-cv
8 Jun 2026
Hardware

LuMamba: Latent Unified Mamba for Electrode Topology-Invariant and Efficient EEG Modeling

DGX agent

arXiv:2603.19100v2 Announce Type: replace Abstract: Electroencephalography (EEG) enables non-invasive monitoring of brain activity across clinical and neurotechnology applications, yet building founda

hardwarearxiv-cs-ai
8 Jun 2026
Hardware

MVSegNet: A Lightweight Boundary-Aware Network for Fetal Lateral Ventricle Segmentation and Atrial Width Estimation in Prenatal Ultrasound

DGX agent

arXiv:2606.06958v1 Announce Type: new Abstract: Fetal ventriculomegaly is assessed by measuring the atrial width of the lateral ventricle in prenatal ultrasound. Accurate segmentation is essential for

hardwarearxiv-cs-cv
8 Jun 2026
Research

Predictive Style Matching: Natural and Robust Humanoid Locomotion

DGX agent

arXiv:2606.07083v1 Announce Type: new Abstract: Reinforcement learning has become the prevailing approach to humanoid locomotion control: policies transfer reliably from simulation to hardware and rec

researcharxiv-cs-ro
8 Jun 2026
Model Releases

RhinoVLA Technical Report

DGX agent

arXiv:2606.07383v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for robotic manipulation, but real-time deployment on edge hardware remains challengin

model-releasesarxiv-cs-lg
8 Jun 2026
Hardware

Scalable Joint Resource Allocation for SLO-Constrained LLM Inference in Heterogeneous GPU Clouds

DGX agent

arXiv:2604.07472v2 Announce Type: replace Abstract: Serving large language model (LLM) inference in cloud environments requires jointly optimizing model selection, GPU provisioning, parallelism config

hardwarearxiv-cs-lg
8 Jun 2026
Hardware

TorchKM: A GPU-Oriented Library for Kernel Learning and Model Selection

DGX agent

arXiv:2606.06742v1 Announce Type: new Abstract: TorchKM is an open-source library for kernel machines, including support vector machines, kernel logistic regression, and kernel quantile regression, wi

hardwarearxiv-cs-lg
8 Jun 2026
Hardware

VIRTUS-FPP: Virtual Sensor Modeling for Fringe Projection Profilometry in NVIDIA Isaac Sim

DGX agent

arXiv:2509.22685v2 Announce Type: replace-cross Abstract: Fringe projection profilometry (FPP) is a high-precision structured-light sensing technique for 3D surface reconstruction, yet its practical d

hardwarearxiv-cs-cv
8 Jun 2026
Hardware

Beyond Code Pairs: Dialogue-Based Data Generation for LLM Code Translation

DGX agent

arXiv:2512.03086v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown remarkable capabilities in code translation, yet their performance deteriorates in low-resource progra

hardwarearxiv-cs-ai
6 Jun 2026
Hardware

CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe

DGX agent

arXiv:2604.01489v2 Announce Type: replace-cross Abstract: High-performance GPU kernels are critical to modern machine learning systems, yet developing them remains a manual, expert-driven process. Rec

hardwarearxiv-cs-ai
6 Jun 2026
Hardware

LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels

DGX agent

arXiv:2603.19312v3 Announce Type: replace-cross Abstract: Joint Embedding Predictive Architectures (JEPAs) offer a compelling framework for learning world models in compact latent spaces, yet existing

hardwarearxiv-cs-ai
6 Jun 2026
Hardware

RedKnot: Efficient Long-Context LLM Serving with Head-Aware KV Reuse and SegPagedAttention

DGX agent

arXiv:2606.06256v1 Announce Type: new Abstract: As the input length of large language model (LLM) serving continues to grow, the KV cache has become a dominant bottleneck in AI infrastructure. It limi

hardwarearxiv-cs-ai
6 Jun 2026
Hardware

Flash-WAM: Modality-Aware Distillation for World Action Models

DGX agent

arXiv:2606.05254v1 Announce Type: cross Abstract: World-action models (WAMs) jointly generate future video and robot actions through iterative diffusion, achieving strong performance on manipulation b

hardwarearxiv-cs-cv
5 Jun 2026
Hardware

Gotta Grow Fast: Design and Benchmarking of a Tip Mount for High-Speed Vine Robots

DGX agent

arXiv:2606.06040v1 Announce Type: new Abstract: Soft, growing vine robots extend through tip eversion, a mechanism that enables navigation through cluttered environments. However, integrating cameras

hardwarearxiv-cs-ro
5 Jun 2026
Hardware

GS-NFS: Bandwidth-adaptive Streaming of Dynamic Gaussian Splats and Point Clouds

DGX agent

arXiv:2606.05650v1 Announce Type: cross Abstract: Dynamic 3D Gaussian Splatting (3DGS) holds great promise as a 3D video streaming technology since it can represent complex 3D scenes with high fidelit

hardwarearxiv-cs-cv
5 Jun 2026
Hardware

Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection

DGX agent

arXiv:2506.10601v2 Announce Type: replace Abstract: Given its ability to reduce annotation costs, weakly supervised learning based on single-point annotations has emerged as a research focus in orient

hardwarearxiv-cs-cv
5 Jun 2026
Hardware

Uncertainty-Aware Adaptive Sensor Fusion for Autonomous Navigation

DGX agent

arXiv:2606.05437v1 Announce Type: cross Abstract: This work introduces a hybrid deep learning approach integrated with an Unscented Kalman Filter (UKF) to enhance pose estimation accuracy in Visual-In

hardwarearxiv-cs-cv
5 Jun 2026
Hardware

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

DGX agent

arXiv:2606.05979v1 Announce Type: new Abstract: We propose world-language-action (WLA) models as a new class of embodied foundation models. WLA takes textual instructions, images, and robot states as

hardwarearxiv-cs-ro
5 Jun 2026
Hardware

AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2606.04484v1 Announce Type: new Abstract: We present AgentJet, a distributed swarm training framework for large language model (LLM) agent reinforcement learning. Unlike centralized frameworks t

hardwarearxiv-cs-ai
4 Jun 2026
← Previous
1…1213141516…55
Next →