AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
Hardware

BitLogic: Training Framework for Gradient-Based FPGA-Native Neural Networks

DGX agent

arXiv:2602.07400v2 Announce Type: replace Abstract: Gradient-based LUT- and logic-gate-based neural networks (LUTNet, LogicNets, DiffLogic, PolyLUT, NeuraLUT, WARP-LUT, DWN, LILogicNet, LightLUT) repl

hardwarearxiv-cs-lg
8 Jul 2026
Hardware
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

DBNN: Neural Spike Classification Using a Deep Binarized Neural Network

DGX agent

arXiv:2607.05590v1 Announce Type: cross Abstract: Implantable brain-computer interfaces require on-node spike sorting to reduce telemetry bandwidth and power while maintaining reliable neural decoding

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

Life Cycle Assessment of Pre-training the Lucie 7B Open-Source Large Language Model on the Jean Zay Supercomputer

DGX agent

arXiv:2607.05408v1 Announce Type: cross Abstract: The environmental impact of training large language models (LLMs) is increasingly scrutinised, yet most published estimates focus on operational energ

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation

DGX agent

arXiv:2607.06558v1 Announce Type: new Abstract: Scaling robot learning requires massive, diverse trajectory data, yet collection is currently bottlenecked by physical teleoperation, where every demons

hardwarearxiv-cs-ro
8 Jul 2026
Hardware

MLSYSIM: First-Principles Infrastructure Modeling for Machine Learning Systems

DGX agent

arXiv:2607.02558v1 Announce Type: cross Abstract: As machine learning shifts from laboratory curiosity to critical infrastructure, the systems that sustain it span an extraordinary range, from sub-mil

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

WattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs

DGX agent

arXiv:2607.02391v1 Announce Type: cross Abstract: Large Language Model (LLM) inference workloads are a rapidly growing contributor to data center energy consumption. Optimizing these deployments requi

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

CSAR: Containerized System Architecture for Robotics

DGX agent

arXiv:2606.30293v1 Announce Type: new Abstract: Robotic applications increasingly rely on distributed computational infrastructures that combine embedded devices, edge servers, and cloud resources. Th

hardwarearxiv-cs-ro
30 Jun 2026
Hardware

Optimizing CUDA like a Human: Micro-Profiling Tools as Expert Surrogates for LLM-Based GPU Kernel Optimization

DGX agent

arXiv:2606.26453v1 Announce Type: new Abstract: We present KernelPro, a closed-loop multi-agent system that automatically generates, profiles, and iteratively optimizes GPU kernel code by integrating

hardwarearxiv-cs-lg
26 Jun 2026
Industry

The Inference Alpha: Maximizing Frontier Models on AMD

DGX agent

This article discusses strategies for optimizing the performance of advanced AI frontier models when running on AMD hardware infrastructure. It likely covers deployment best practices, hardware config

industrydigitalocean
10 Jun 2026
Agents

Towards Autonomous Accelerator Design: FPGA Accelerator Generation with SECDA

DGX agent

arXiv:2606.11117v1 Announce Type: cross Abstract: Designing FPGA-based accelerators for modern artificial intelligence workloads requires exploring a large and complex hardware design space that invol

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

DGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

model-releasesarxiv-cs-lg
9 Jun 2026
Hardware

Towards Automated Kernel Generation in the Era of LLMs

DGX agent

arXiv:2601.15727v3 Announce Type: replace Abstract: The performance of modern AI systems is fundamentally constrained by the quality of their underlying GPU kernels, which translate high-level algorit

hardwarearxiv-cs-lg
9 Jun 2026
Model Releases

vla.cpp: A Unified Inference Runtime for Vision-Language-Action Models

DGX agent

arXiv:2606.08094v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically shipped as Python/PyTorch stacks that assume a workstation-class GPU, a mismatch for the hardware

model-releasesarxiv-cs-ai
9 Jun 2026
Local Ai

What are the most capable LLM models I can run on my laptop?

DGX agent

A discussion on r/ollama exploring which high-performance LLM models can be effectively run locally on standard laptop hardware , likely covering model size comparisons, hardware requirements, and per

local-air-ollama
5 Jun 2026
Hardware

Bit-Exact AI Inference Verification Without Performance Tradeoffs

DGX agent

arXiv:2606.00279v1 Announce Type: cross Abstract: Verifying claims about AI workloads is a pre- requisite for credible AI governance of covert adversaries (who comply with monitoring only when detecti

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Bastion: Budget-Aware Speculative Decoding with Tree-structured Block Diffusion Drafting

DGX agent

arXiv:2605.29727v1 Announce Type: new Abstract: Block-diffusion drafters have recently emerged as a powerful alternative for speculative decoding by predicting multiple future-token distributions in a

hardwarearxiv-cs-lg
29 May 2026
Hardware

Inference Time Context Sparsity: Illusion or Opportunity?

DGX agent

arXiv:2605.24168v1 Announce Type: new Abstract: Sparsity has long been a central theme in LLM efficiency, but its role in context processing remains unresolved. As LLM workloads shift toward longer co

hardwarearxiv-cs-ai
26 May 2026
Hardware

ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration

DGX agent

arXiv:2605.22015v1 Announce Type: new Abstract: Diffusion Transformer (DiT) has emerged as a powerful model architecture for generating high-quality images and videos. In the case of video DiT, 3D Spa

hardwarearxiv-cs-cv
22 May 2026
Hardware

COBALT: Crowdsourcing Robot Learning via Cloud-Based Teleoperation with Smartphones

DGX agent

arXiv:2605.19138v1 Announce Type: cross Abstract: The scarcity of large-scale, high-quality demonstration data remains a bottleneck in scaling imitation learning for robotic manipulation. We present C

hardwarearxiv-cs-ai
20 May 2026
Hardware

How Imgix processes 8 billion images daily with G4 VMs powered by NVIDIA Blackwell

DGX agent

The modern web is extremely visual. People are busy and easily-distracted, and smart companies know they have just seconds to attract would-be customers with compelling images, videos, animations, and

hardwaregoogle-cloud-ai
12 May 2026
Hardware

GATO: GPU-Accelerated and Batched Trajectory Optimization for Scalable Edge Model Predictive Control

DGX agent

arXiv:2510.07625v2 Announce Type: replace Abstract: While Model Predictive Control (MPC) delivers strong performance across robotics applications, solving the underlying (batches of) nonlinear traject

hardwarearxiv-cs-ro
11 May 2026
Hardware

LLMSpace: Carbon Footprint Modeling for Large Language Model Inference on LEO Satellites

DGX agent

arXiv:2605.05615v2 Announce Type: replace Abstract: Large language models (LLMs) impose rapidly growing energy demands, creating an emerging energy and carbon crisis driven by large-scale inference. S

hardwarearxiv-cs-lg
11 May 2026
Hardware

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

DGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

hardwarearxiv-cs-lg
11 May 2026
Hardware

When Agents Handle Secrets: A Survey of Confidential Computing for Agentic AI

DGX agent

arXiv:2605.03213v1 Announce Type: cross Abstract: Agentic AI systems, specifically LLM-driven agents that plan, invoke tools, maintain persistent memory, and delegate tasks to peer agents via protocol

hardwarearxiv-cs-ai
7 May 2026
Hardware

A Semantic Autonomy Framework for VLM-Integrated Indoor Mobile Robots: Hybrid Deterministic Reasoning and Cross-Robot Adaptive Memory

DGX agent

arXiv:2605.02525v1 Announce Type: new Abstract: Autonomous indoor mobile robots can navigate reliably to metric coordinates using established frameworks such as ROS 2 Navigation 2, yet they lack the a

hardwarearxiv-cs-ro
5 May 2026
Hardware

The Quantization Trap: Breaking Linear Scaling Laws in Multi-Hop Reasoning

DGX agent

arXiv:2602.13595v2 Announce Type: replace Abstract: Neural scaling laws provide a predictable recipe for AI advancement: reducing numerical precision should linearly improve computational efficiency a

hardwarearxiv-cs-ai
5 May 2026
Hardware

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization

DGX agent

arXiv:2605.02262v1 Announce Type: cross Abstract: Recently, video language models (VLMs) have been applied in various fields. However, the visual token sequence of the VLM is too long, which may cause

hardwarearxiv-cs-cl
5 May 2026
Hardware

EdgeFM: Efficient Edge Inference for Vision-Language Models

DGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

hardwarearxiv-cs-cv
1 May 2026
Model Releases

EdgeSpike: Spiking Neural Networks for Low-Power Autonomous Sensing in Edge IoT Architectures

DGX agent

arXiv:2604.27004v1 Announce Type: cross Abstract: We propose EdgeSpike, a co-designed spiking neural network (SNN) framework for autonomous low-power sensing in edge Internet of Things (IoT) architect

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

asRoBallet: Closing the Sim2Real Gap via Friction-Aware Reinforcement Learning for Underactuated Spherical Dynamics

DGX agent

arXiv:2604.24916v1 Announce Type: new Abstract: We introduce asRoBallet, to the best of our knowledge, the first successful deployment of reinforcement learning (RL) on a humanoid ballbot hardware. Hi

model-releasesarxiv-cs-ro
29 Apr 2026
Local Ai

@ClementDelangue I got mine too yesterday, fitting nicely next to my DGX Spark.

DGX agent

Clement Delangue, CEO of Hugging Face, posted about receiving a piece of hardware that fits alongside a DGX Spark system, likely referring to an AI accelerator or GPU device. The post suggests a perso

local-aiclem-delangue--x
29 Apr 2026
Hardware

Latent Inter-Frame Pruning: A Training-Free Method Bridging Traditional Video Compression and Modern Diffusion Transformers for Efficient Generation

DGX agent

arXiv:2604.23858v1 Announce Type: new Abstract: Video generation, while capable of generating realistic videos, is computationally expensive and slow, prohibiting real-time applications. In this paper

hardwarearxiv-cs-cv
28 Apr 2026
Applications

HGQ-LUT: Fast LUT-Aware Training and Efficient Architectures for DNN Inference

DGX agent

arXiv:2604.22293v1 Announce Type: cross Abstract: Lookup-table (LUT) based neural networks can deliver ultra-low latency and excellent hardware efficiency on FPGAs by mapping arithmetic operations dir

applicationsarxiv-cs-lg
27 Apr 2026
Hardware

DAVIS, APRIL 25, 2026 — InferenceX has added DeepSeekv4 for @vllm_project 's day 0 support for GB200 disagg! Great work to @flowpow123 @roge…

DGX agent

DAVIS, APRIL 25, 2026 — InferenceX has added DeepSeekv4 for @vllm_project 's day 0 support for GB200 disagg! Great work to @flowpow123 @rogerw0108 @NVIDIAAIDev @inferact for the fast support and engin

hardwaredylan-patel--x
25 Apr 2026
Hardware

A Delta-Aware Orchestration Framework for Scalable Multi-Agent Edge Computing

DGX agent

arXiv:2604.20129v1 Announce Type: new Abstract: The Synergistic Collapse occurs when scaling beyond 100 agents causes superlinear performance degradation that individual optimizations cannot prevent.

hardwarearxiv-cs-lg
23 Apr 2026
Hardware

Amortized Vine Copulas for High-Dimensional Density and Information Estimation

DGX agent

arXiv:2604.20568v1 Announce Type: new Abstract: Modeling high-dimensional dependencies while keeping likelihoods tractable remains challenging. Classical vine-copula pipelines are interpretable but ca

hardwarearxiv-cs-lg
23 Apr 2026
Hardware

'Funny and distressingly realistic...propelled by awesome characters and inventive twists”— @andyweirauthor Silicon Valley invents the time …

DGX agent

'Funny and distressingly realistic...propelled by awesome characters and inventive twists”— @andyweirauthor Silicon Valley invents the time machine in my upcoming book PARADOX INC, now available for p

hardwareswyx--x
22 Apr 2026
Hardware

PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment

DGX agent

arXiv:2604.19129v1 Announce Type: new Abstract: Existing facial reenactment methods struggle with a trade-off between expressiveness and fine-grained controllability. Holistic facial reenactment model

hardwarearxiv-cs-cv
22 Apr 2026
Hardware

SpikeMLLM: Spike-based Multimodal Large Language Models via Modality-Specific Temporal Scales and Temporal Compression

DGX agent

arXiv:2604.18610v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress but incur substantial computational overhead and energy consumption during

hardwarearxiv-cs-ai
22 Apr 2026
Industry

A look at Apple's management changes and several key executives' futures; sources say Mike Rockwell has considered leaving or taking an advisory role next year (Mark Gurman/Bloomberg)

DGX agent

Mark Gurman / Bloomberg: A look at Apple's management changes and several key executives' futures; sources say Mike Rockwell has considered leaving or taking an advisory role next year — John Ternus,

industrytechmeme
21 Apr 2026
Hardware

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation

DGX agent

arXiv:2604.18348v1 Announce Type: new Abstract: Video diffusion transformers (DiTs) suffer from prohibitive inference latency due to quadratic attention complexity. Existing sparse attention methods e

hardwarearxiv-cs-cv
21 Apr 2026
Hardware

Framework’s first eGPUs turn its laptop into a desktop PC

DGX agent

Remember when Framework made the first laptop where you can easily upgrade its entire internal video card in three minutes flat? The company's getting into the external graphics game, too. As promised

hardwarethe-verge-ai
21 Apr 2026
Hardware

Real-Time Structural Detection for Indoor Navigation from 3D LiDAR Using Bird's-Eye-View Images

DGX agent

arXiv:2603.19830v2 Announce Type: replace Abstract: Efficient structural perception is essential for mapping and autonomous navigation on resource-constrained robots. Existing 3D methods are computati

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

CPU Optimization of a Monocular 3D Biomechanics Pipeline for Low-Resource Deployment

DGX agent

arXiv:2604.15665v1 Announce Type: new Abstract: Markerless 3D movement analysis from monocular video enables accessible biomechanical assessment in clinical and sports settings. However, most research

hardwarearxiv-cs-cv
20 Apr 2026
Hardware

Silicon Valley has forgotten what normal people want

DGX agent

One of the most mortifying things about knowing a lot of techies is listening to them tell me excitedly about some very important discovery that they believe they have made. Recently, I ran into an ac

hardwarethe-verge-ai
20 Apr 2026
Hardware

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

DGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

DGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

At SemiAnalysis, we're quite tired of the minimalist style webapps and landing pages that have become so commonplace lately. Today, we're in…

DGX agent

At SemiAnalysis, we're quite tired of the minimalist style webapps and landing pages that have become so commonplace lately. Today, we're introducing Minecraft mode on inferencex dot com so that you c

hardwaredylan-patel--x
17 Apr 2026
← Previous
1…678910…93
Next →