AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
Hardware

POLAR: Online Learning for LoRA Adapter Caching and Routing in Edge LLM Serving

DGX agent

arXiv:2604.16583v1 Announce Type: new Abstract: Edge deployment of large language models (LLMs) increasingly relies on libraries of lightweight LoRA adapters, yet GPU/DRAM can keep only a small reside

hardwarearxiv-cs-lg
21 Apr 2026
Hardware
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Probabilistic Programs of Thought

DGX agent

arXiv:2604.17290v1 Announce Type: new Abstract: LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to rea

hardwarearxiv-cs-cl
21 Apr 2026
Hardware

RACE Attention: A Strictly Linear-Time Attention Layer for Training on Outrageously Large Contexts

DGX agent

arXiv:2510.04008v5 Announce Type: replace Abstract: Softmax Attention has a quadratic time complexity in sequence length, which becomes prohibitive to run at long contexts, even with highly optimized

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

RainFusion2.0: Temporal-Spatial Awareness and Hardware-Efficient Block-wise Sparse Attention

DGX agent

arXiv:2512.24086v2 Announce Type: replace Abstract: In video and image generation tasks, Diffusion Transformer (DiT) models incur extremely high computational costs due to attention mechanisms, which

hardwarearxiv-cs-cv
21 Apr 2026
Hardware

Real-Time Structural Detection for Indoor Navigation from 3D LiDAR Using Bird's-Eye-View Images

DGX agent

arXiv:2603.19830v2 Announce Type: replace Abstract: Efficient structural perception is essential for mapping and autonomous navigation on resource-constrained robots. Existing 3D methods are computati

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

Scalable and Adaptive Parallel Training of Graph Transformer on Large Graphs

DGX agent

arXiv:2604.16715v1 Announce Type: cross Abstract: Graph foundation models have demonstrated remarkable adaptability across diverse downstream tasks through large-scale pretraining on graphs. However,

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

DGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

hardwarearxiv-cs-lg
21 Apr 2026
Hardware

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’…

DGX agent

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’s leading product and distribution to expert software engine

hardwareelon-musk--x
21 Apr 2026
Hardware

Web-Gewu: A Browser-Based Interactive Playground for Robot Reinforcement Learning

DGX agent

arXiv:2604.17050v1 Announce Type: new Abstract: With the rapid development of embodied intelligence, robotics education faces a dual challenge: high computational barriers and cumbersome environment c

hardwarearxiv-cs-ro
21 Apr 2026
Hardware

AutoFlows++: Hierarchical Message Flow Mining for System on Chip Designs

DGX agent

arXiv:2604.15359v1 Announce Type: cross Abstract: Understanding communication behavior in modern system-on-chip (SoC) designs is critical for functional verification, performance analysis, and post-si

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Autonomous AI at Scale: Adobe Agents Unlock Breakthrough Creative Intelligence With NVIDIA and WPP

DGX agent

AI agents are transforming how work gets done across all industries, accelerating everything from content creation to decision-making. NVIDIA’s expanded strategic collaborations with Adobe and WPP are

hardwarenvidia-blog
20 Apr 2026
Hardware

Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~$2.6B in its IPO, the largest in the city so far this year (Eunice Xu/South China Morning Post)

DGX agent

Eunice Xu / South China Morning Post: Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~2.6B in its IPO, the largest in the city so far this year — Nvidia suppl

hardwaretechmeme
20 Apr 2026
Hardware

CPU Optimization of a Monocular 3D Biomechanics Pipeline for Low-Resource Deployment

DGX agent

arXiv:2604.15665v1 Announce Type: new Abstract: Markerless 3D movement analysis from monocular video enables accessible biomechanical assessment in clinical and sports settings. However, most research

hardwarearxiv-cs-cv
20 Apr 2026
Hardware

How Much Do GPU Clusters Really Cost?

DGX agent

This article from SemiAnalysis analyzes the total cost of ownership for GPU clusters, likely covering hardware expenses, infrastructure requirements, power consumption, and operational overhead beyond

hardwaresemianalysis
20 Apr 2026
Hardware

Johny Sroji, the GOAT who made Apple silicon program what it is, is now in charge of all hardware at Apple as Chief Hardware Officer

DGX agent

Johny Sroji, the GOAT who made Apple silicon program what it is, is now in charge of all hardware at Apple as Chief Hardware Officer Media Johny Srouji Taking Over as Apple's Chief Hardware Officer as

hardwaredylan-patel--x
20 Apr 2026
Hardware

Lossless Compression via Chained Lightweight Neural Predictors with Information Inheritance

DGX agent

arXiv:2604.15472v1 Announce Type: cross Abstract: This paper is dedicated to lossless data compression with probability estimation using neural networks. First, we propose a probability estimation arc

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Mitigating Indirect AGENTS.md Injection Attacks in Agentic Environments

DGX agent

NVIDIA researchers discovered a vulnerability in AI coding assistants where malicious software dependencies can inject harmful instructions into AGENTS.md configuration files, allowing attackers to re

hardwarenvidia-developer
20 Apr 2026
Hardware

NeuroMesh: A Unified Neural Inference Framework for Decentralized Multi-Robot Collaboration

DGX agent

arXiv:2604.15475v1 Announce Type: new Abstract: Deploying learned multi-robot models on heterogeneous robots remains challenging due to hardware heterogeneity, communication constraints, and the lack

hardwarearxiv-cs-ro
20 Apr 2026
Hardware

NVIDIA and Partners Showcase the Future of AI-Driven Manufacturing at Hannover Messe 2026

DGX agent

Manufacturing is at an inflection point. Across every major industrial economy, the pressure to do more with less — due to faster design cycles, leaner operations and strain on skilled labor pools — i

hardwarenvidia-blog
20 Apr 2026
Hardware

PyLO: Towards Accessible Learned Optimizers in PyTorch

DGX agent

arXiv:2506.10315v3 Announce Type: replace Abstract: Learned optimizers have been an active research topic over the past decade, with increasing progress toward practical, general-purpose optimizers th

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Scalable Posterior Uncertainty for Flexible Density-Based Clustering

DGX agent

arXiv:2603.03188v2 Announce Type: replace-cross Abstract: We introduce a novel framework for uncertainty quantification in clustering that combines martingale posterior distributions with density-base

hardwarearxiv-cs-lg
20 Apr 2026
Hardware

Silicon Valley has forgotten what normal people want

DGX agent

One of the most mortifying things about knowing a lot of techies is listening to them tell me excitedly about some very important discovery that they believe they have made. Recently, I ran into an ac

hardwarethe-verge-ai
20 Apr 2026
Hardware

SK hynix says it has begun mass production of the 192GB SOCAMM2, a next-gen LPDDR5X low-power DRAM module designed particularly for Nvidia's Vera Rubin (The Korea Herald)

DGX agent

The Korea Herald: SK hynix says it has begun mass production of the 192GB SOCAMM2, a next-gen LPDDR5X low-power DRAM module designed particularly for Nvidia's Vera Rubin — SK hynix Inc. said Monday it

hardwaretechmeme
20 Apr 2026
Hardware

Taming Asynchronous CPU-GPU Coupling for Frequency-aware Latency Estimation on Mobile Edge

DGX agent

arXiv:2604.15357v1 Announce Type: cross Abstract: Precise estimation of model inference latency is crucial for time-critical mobile edge applications, enabling devices to calculate latency margins aga

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

DGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

hardwarearxiv-cs-ai
20 Apr 2026
Hardware

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, …

DGX agent

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, multi-modal hierarchical caching, and prefill-decode disaggr

hardwarekimi-moonshot--x
20 Apr 2026
Hardware

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was th…

DGX agent

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was the most popular model on our training platform. K2.6 on Firew

hardwarefireworks-ai--x
20 Apr 2026
Hardware

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed…

DGX agent

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed with building bigger LLMs. Trillions of parameters. Billion

hardwareyann-lecun--x
20 Apr 2026
Hardware

A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research (Abram Brown/The Information)

DGX agent

Abram Brown / The Information: A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research — For Dylan

hardwaretechmeme
19 Apr 2026
Hardware

A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M in 2026 (Emily Shugerman/The San Francisco ...)

DGX agent

Emily Shugerman / The San Francisco Standard: A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M i

hardwaretechmeme
19 Apr 2026
Hardware

Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models (Qianer Liu/The Information)

DGX agent

Qianer Liu / The Information: Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models — Google is in talk

hardwaretechmeme
19 Apr 2026
Hardware

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge back…

DGX agent

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge backlash against AI, he cobbles together this cope - to try to p

hardwaregary-marcus--x
19 Apr 2026
Hardware

Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection (Reuters)

DGX agent

Reuters: Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection — Hesai (2525.HK), China's leadi

hardwaretechmeme
18 Apr 2026
Hardware

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on…

DGX agent

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on @latentspacepod with Alessio Fanelli (@FanaHOVA) and @swyx

hardwareswyx--x
17 Apr 2026
Hardware

A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the interview (Zvi Mowshowitz/Don't Worry About the Vase)

DGX agent

Zvi Mowshowitz / Don't Worry About the Vase: A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the inter

hardwaretechmeme
17 Apr 2026
Hardware

Accelerate Clean, Modular, Nuclear Reactor Design with AI Physics

DGX agent

NVIDIA released PhysicsNeMo, an open-source AI framework designed to speed up nuclear reactor design through physics-based simulations. The framework uses machine learning-based multiphysics emulators

hardwarenvidia-developer
17 Apr 2026
Hardware

AdaSplash-2: Faster Differentiable Sparse Attention

DGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle

DGX agent

Quantum computing keeps gaining momentum even though practically it’s years away from wide commercialization, and World Quantum Day April 14 provided an excuse for a lot of announcements. Among them:

hardwaresiliconangle
17 Apr 2026
Hardware

At SemiAnalysis, we're quite tired of the minimalist style webapps and landing pages that have become so commonplace lately. Today, we're in…

DGX agent

At SemiAnalysis, we're quite tired of the minimalist style webapps and landing pages that have become so commonplace lately. Today, we're introducing Minecraft mode on inferencex dot com so that you c

hardwaredylan-patel--x
17 Apr 2026
Hardware

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

DGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

cuRoboV2: Dynamics-Aware Motion Generation with Depth-Fused Distance Fields for High-DoF Robots

DGX agent

arXiv:2603.05493v2 Announce Type: replace Abstract: Effective robot autonomy requires motion generation that is safe, feasible, and reactive. Current methods are fragmented: fast planners output physi

hardwarearxiv-cs-ro
17 Apr 2026
Hardware

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

DGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

hardwarearxiv-cs-cl
17 Apr 2026
Hardware

DEEP-GAP: Deep-learning Evaluation of Execution Parallelism in GPU Architectural Performance

DGX agent

arXiv:2604.14552v1 Announce Type: cross Abstract: Modern datacenters increasingly rely on low-power, single-slot inference accelerators to balance performance, energy efficiency, and rack density cons

hardwarearxiv-cs-lg
17 Apr 2026
Hardware

Dell and Nvidia turn AI infrastructure into the new power center of enterprise tech

DGX agent

As artificial intelligence scales, AI infrastructure is becoming the deciding factor in whether enterprise AI delivers real value or stalls. AI is shifting from experimentation to unified systems wher

hardwaresiliconangle
17 Apr 2026
Hardware

Friends, housemates, officemates. Maybe lovers?

DGX agent

This post likely explores the various relationship dynamics and social connections people develop in shared living and working spaces, with a speculative or humorous tone suggested by the 'Maybe lover

hardwaredylan-patel--x
17 Apr 2026
Hardware

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability

DGX agent

arXiv:2604.13048v1 Announce Type: cross Abstract: Modern cloud-native platforms expose thousands of time series metrics through systems like Prometheus, yet formulating correct queries in domain-speci

hardwarearxiv-cs-ai
17 Apr 2026
Hardware

Hangzhou-based Manycore shares rose 187% early in its Hong Kong debut after raising $156M in its IPO; it is pivoting to selling AI training data to robot makers (Bloomberg)

DGX agent

Bloomberg: Hangzhou-based Manycore shares rose 187% early in its Hong Kong debut after raising $156M in its IPO; it is pivoting to selling AI training data to robot makers — Manycore Tech Inc. built a

hardwaretechmeme
17 Apr 2026
← Previous
1…3132333435…37
Next →