AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
Hardware

Semantic-Aware Adaptive Visual Memory for Streaming Video Understanding

DGX agent

arXiv:2605.07897v1 Announce Type: cross Abstract: Online streaming video understanding requires models to process continuous visual inputs and respond to user queries in real time, where the unbounded

hardwarearxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

SemanticDialect: Semantic-Aware Mixed-Format Quantization for Video Diffusion Transformers

DGX agent

arXiv:2603.02883v3 Announce Type: replace Abstract: Diffusion Transformers (DiTs) achieve state-of-the-art video generation quality, but their substantial memory and computational footprints hinder ed

hardwarearxiv-cs-cv
11 May 2026
Hardware

🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architec…

DGX agent

🧵 Slime: The Most Elegant & Comfortable RL Training Framework Ever A deep dive into why Slime redefines LLM RL training with clean architecture & production-grade engineering ✨ Insights from Zhihu con

hardwarezhipu-ai--x
11 May 2026
Hardware

SOCKET: SOft Collision Kernel EsTimator for Sparse Attention

DGX agent

arXiv:2602.06283v2 Announce Type: replace Abstract: Exploiting sparsity during long-context inference is key to scaling large language models, as attention dominates the cost of autoregressive decodin

hardwarearxiv-cs-lg
11 May 2026
Hardware

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache

DGX agent

arXiv:2605.06763v1 Announce Type: new Abstract: Sparse attention improves LLM inference efficiency by selecting a subset of key-value entries, but at the cost of potential accuracy degradation. In par

hardwarearxiv-cs-lg
11 May 2026
Hardware

Sparser, Faster, Lighter Transformer Language Models

DGX agent

arXiv:2603.23198v2 Announce Type: replace-cross Abstract: Scaling autoregressive large language models (LLMs) has driven unprecedented progress but comes with vast computational costs. In this work, w

hardwarearxiv-cs-cl
11 May 2026
Hardware

SpikingBrain: Spiking Brain-inspired Large Models

DGX agent

arXiv:2509.05276v4 Announce Type: replace-cross Abstract: Mainstream Transformer-based large language models face major efficiency bottlenecks: training computation scales quadratically with sequence

hardwarearxiv-cs-ai
11 May 2026
Hardware

Three things in AI to watch, according to a Nobel-winning economist

DGX agent

This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. A few months before he was awarded the Nobel Prize in economic

hardwaremit-tech-review
11 May 2026
Hardware

Towards Billion-scale Multi-modal Biometric Search

DGX agent

arXiv:2605.07655v1 Announce Type: cross Abstract: Searching a multi-biometric database of a billion records for a country-level identity system requires pushing the limits of all aspects of a biometri

hardwarearxiv-cs-ai
11 May 2026
Hardware

Versatile yet Efficient Network Traffic Analysis: Offloading Network Foundation Model to SmartNIC

DGX agent

arXiv:2508.02001v2 Announce Type: replace-cross Abstract: Pervasive encryption makes large-scale labeling infeasible for traffic analysis, while security operations demand edge analysis to avert servi

hardwarearxiv-cs-lg
11 May 2026
Hardware

XiYOLO: Energy-Aware Object Detection via Iterative Architecture Search and Scaling

DGX agent

arXiv:2605.06927v1 Announce Type: cross Abstract: Object detection on heterogeneous edge devices must satisfy strict energy, latency, and memory constraints while still providing reliable perception f

hardwarearxiv-cs-ai
11 May 2026
Hardware

Zero-Shot Neural Network Evaluation with Sample-Wise Activation Patterns

DGX agent

arXiv:2605.07378v1 Announce Type: new Abstract: Zero-shot proxies, also known as training-free metrics, are widely adopted to reduce the computational overhead in neural network evaluation for scenari

hardwarearxiv-cs-lg
11 May 2026
Hardware

Arm, the UK and Apple

DGX agent

This article likely examines the relationship between Arm Holdings (the British semiconductor design company), the UK government, and Apple, possibly covering topics such as Apple's use of Arm-based c

hardwarethe-chip-letter
10 May 2026
Hardware

I was talking to a room of senior accountants a couple months ago and 10% had OpenClaw installations. Of course there are far more non-users…

DGX agent

I was talking to a room of senior accountants a couple months ago and 10% had OpenClaw installations. Of course there are far more non-users and firms lag behind their people, but there is a sort of S

hardwareethan-mollick--x
10 May 2026
Hardware

Nvidia, AI factories and the transition to accelerated computing

DGX agent

The market is trying to price a transition it hasn’t fully internalized. It sees Nvidia Corp.’s market cap with a five-handle and assumes the valuation is too high to grow further. We believe that’s t

hardwaresiliconangle
10 May 2026
Hardware

‘Your Career Starts at the Beginning of the AI Revolution,’ NVIDIA CEO Tells Graduates

DGX agent

“You are entering the world at an extraordinary moment,” NVIDIA founder and CEO Jensen Huang told graduates as he delivered the keynote address at Carnegie Mellon University’s 128th commencement cerem

hardwarenvidia-blog
10 May 2026
Hardware

Yann LeCun closed $1.03B for AMI Labs on March 10. Three days later, this paper dropped from his NYU collaborators. 15M parameters. Single G…

DGX agent

Yann LeCun closed 1.03B for AMI Labs on March 10. Three days later, this paper dropped from his NYU collaborators. 15M parameters. Single GPU. A few hours of training. LeWorldModel is the first JEPA t

hardwareyann-lecun--x
9 May 2026
Hardware

AI data center firm IREN’s stock soars after it strikes $2.1B deal with Nvidia

DGX agent

Neocloud company IREN Ltd. has secured a 2.1 billion commitment from the chipmaker Nvidia Corp. as part of a new data center partnership aimed at artificial intelligence workloads. The partners plan t

hardwaresiliconangle
8 May 2026
Hardware

Deploy and inference any model from HuggingFace

DGX agent

Learn how to deploy any Hugging Face model in one session using Goose and Together's Dedicated Container Inference. Skip the setup complexity — one prompt gets your model running in a production-grade

hardwaretogether-ai-blog
8 May 2026
Hardware

Great collab with @SakanaAILabs on an #ICML26 paper about sparse transformer kernels + formats optimized for modern NVIDIA GPU execution. • …

DGX agent

Great collab with @SakanaAILabs on an #ICML26 paper about sparse transformer kernels + formats optimized for modern NVIDIA GPU execution. • TwELL sparse packing • Fused CUDA kernels • 20%+ inference/t

hardwaredavid-ha--x
8 May 2026
Hardware

Improving Bash Generation in Small Language Models with Grammar-Constrained Decoding

DGX agent

Grammar-constrained decoding modifies language model generation by applying grammar constraints at each step to block structurally invalid tokens , ensuring syntactically correct Bash command generati

hardwarenvidia-developer
8 May 2026
Hardware

OPINION: Ever since private equity bought the @Veritasium YouTube channel, the content just hasn't been as good. Used to be a GREAT channel,…

DGX agent

OPINION: Ever since private equity bought the @Veritasium YouTube channel, the content just hasn't been as good. Used to be a GREAT channel, maybe the best Science channel on YouTube, and now? SAD! Ve

hardwaredylan-patel--x
8 May 2026
Hardware

Tickets are now available for the 4th annual AI Film Festival: June 11 at Alice Tully Hall at Lincoln Center in NYC, and June 18 at The Broa…

DGX agent

Tickets are now available for the 4th annual AI Film Festival: June 11 at Alice Tully Hall at Lincoln Center in NYC, and June 18 at The Broad Stage in LA. It’s the biggest and most important celebrati

hardwarecristobal-valenzuela--x
8 May 2026
Hardware

With faster node startup for GKE, say goodbye to cold-start latency

DGX agent

We’ve rolled out a significant update to Google Kubernetes Engine (GKE) that solves one of the most annoying problems in cloud infrastructure: cold start latency. GKE now has up to 4x faster node star

hardwaregoogle-cloud-ai
8 May 2026
Hardware

A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers

DGX agent

arXiv:2605.04074v1 Announce Type: new Abstract: AI data centers experience rapid fluctuations in power demand due to the heterogeneity of computational tasks that they have to support. For example, th

hardwarearxiv-cs-lg
7 May 2026
Hardware

A Queueing-Theoretic Framework for Stability Analysis of LLM Inference with KV Cache Memory Constraints

DGX agent

arXiv:2605.04595v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) has created significant challenges for efficient inference at scale. Unlike traditional workloads, LL

hardwarearxiv-cs-lg
7 May 2026
Hardware

Achieving Peak System and Workload Efficiency on NVIDIA GB200 NVL72 with Slurm Block Scheduling

DGX agent

This article discusses optimization techniques for maximizing efficiency on NVIDIA's GB200 NVL72 system using Slurm block scheduling, a job scheduling approach designed to improve resource utilization

hardwarenvidia-developer
7 May 2026
Hardware

Budget-aware Auto Optimizer Configurator

DGX agent

arXiv:2605.04711v1 Announce Type: cross Abstract: Optimizer states occupy massive GPU memory in large-scale model training. However, gradients in different network blocks exhibit distinct behaviors, s

hardwarearxiv-cs-lg
7 May 2026
Hardware

Coral: Cost-Efficient Multi-LLM Serving over Heterogeneous Cloud GPUs

DGX agent

arXiv:2605.04357v1 Announce Type: cross Abstract: The usage of large language models (LLMs) has grown increasingly fragmented, with no single model dominating. Meanwhile, cloud providers offer a wide

hardwarearxiv-cs-cl
7 May 2026
Hardware

CuBridge: An LLM-Based Framework for Understanding and Reconstructing High-Performance Attention Kernels

DGX agent

arXiv:2605.05023v1 Announce Type: new Abstract: Efficient CUDA implementations of attention mechanisms are critical to modern deep learning systems, yet supporting diverse and evolving attention varia

hardwarearxiv-cs-lg
7 May 2026
Hardware

Deep Wave Network for Modeling Multi-Scale Physical Dynamics

DGX agent

arXiv:2605.04198v1 Announce Type: new Abstract: Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-scien

hardwarearxiv-cs-lg
7 May 2026
Hardware

DualTCN: A Physics-Constrained Temporal Convolutional Network for 2 Time-Domain Marine CSEM Inversion

DGX agent

arXiv:2605.04997v1 Announce Type: new Abstract: DualTCN is the first deep-learning framework for inverting time-domain marine controlled-source electromagnetic (MCSEM) transient data. Moving away from

hardwarearxiv-cs-lg
7 May 2026
Hardware

Hot take on Elon’s surprise decision to rent 30 megawatts of compute to Anthropic: 1. It’s a tacit concession that xAI is not all that close…

DGX agent

Hot take on Elon’s surprise decision to rent 30 megawatts of compute to Anthropic: 1. It’s a tacit concession that xAI is not all that close to AGI (despite what he suggested last year). 2. It’s more

hardwaregary-marcus--x
7 May 2026
Hardware

.@huggingface's agentic robotics app store for Reachy Mini is a big step toward more accessible physical AI. 🙌 Excited to see NVIDIA Isaac …

DGX agent

.@huggingface's agentic robotics app store for Reachy Mini is a big step toward more accessible physical AI. 🙌 Excited to see NVIDIA Isaac GR00T N integrated with Hugging Face LeRobot, helping develop

hardwareclem-delangue--x
7 May 2026
Hardware

I wish you could filter resteruant reviews and rankings on Google maps searches for by race Gotta authentic ethnic food max Too much overly …

DGX agent

This post discusses a request for filtering functionality on Google Maps that would allow users to search for restaurants by ethnic cuisine type or authenticity level, suggesting a desire for better t

hardwaredylan-patel--x
7 May 2026
Hardware

Lambda closes $1 billion senior secured credit facility to meet gigawatt-scale AI infrastructure demand

DGX agent

Lambda Labs secured a $1 billion senior secured credit facility to fund the expansion of its AI infrastructure and meet growing demand for gigawatt-scale computational resources. This financing enable

hardwarelambda-labs
7 May 2026
Hardware

Linked and Loaded: Gaijin Single Sign-On Now Available on GeForce NOW

DGX agent

Less typing, more tanking. Faster logins mean more time in the gaming action — and this week provides GeForce NOW members with a smoother path straight into the battlefield. Cloud gaming is all about

hardwarenvidia-blog
7 May 2026
Hardware

Massively Parallel Exact Inference for Hawkes Processes

DGX agent

arXiv:2604.01342v2 Announce Type: replace Abstract: Multivariate Hawkes processes are a widely used class of self-exciting point processes, but maximum likelihood estimation naively scales as O(N^2) i

hardwarearxiv-cs-lg
7 May 2026
Hardware

Moonshot, the Chinese AI startup behind Kimi chatbot, raised ~2B at a 20B+ valuation led by Meituan's venture arm; Moonshot's ARR topped $200M in April 2026 (Zheping Huang/Bloomberg)

DGX agent

Zheping Huang / Bloomberg: Moonshot, the Chinese AI startup behind Kimi chatbot, raised ~2B at a 20B+ valuation led by Meituan's venture arm; Moonshot's ARR topped 200M in April 2026 — Moonshot AI has

hardwaretechmeme
7 May 2026
Hardware

Nvidia and data center operator IREN announce a deal to deploy up to 5 GW of AI infrastructure; Nvidia can invest $2.1B into IREN; IREN jumps 9%+ after hours (Jonathan Vanian/CNBC)

DGX agent

Jonathan Vanian / CNBC: Nvidia and data center operator IREN announce a deal to deploy up to 5 GW of AI infrastructure; Nvidia can invest $2.1B into IREN; IREN jumps 9%+ after hours — IREN shares surg

hardwaretechmeme
7 May 2026
Hardware

Nvidia up 2.6% on news that xAI is flooding the market with 220,000 secondhand GPUs.

DGX agent

Nvidia's stock rose 2.6% following news that xAI is selling 220,000 used GPUs on the secondhand market, potentially indicating a shift in AI hardware demand or xAI's operational priorities. The report

hardwaregary-marcus--x
7 May 2026
Hardware

OptiLookUp: An Optical ROM-Based Loop up Table Engine for Photonic Accelerators

DGX agent

arXiv:2605.03241v1 Announce Type: cross Abstract: Read-only memory (ROM) provides deterministic access to predefined data mappings. Extending ROM concepts to the optical domain enables high-bandwidth,

hardwarearxiv-cs-ai
7 May 2026
Hardware

Piper: Efficient Large-Scale MoE Training via Resource Modeling and Pipelined Hybrid Parallelism

DGX agent

arXiv:2605.05049v1 Announce Type: cross Abstract: Frontier models increasingly adopt Mixture-of-Experts (MoE) architectures to achieve large-model performance at reduced cost. However, training MoE mo

hardwarearxiv-cs-lg
7 May 2026
Hardware

Powering the Next American Century: US Energy Secretary Chris Wright and NVIDIA’s Ian Buck on the Genesis Mission

DGX agent

AI will help build the energy it needs. That’s the case U.S. Energy Secretary Chris Wright and NVIDIA Vice President of Hyperscale and High-Performance Computing Ian Buck made Thursday morning at the

hardwarenvidia-blog
7 May 2026
Hardware

Quadrature-TreeSHAP: Depth-Independent TreeSHAP and Shapley Interactions

DGX agent

arXiv:2605.04497v1 Announce Type: new Abstract: Shapley values are a standard tool for explaining predictions of tree ensembles, with Path-Dependent SHAP being the most widely used variant. Despite su

hardwarearxiv-cs-lg
7 May 2026
Hardware

Quantum Motion raises $160M to build faster quantum chips

DGX agent

Quantum computer maker Quantum Motion Ltd. today announced that it has raised 160 million in funding to enhance its silicon-based qubit technology. The Series C round was led by DCVC and Kembara. It c

hardwaresiliconangle
7 May 2026
Hardware

Real-Time Performance Monitoring and Faster Debugging with NCCL Inspector and Prometheus

DGX agent

NCCL Inspector is a profiling plugin that provides detailed, per-communicator, per-collective performance and metadata logging, designed to help users analyze and debug NCCL collective operations by g

hardwarenvidia-developer
7 May 2026
Hardware

Secure short-term GPU capacity for ML workloads with EC2 Capacity Blocks for ML and SageMaker training plans

DGX agent

In this post, you will learn how to secure reserved GPU capacity for short-term workloads using Amazon Elastic Compute Cloud (Amazon EC2) Capacity Blocks for ML and Amazon SageMaker training plans. Th

hardwareaws-ml-blog
7 May 2026
← Previous
1…2526272829…37
Next →