AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
Hardware

Choosing the right orchestration layer for your AI use cases

DGX agent

Compute scarcity is not the only struggle AI teams face. Optimal utilization is also key. Friction also emerges when they outgrow informal coordination methods such as shared spreadsheets, manual SSH

hardwarelambda-labs
4 Aug 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

early sparks of rsi? Ali Taha @waterloo_intern explains how http://Z.ai's GLM-5.2 (@Zai_org) profiled its SGLang serving path and rewrote bo…

DGX agent

early sparks of rsi? Ali Taha @waterloo_intern explains how http://Z.ai's GLM-5.2 (@Zai_org) profiled its SGLang serving path and rewrote bottlenecked GPU kernels, (still lacks reliable judgment) grea

hardwareswyx--x
4 Aug 2026
Hardware

Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super

DGX agent

NVIDIA Alpamayo 2 Super is a publicly available 34‑billion‑parameter vision–language–action model that merges a 32‑B Cosmos 3 Super Reasoner with a 2‑B action‑expert diffusion network. It produces uni

hardwarenvidia-developer
4 Aug 2026
Hardware

Has anyone been working on a solid setup for DSV4F on x2+ R9700s?

DGX agent

I'm hoping that one of you guys has been working on an inference engine or has somehow found improvements to running DSV4F on RDNA4 multi-GPU setups. I am currently building a custom inference engine

hardwarer-localllama
4 Aug 2026
Hardware

HetRoute Heterogeneous and Cost-aware Collaborative Routing Framework for Distributed Edge MoE Inference

DGX agent

arXiv:2608.00577v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have become a dominant architecture for large-scale AI services, yet deploying them over geo-distributed heterogeneous

hardwarearxiv-cs-cl
4 Aug 2026
Hardware

MiniWorld: Democratizing the Training of Video World Models from Scratch

DGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

hardwarearxiv-cs-cv
4 Aug 2026
Hardware

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

DGX agent

For robotaxis and other autonomous vehicles (AVs), the hardest problems aren’t the everyday scenarios. They’re the rare, complex situations that are difficult to anticipate and train for. Handling the

hardwarenvidia-blog
4 Aug 2026
Hardware

NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US

DGX agent

NVIDIA is participating in the U.S. National Science Foundation’s (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanc

hardwarenvidia-blog
4 Aug 2026
Hardware

Nvidia open sources cuFile API, accelerating GPU read/write capability for high-speed storage

DGX agent

As artificial intelligence applications become ever hungrier for faster access to data, Nvidia Corp. today announced it is open-sourcing the application programming interface for its powerful cuFile v

hardwaresiliconangle
4 Aug 2026
Hardware

Open Secure AI Alliance proposes SAFE guidelines as membership tops 120

DGX agent

The Open Secure AI Alliance today proposed a set of guidelines for reporting cybersecurity incidents involving artificial intelligence agents, one week after the group was formed. The proposal is call

hardwaresiliconangle
4 Aug 2026
Hardware

Partial FC: Training 10 Million Identities on a Single Machine

DGX agent

arXiv:2010.05222v3 Announce Type: replace Abstract: Training face recognition models with millions of identities is challenging because classifier storage, logit memory, and computation grow linearly

hardwarearxiv-cs-cv
4 Aug 2026
Hardware

Sources: Anthropic agreed to a $10B deal for computing capacity in Norway from Nvidia-backed AI cloud startup Volta Infra, which says the deal is for six years (Bloomberg)

DGX agent

Bloomberg: Sources: Anthropic agreed to a 10B deal for computing capacity in Norway from Nvidia-backed AI cloud startup Volta Infra, which says the deal is for six years — Anthropic PBC has struck a 1

hardwaretechmeme
4 Aug 2026
Hardware

SparseKAN: Compressing Kolmogorov--Arnold Networks Across Basis Functions, Neurons, and Bits

DGX agent

arXiv:2608.00859v1 Announce Type: new Abstract: Kolmogorov--Arnold Networks (KANs) replace scalar edge weights with learnable univariate functions parameterized by multiple basis coefficients. This in

hardwarearxiv-cs-lg
4 Aug 2026
Hardware

Stipple: Real-Time Incremental Gaussian Splatting with Visual-Inertial Tracking

DGX agent

arXiv:2608.00931v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) provides efficient rendering of photo-realistic scenes, but its heavy preprocessing and training steps make it a poor fit

hardwarearxiv-cs-cv
4 Aug 2026
Hardware

Toward Geometry-Scalable Whole-Body Touch for Humanoids: A 3D-Printed Conformal EIT Skin

DGX agent

arXiv:2608.02080v1 Announce Type: new Abstract: Whole-body tactile sensing is a prerequisite for humanoids that operate in contact-rich human environments, but conventional taxel arrays scale poorly w

hardwarearxiv-cs-ro
4 Aug 2026
Hardware

DeltaServe: Host-Agnostic Co-Serving of Inference and Fine-Tuning for LLMs

DGX agent

arXiv:2607.28848v1 Announce Type: cross Abstract: LLM serving systems are provisioned for peak load to meet strict latency targets, leaving substantial GPU compute idle whenever traffic falls below pe

hardwarearxiv-cs-lg
3 Aug 2026
Hardware

Event-Based Upper-Body Humanoid Teleoperation Under Challenging Illumination

DGX agent

arXiv:2607.29227v1 Announce Type: new Abstract: We present a real-time upper-body human-to-humanoid motion imitation framework driven by neuromorphic event-based vision. This work addresses practical

hardwarearxiv-cs-ro
3 Aug 2026
Hardware

GPU-Accelerated ANNS: Quantized for Speed, Built for Change

DGX agent

arXiv:2601.07048v5 Announce Type: replace-cross Abstract: Approximate nearest neighbor search (ANNS) is a core problem in machine learning and information retrieval applications. GPUs offer a promisin

hardwarearxiv-cs-ai
3 Aug 2026
Hardware

How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure

DGX agent

A combined architecture using KAI Scheduler and vCluster enables multiple teams to run fully isolated Kubernetes tenant clusters on a shared GPU node, each with its own control plane, RBAC, CRDs, and

hardwarenvidia-developer
3 Aug 2026
Hardware

Implicit Machine Learning Force Fields Accelerate Molecular Dynamics Simulations

DGX agent

arXiv:2607.29158v1 Announce Type: cross Abstract: We introduce implicit machine learning force fields (I-MLFFs), which replace explicit stacks of neural network layers with self-consistent fixed-point

hardwarearxiv-cs-ai
3 Aug 2026
Hardware

London-based AI chip startup Olix raised 312M led by Fundomo, with participation from Arm, at a 3.3B valuation, up from 1B+ after raising 220M in February (Tim Bradshaw/Financial Times)

DGX agent

Tim Bradshaw / Financial Times: London-based AI chip startup Olix raised 312M led by Fundomo, with participation from Arm, at a 3.3B valuation, up from 1B+ after raising 220M in February — British ent

hardwaretechmeme
3 Aug 2026
Hardware

NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage

DGX agent

NVIDIA’s Vera BlueField‑4 STX Storage Processor combines an 88‑core Armv9.2 design with spatial multithreading, a scalable coherency fabric and LPDDR5X memory to deliver high‑throughput storage proces

hardwarenvidia-developer
3 Aug 2026
Hardware

OsteoCAD: A Human-in-the-Loop Cloud-Edge Framework for Bone Tumor Segmentation

DGX agent

arXiv:2607.29266v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Deep Learning (DL) have notably advanced medical image analysis, yet many health- care organizations struggle to adop

hardwarearxiv-cs-ai
3 Aug 2026
Hardware

Point2Radio: A Foundation Model for Cross-Scene Radio Fields from Material-Aware Point Clouds

DGX agent

arXiv:2607.28994v1 Announce Type: cross Abstract: High-fidelity radio fields are typically simulated for every scene--transmitter configuration or fitted separately to each scene, failing to exploit p

hardwarearxiv-cs-ai
3 Aug 2026
Hardware

Studying quantization trade-offs for efficient inference deployment in machine translation

DGX agent

arXiv:2607.29397v1 Announce Type: new Abstract: Deploying large language models in realistic server environments poses challenges, as the system needs to provide high-quality responses with low latenc

hardwarearxiv-cs-cl
3 Aug 2026
Hardware

The Inference Engineering Masterclass: 10x faster models, quantization, speculative decoding, Rubin, & self-optimizing AI https://www.latent…

DGX agent

The Inference Engineering Masterclass: 10x faster models, quantization, speculative decoding, Rubin, & self-optimizing AI https://www.latent.space/p/inference-eng @Baseten @philipkiely and @waterloo_i

hardwareswyx--x
3 Aug 2026
Hardware

The math still ain’t mathing.

DGX agent

The math still ain’t mathing. Recently we estimated global (ex China) AI revenues of around 200bn annualised, based on four different sources. Just come across a fifth, based on Nvidia inference sales

hardwaregary-marcus--x
3 Aug 2026
Hardware

TokTier: Exact Stateful Tokenization for Agentic LLM Serving

DGX agent

arXiv:2607.29678v1 Announce Type: new Abstract: LLM serving systems cache prompt KV state, yet most front ends still re-tokenize the full request text on every call. The cost lands on coding agents, w

hardwarearxiv-cs-cl
3 Aug 2026
Hardware

Topology-Aware Data Movement for Disaggregated GPU Inference

DGX agent

arXiv:2607.28633v1 Announce Type: cross Abstract: Disaggregated LLM inference creates a datacenter networking problem that no existing system solves correctly. When prefill and decode run on separate

hardwarearxiv-cs-ai
3 Aug 2026
Hardware

Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models! With the help of @nvidia's Nemo Relay and sever…

DGX agent

Hermes Agent is now dramatically more efficient, especially for smaller/weaker/local models! With the help of @nvidia's Nemo Relay and several other strategies Hermes was able to identify a ton of opt

hardwarenous-research--x
2 Aug 2026
Hardware

Hugging Face CEO Clément Delague says “AI is actually an opportunity to fix a lot of the cybersecurity problems” because his company used Nv…

DGX agent

Hugging Face CEO Clément Delague says “AI is actually an opportunity to fix a lot of the cybersecurity problems” because his company used Nvidia’s version of a Chinese open model to defend itself agai

hardwareclem-delangue--x
2 Aug 2026
Hardware

I pushed Kimi K3 onto one CPU with 8 GB of RAM

DGX agent

I deployed K3 on 32 H100s at work a couple of weeks ago and then got annoyed that there was no way to poke at it on my own machine. So I wrote an inference engine for it in C99. Nothing clever going o

hardwarer-localllama
2 Aug 2026
Hardware

The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were …

DGX agent

The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were built for. It's illegal to collude and slow down AI progress

hardwaredylan-patel--x
2 Aug 2026
Hardware

Watching @ClementDelangue on @FaceTheNation discussing agentic hacking. “Preventing releases does not work; concentrating behind closed door…

DGX agent

Watching @ClementDelangue on @FaceTheNation discussing agentic hacking. “Preventing releases does not work; concentrating behind closed doors in just a few organizations doesnt work. What worked in th

hardwareclem-delangue--x
2 Aug 2026
Hardware

Banning open source model, will 'remove capabilities for the defenders' When OpenAI's models breached Huggingface, a Chinese open model is w…

DGX agent

Banning open source model, will 'remove capabilities for the defenders' When OpenAI's models breached Huggingface, a Chinese open model is what cleaned up the mess. Anthropic's Fable 5 refused, so Hug

hardwareclem-delangue--x
1 Aug 2026
Hardware

Vivid depiction of a dark scenario for Nvidia. “Indeed, CEO Jensen Huang has been so aggressive in funding the AI ecosystem that the company…

DGX agent

Vivid depiction of a dark scenario for Nvidia. “Indeed, CEO Jensen Huang has been so aggressive in funding the AI ecosystem that the company is functioning almost like a bank; with annual free cash fl

hardwaregary-marcus--x
1 Aug 2026
Hardware

A Query-Efficient Stochastic Volume Rendering Framework for Time-Varying Implicit Neural Volumes

DGX agent

arXiv:2607.28047v1 Announce Type: cross Abstract: Time-varying implicit neural representations (INRs) provide a compact representation of scientific volumes and, for modalities such as dynamic X-ray c

hardwarearxiv-cs-lg
31 Jul 2026
Hardware

AgenticCANN: Automated Ascend C Operator Generation via Knowledge-Augmented Agentic Evolution

DGX agent

arXiv:2607.26661v1 Announce Type: new Abstract: Ascend C operator optimization is critical for NPU (Neural Processing Unit) inference performance but requires deep hardware expertise.While large langu

hardwarearxiv-cs-ai
31 Jul 2026
Hardware

Autoscaling endpoints for LLM inference

DGX agent

GPU utilization can read healthy while your queue backs up, and a new replica takes minutes to warm. Here's how to pick autoscaling metrics, tune scale-up/down windows, and budget for cold starts on d

hardwaretogether-ai-blog
31 Jul 2026
Hardware

Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference

DGX agent

The article shows that dense‑attention performance in long‑context inference is governed by group size (query heads per KV head), head dimension, and sequence length, with prefill being compute‑bound

hardwarenvidia-developer
31 Jul 2026
Hardware

Complementary Matrix-Gated QKAN Fast-Weight Programmers for Quantum Dynamics Forecasting

DGX agent

arXiv:2607.27945v1 Announce Type: cross Abstract: Sequence models must decide what to write into memory and what to retain. In quantum and quantum-inspired sequence learning, nonlinear recurrent updat

hardwarearxiv-cs-lg
31 Jul 2026
Hardware

FARI: Robust One-Step Inversion for Watermarking in Diffusion Models

DGX agent

arXiv:2607.26723v1 Announce Type: cross Abstract: Inversion-based watermarking is a promising approach to authenticate diffusion-generated images, yet practical use is bottlenecked by inversion that i

hardwarearxiv-cs-ai
31 Jul 2026
Hardware

FasTac: A Curved Multispectral Vision-Based Tactile Sensor for High-Speed High-Precision 3D Shape and Force Perception

DGX agent

arXiv:2607.28416v1 Announce Type: new Abstract: Curved tactile fingertips for dexterous manipulation must resolve fine contact geometry, distinguish normal and tangential loads, and capture transient

hardwarearxiv-cs-ro
31 Jul 2026
Hardware

FinCacheServe: Dependency-Consistent Answer Reuse for Cost-Efficient RAG Serving over Mutable Enterprise Documents

DGX agent

arXiv:2607.26076v1 Announce Type: cross Abstract: Retrieval-augmented generation services over mutable enterprise documents repeatedly execute semantically equivalent analysis requests. Answer reuse c

hardwarearxiv-cs-ai
31 Jul 2026
Hardware

GoGoTB: Agentic RTL Verification with Specification-Grounded Coverage Closure

DGX agent

arXiv:2607.26181v1 Announce Type: new Abstract: Functional verification dominates integrated circuit (IC) front-end engineering effort, and a single missed bug that escapes to silicon can trigger a co

hardwarearxiv-cs-ai
31 Jul 2026
Hardware

May have found the highest and best use case of Flux 3 - generating GPU ASMR ✨ For everyone who has been asking for access, it’s available N…

DGX agent

Justine Moore announced on July 31, 2026 that Flux 3’s latest iteration excels at generating GPU‑based ASMR content. She confirmed that this capability is now available in an early preview on the Nous

hardwarenous-research--x
31 Jul 2026
Hardware

Nscale buys AI infrastructure optimization startup Anyscale for reported $1.65B

DGX agent

Data center builder Nscale Global Holdings Ltd. today announced plans to acquire Anyscale Inc., a venture-backed provider of artificial intelligence software. The terms of the deal were not disclosed.

hardwaresiliconangle
31 Jul 2026
Hardware

NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek

DGX agent

NVIDIA Video Codec SDK 13.1 adds AV1 hierarchical reference mode supporting up to 31 B‑frames and efficient iterative tuning that delivers significant bitrate savings in CQ and VBR modes. It enhances

hardwarenvidia-developer
31 Jul 2026
← Previous
12345…37
Next →