AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
Hardware

CARB: A Characterization-Guided Framework for CNN Inference Cost Prediction and Deployment Screening

DGX agent

arXiv:2608.10506v1 Announce Type: cross Abstract: Accurate pre-deployment estimation of CNN inference cost--energy, latency, and peak memory--is increasingly critical as models are deployed on resourc

hardwarearxiv-cs-lg
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

EvoMem: Memory-Augmented Evolution for Code Optimization

DGX agent

arXiv:2608.10795v1 Announce Type: new Abstract: Successful mutation strategies in evolutionary code search may contain reusable knowledge that is useful beyond a single run, and in some cases may tran

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

Ex-Omni-2D: Expressive Omni-Modal Dialogue Models with Native Visual Presence

DGX agent

arXiv:2608.10720v1 Announce Type: new Abstract: Omni-modal dialogue models can understand multimodal inputs and synthesize spoken replies, yet their responses remain visually disembodied. We introduce

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

From Detection to Understanding: TAR and TAR-Bench for Multi-Task Traffic Anomaly Reasoning

DGX agent

arXiv:2608.10317v1 Announce Type: new Abstract: We present TAR (Traffic Anomaly Reasoning) and TAR-Bench datasets, resources for training and evaluating video-language models beyond anomaly detection.

hardwarearxiv-cs-cv
12 Aug 2026
Hardware

Hand-Written PTX Tensor-Core GEMM Kernels: A Multi-Precision Study on NVIDIA L4

DGX agent

arXiv:2608.10103v1 Announce Type: cross Abstract: High-performance Tensor Core kernels rely on a low-level PTX pipeline built from asynchronous data movement with cp.async, warp-level matrix loads wit

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

HyWA: Architecture-Preserving Personalized Voice Activity Detection for Full-Duplex Voice Assistants

DGX agent

arXiv:2510.12947v3 Announce Type: replace-cross Abstract: Voice activity detection (VAD) serves as an early gate in voice-assistant pipelines for smart devices. Because conventional VADs respond to sp

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

IBM inks $240M infrastructure deal with AI-optimized cloud operator Together AI

DGX agent

IBM Corp. will provide infrastructure to artificial intelligence startup Together AI Inc. as part of a 240 million deal announced today. The partnership comes a few weeks after the latter company rais

hardwaresiliconangle
12 Aug 2026
Hardware

Is the future of AI selling hardware for Open Source/Models?

DGX agent

I’m not super knowledgeable of the entire AI industry, but as we see this industry grow and the Cold War that is happening between the US and China on AI development, I can’t help but notice what US c

hardwarer-localllama
12 Aug 2026
Hardware

Neuroevolution Arena: Nested Ecological Evaluation of Update-and-Inheritance Regimes across Neural Architectures

DGX agent

arXiv:2608.10323v1 Announce Type: new Abstract: Competitive artificial-life systems can rank trained controllers differently under training and ecological evaluation. We present Neuroevolution Arena,

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

NVIDIA AI Factory Compute Is Becoming an Investable Asset Class

DGX agent

We announced partnerships with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR to establish independent financing platforms designed to mobilize over $500 billion of third-party capit

hardwarenvidia-blog
12 Aug 2026
Hardware

TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling

DGX agent

arXiv:2608.10402v1 Announce Type: new Abstract: Reinforcement learning (RL) for large language models is moving toward multi-turn agentic workloads, where rollout tasks repeatedly pause for external e

hardwarearxiv-cs-lg
12 Aug 2026
Hardware

Tiered KV cache for large LLMs on Amazon SageMaker HyperPod with Curvine

DGX agent

Running large language model inference at scale forces a KV cache trade-off: oversized GPU instances or slow time-to-first-token. This post builds a tiered KV cache on Amazon SageMaker HyperPod that e

hardwareaws-ml-blog
12 Aug 2026
Hardware

TransitReID: Transit OD Data Collection with Occlusion-Resistant Dynamic Passenger Re-Identification

DGX agent

arXiv:2504.11500v3 Announce Type: replace-cross Abstract: Transit Origin-Destination (OD) data are fundamental for optimizing public transit services, yet current collection methods, such as manual su

hardwarearxiv-cs-ai
12 Aug 2026
Hardware

What do you guys do for GPU Kernels?

DGX agent

I'm trying to figure out GPU Kernel optimization on older hardware like SM80(ampere) . Is there tools you guys use? Or frameworks? Im waiting for this framework https://www.reddit.com/r/LocalLLaMA/com

hardwarer-localllama
12 Aug 2026
Hardware

Aero Realtime: Fully Aligned Input-Output Streams for Low-Latency Streaming Multimodal Generation

DGX agent

arXiv:2608.08469v1 Announce Type: new Abstract: Existing streaming multimodal models process observations incrementally but still follow a turn-based prefill-then-decode pattern, making them non-duple

hardwarearxiv-cs-ai
11 Aug 2026
Hardware

AraSSM: A bidirectional state-space encoder for Arabic masked language modeling

DGX agent

arXiv:2608.08256v1 Announce Type: new Abstract: Pretrained Transformer encoders such as AraBERT, MARBERT, and CAMeLBERT have become the standard backbone for Arabic natural language understanding, but

hardwarearxiv-cs-cl
11 Aug 2026
Hardware

Beyond Isotropic Assumptions: Continuity-Constrained Segmentation and GPU Morphometry for Nanoscale GBM Analysis

DGX agent

arXiv:2608.07575v1 Announce Type: new Abstract: Confocal microscopy of optically cleared and swelled tissue resolves complex biological structures in 3D, but such acquisitions are highly anisotropic:

hardwarearxiv-cs-cv
11 Aug 2026
Hardware

Big Tech's AI boom echoes the 1870s railroad buildout, and Nvidia shifting risk to institutional capital may expose investors if AI revenues fail to materialize (Ben Thompson/Stratechery)

DGX agent

Ben Thompson / Stratechery: Big Tech's AI boom echoes the 1870s railroad buildout, and Nvidia shifting risk to institutional capital may expose investors if AI revenues fail to materialize — On Januar

hardwaretechmeme
11 Aug 2026
Hardware

Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching

DGX agent

arXiv:2608.09444v1 Announce Type: cross Abstract: A main promise of looped language models (LMs) is depth-adaptive inference. By iterating a block of shared layers a variable number of times, the mode

hardwarearxiv-cs-cl
11 Aug 2026
Hardware

Design Space of Self--Consistent Electrostatic Machine Learning Interatomic Potentials

DGX agent

arXiv:2603.14700v2 Announce Type: replace-cross Abstract: Machine learning interatomic potentials (MLIPs) have become widely used tools in atomistic simulations. For much of the history of this field,

hardwarearxiv-cs-lg
11 Aug 2026
Hardware

Differentiate the Solver, Not the Equation: Reverse-Sweep Adjoints for Block Implicit Simulation

DGX agent

arXiv:2608.08559v1 Announce Type: cross Abstract: Differentiable simulation is a key component in learning, control, and inverse problems, where gradients through nonlinear implicit solvers are requir

hardwarearxiv-cs-lg
11 Aug 2026
Hardware

EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference

DGX agent

arXiv:2608.07964v1 Announce Type: cross Abstract: Load Balancing has emerged as a critical problem in expert-parallel distributed inference of Mixture-of-Experts (MoE) models. As routing distributions

hardwarearxiv-cs-ai
11 Aug 2026
Hardware

Eco-SoC: A Sustainable VLSI Architecture for Energy-Proportional Artificial Intelligence

DGX agent

arXiv:2608.08761v1 Announce Type: cross Abstract: In an era defined by escalating climate change and the pervasive deployment of edge intelligence, the environmental cost of semiconductor manufacturin

hardwarearxiv-cs-ai
11 Aug 2026
Hardware

ERF-GS: Reconstructing Fast Motion from Disjoint Event-RGB Viewpoints

DGX agent

arXiv:2608.08531v1 Announce Type: new Abstract: Deep learning-driven representations such as neural radiance fields (NeRFs) and 3D Gaussian splatting (3DGS) have revolutionized the field of dynamic 3D

hardwarearxiv-cs-cv
11 Aug 2026
Hardware

EsaacSim: A Multimodal Event Camera Add-on for NVIDIA Isaac Sim

DGX agent

arXiv:2608.08522v1 Announce Type: new Abstract: Event-based vision is becoming an increasingly important sensing paradigm for robotics, yet its adoption remains limited by sensor availability and the

hardwarearxiv-cs-ro
11 Aug 2026
Hardware

FlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal Applications

DGX agent

arXiv:2607.18171v2 Announce Type: replace Abstract: Real-time multimodal applications, including voice agents and interactive video generation, compose heterogeneous models into pipelines whose effici

hardwarearxiv-cs-lg
11 Aug 2026
Hardware

From Approachability Residuals to Anytime-Valid Evidence: The Online Convex Geometry of Testing by Betting

DGX agent

arXiv:2608.09450v1 Announce Type: new Abstract: Betting-based sequential tests and Blackwell approachability are linked by a rate-explicit reduction through support-function residuals. For a compact c

hardwarearxiv-cs-lg
11 Aug 2026
Hardware

Holder Signed Distance: A Differentiable, Signed, Parallelizable Metric for Robotics

DGX agent

arXiv:2608.07707v1 Announce Type: new Abstract: Computing distances between sets is essential in robotic motion planning and control, where differentiable gradients enable real-time optimization. The

hardwarearxiv-cs-ro
11 Aug 2026
Hardware

IBM and Together AI sign a $240M, multi-year deal to build an AI inference cluster on IBM Cloud, using Nvidia's HGX B300 systems, to support open-source models (Anhata Rooprai/Reuters)

DGX agent

Anhata Rooprai / Reuters: IBM and Together AI sign a 240M, multi-year deal to build an AI inference cluster on IBM Cloud, using Nvidia's HGX B300 systems, to support open-source models — IBM (IBM.N) a

hardwaretechmeme
11 Aug 2026
Hardware

Mesh-Attention: A New Communication-Efficient Distributed Attention with Improved Data Locality

DGX agent

arXiv:2512.20968v2 Announce Type: replace-cross Abstract: Distributed attention is essential for scaling large language models (LLMs) to long contexts, yet existing methods either have limited paralle

hardwarearxiv-cs-ai
11 Aug 2026
Hardware

More on our partnership: https://newsroom.ibm.com/2026-08-11-IBM-and-Together-AI-Sign-Multi-Year-Agreement-to-Scale-Open-Source-AI-Inference…

DGX agent

In August 2026, Together AI entered into a multi‑year partnership with IBM and NVIDIA to deliver enterprise‑grade open‑source AI inference on IBM Cloud. The collaboration deploys a dedicated NVIDIA B3

hardwaretogether-ai--x
11 Aug 2026
Hardware

Multi-tier storage rewrites the economics of AI inference

DGX agent

As inference becomes the dominant workload in AI infrastructure, multi-tier storage architectures are emerging as a key method for cost control and enhanced performance. These architectures combine fl

hardwaresiliconangle
11 Aug 2026
Hardware

🦔Nvidia announced agreements yesterday with the six biggest names in private capital, Apollo, Blackstone, BlackRock, Brookfield, Goldman Sa…

DGX agent

🦔Nvidia announced agreements yesterday with the six biggest names in private capital, Apollo, Blackstone, BlackRock, Brookfield, Goldman Sachs, and KKR, to raise over $500 billion so its own customers

hardwaregary-marcus--x
11 Aug 2026
Hardware

NVIDIA JetPack 7.2.1 Adds Agentic Video Skills and T3000 Emulation

DGX agent

NVIDIA JetPack 7.2.1 adds agentic video skills with the unified jetson‑videosdk, allowing programmable, device-aware video workflows that link developer intent to live device discovery and performance

hardwarenvidia-developer
11 Aug 2026
Hardware

Nvidia Nemo Switchyard

DGX agent

https://github.com/NVIDIA-NeMo/Switchyard Finally an open source LLM router. An alternative to openrouter fusion and Sakana Fugu. Doesn't look like it does exactly what Sakana Fugu does according to i

hardwarer-localllama
11 Aug 2026
Hardware

Nvidia taps Wall Street for a half-trillion dollars to fuel global AI infrastructure buildout

DGX agent

Some of Wall Street’s biggest financial firms are partnering with Nvidia Corp. to pour a half-trillion dollars of funding into the artificial intelligence industry’s massive infrastructure buildout. N

hardwaresiliconangle
11 Aug 2026
Hardware

Personalized AI startup River AI raises $1.1B from consortium backed by Nvidia, AMD

DGX agent

River AI Inc., a startup that helps enterprises customize open-source artificial intelligence models, has raised 1.1 billion in early-stage funding. The company stated in today’s announcement that it

hardwaresiliconangle
11 Aug 2026
Hardware

Physics-Informed Condition Monitoring of SiC Power Modules

DGX agent

arXiv:2608.08363v1 Announce Type: cross Abstract: Silicon carbide (SiC) power modules are increasingly deployed in automotive traction inverters, where condition monitoring is essential to prevent in-

hardwarearxiv-cs-lg
11 Aug 2026
Hardware

Real-time physics inversion for retrieval of sub-pixel wildfire temperatures from VSWIR imaging spectroscopy

DGX agent

arXiv:2608.07580v1 Announce Type: new Abstract: In this work, we present a wildfire temperature retrieval framework for VSWIR imaging spectroscopy data, employed on data from NASA's Airborne Visible I

hardwarearxiv-cs-cv
11 Aug 2026
Hardware

Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard

DGX agent

NVIDIA NeMo Switchyard is a routing platform that directs AI agent workloads to the most suitable specialized or frontier model for each step of a task, balancing performance, cost, and latency. It of

hardwarenvidia-developer
11 Aug 2026
Hardware

StitchCUDA: An Automated Multi-Agents End-to-End GPU Programing Framework with Rubric-based Agentic Reinforcement Learning

DGX agent

arXiv:2603.02637v2 Announce Type: replace-cross Abstract: Modern machine learning (ML) workloads increasingly rely on GPUs, yet achieving high end-to-end performance remains challenging due to depende

hardwarearxiv-cs-cl
11 Aug 2026
Hardware

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization

DGX agent

arXiv:2608.09160v1 Announce Type: new Abstract: Query-Key Normalization (QK-Norm) improves the training stability and quality of modern Large Language Models (LLMs). However, under Tensor Parallelism

hardwarearxiv-cs-lg
11 Aug 2026
Hardware

Together AI is teaming up with @IBM and @nvidia to bring enterprise-grade AI inference to IBM Cloud. A dedicated NVIDIA B300 cluster. Spectr…

DGX agent

Together AI is teaming up with @IBM and @nvidia to bring enterprise-grade AI inference to IBM Cloud. A dedicated NVIDIA B300 cluster. Spectrum-X networking. First of its kind on IBM Cloud, powered by

hardwaretogether-ai--x
11 Aug 2026
Hardware

verdi: retrieval is not transfer for continual world model optimization

DGX agent

arXiv:2608.09537v1 Announce Type: new Abstract: Foundation world models have made remarkable progress in planning, simulation, and embodied intelligence. However, optimizing a pretrained world model t

hardwarearxiv-cs-ai
11 Aug 2026
Hardware

What Irregularity Costs: CUDA C++, Rust, and Triton on a Hash-Blocked GPU Workload

DGX agent

arXiv:2608.08287v1 Announce Type: new Abstract: GPU language comparisons are almost always run on tiled dense linear algebra, where every toolchain is good and the differences are small. We implement

hardwarearxiv-cs-cv
11 Aug 2026
Hardware

Why Scaling AI Compute Performance Requires a New Power Architecture

DGX agent

Every new generation of accelerated computing demands more from the infrastructure underneath it — more compute performance, higher rack density and more efficient, scalable power distribution. The bo

hardwarenvidia-blog
11 Aug 2026
Hardware

2/ New deep dive: Autoscaling endpoints for LLM inference. Dedicated Inference can scale on eight metrics. inflight_requests is the default …

DGX agent

2/ New deep dive: Autoscaling endpoints for LLM inference. Dedicated Inference can scale on eight metrics. inflight_requests is the default because it sees queue pressure before latency degrades. We t

hardwaretogether-ai--x
10 Aug 2026
Hardware

Beyond Visibility: Real-Time Surface Accessibility Fields from Sparse LiDAR

DGX agent

arXiv:2608.06412v1 Announce Type: cross Abstract: Understanding which surfaces in a scene are physically accessible to a given tool is fundamental for robotic interaction, yet 3D perception systems ty

hardwarearxiv-cs-ro
10 Aug 2026
← Previous
123…37
Next →