AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
All
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
Hardware

Understanding GPU Inference Workloads [D]

DGX agent

Hey everyone, I have been looking into how people source compute for their Inference workloads (and in general). I wanted to understand some specific pain points here. If you've used online services l

hardwarer-machinelearning
26 Jul 2026
Hardware

AMD moves beyond challenger status in the race for AI platform leadership

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

AI hardware competition entered a sharper phase this week as Advanced Micro Devices Inc. used its flagship AI event to argue it isn’t merely chasing Nvidia Corp. — it intends to lead the market outrig

hardwaresiliconangle
25 Jul 2026
Hardware

AMD’s Helios strategy turns the GPU battle into a systems contest

DGX agent

The market for data center GPUs is evolving beyond individual chip specifications into a contest over fully integrated rack-scale systems. As AI workloads scale into the gigawatt range, buyers increas

hardwaresiliconangle
25 Jul 2026
Hardware

Can AMD break the CUDA Moat? AMD Advancing AI 2026

DGX agent

AMD’s AI accelerator software stack has progressed sharply, moving from a 0 % chance of catching Nvidia’s CUDA moat in early 2025 to a “great chance” of success by July 2026 after leadership changes a

hardwaresemianalysis
25 Jul 2026
Hardware

don't do this

DGX agent

Title: “don’t do this” refers to a conversation on Twitter where Jerry Liu warns against a particular action, while Julian Schrittwieser comments on a separate thread expressing excitement that Jensen

hardwarejerry-liu--x
25 Jul 2026
Hardware

Hey sir. We are not asking you to open source Anthropic. Just don’t lobby the government to shut down others who do. Jensen never framed oth…

DGX agent

Hey sir. We are not asking you to open source Anthropic. Just don’t lobby the government to shut down others who do. Jensen never framed other chips as “dangerous” or decides who can use CUDA based on

hardwarejeremy-howard--x
25 Jul 2026
Hardware

I tried making a cinematic action trailer using Krea 2 + LTX 2.3

DGX agent

I wanted to challenge myself and see how far I could push Krea 2 and LTX 2.3, so I decided to create a short cinematic action trailer. It ended up being one of the most enjoyable AI projects I've work

hardwarer-stablediffusion
25 Jul 2026
Hardware

“open source shouldn’t be banned” != “everything must be open source” but also to NVidia credit big parts of the toolchain are increasingly …

DGX agent

“open source shouldn’t be banned” != “everything must be open source” but also to NVidia credit big parts of the toolchain are increasingly open source: driver, cutedsl kernels, etc I’m so excited tha

hardwaresonya-huang--x
25 Jul 2026
Hardware

Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI (New York Times)

DGX agent

New York Times: Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI — Anthropic and OpenAI

hardwaretechmeme
25 Jul 2026
Hardware

SVDQuant + native INT8/W4A4 for Krea 2 on ComfyUI — up to 2x faster, works on any modern NVIDIA GPU

DGX agent

Quantized Krea 2 Turbo checkpoints for ComfyUI, up to 2x faster and about a third smaller than the usual FP8 version — no calibration dataset, no quality cliff. How to use it (short version): clone th

hardwarer-stablediffusion
25 Jul 2026
Hardware

AMD takes on Nvidia, US takes on Chinese AI models and AI spending still spooks investors

DGX agent

Maybe Nvidia ultimately won’t win the entire AI enchilada, even as it reminded everyone this week that it offers everything you need to build AI factories. This week, Advanced Micro Devices also annou

hardwaresiliconangle
24 Jul 2026
Hardware

At AI Summit, South Korea Outlines Its AI Future With NVIDIA and Partners

DGX agent

At this week’s AI Summit in San Francisco, South Korean President Jae Myung Lee and some of the country’s top business leaders and researchers are meeting with NVIDIA and ecosystem partners to chart K

hardwarenvidia-blog
24 Jul 2026
Hardware

Congrats to @lmstudio on Bionic! Put open models to work with a local-first agent for docs, coding, voice, and more. Run it today on NVIDIA …

DGX agent

Congrats to @lmstudio on Bionic! Put open models to work with a local-first agent for docs, coding, voice, and more. Run it today on NVIDIA RTX GPUs. 🚀 Meet LM Studio Bionic. The Agent made for Open M

hardwarelm-studio--x
24 Jul 2026
Hardware

Evaluating and Guarding Citation Faithfulness in Agentic Scientific Synthesis

DGX agent

arXiv:2607.20527v1 Announce Type: new Abstract: Agentic LLM systems such as OpenScholar and PaperQA2 read the scientific literature and return cited answers, and both they and their benchmarks already

hardwarearxiv-cs-ai
24 Jul 2026
Hardware

KroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers

DGX agent

arXiv:2607.21446v1 Announce Type: cross Abstract: Post-training quantization (PTQ) of diffusion transformers (DiTs) to W4A4 severely degrades output quality, because activations entering each linear l

hardwarearxiv-cs-cv
24 Jul 2026
Hardware

Meta, Nvidia, Microsoft, a16z, and others sign a letter defending open-source AI; Jensen Huang, in his first X post, says open models strengthen cybersecurity (Leo Schwartz/The Information)

DGX agent

Leo Schwartz / The Information: Meta, Nvidia, Microsoft, a16z, and others sign a letter defending open-source AI; Jensen Huang, in his first X post, says open models strengthen cybersecurity — Many of

hardwaretechmeme
24 Jul 2026
Hardware

ModelExpress: Distributing Model Artifacts at the Speed of Light

DGX agent

NVIDIA ModelExpress (MX) is an agentic AI platform that streamlines the lifecycle of large‑model weights by automatically locating and using the fastest path to load them—prioritizing GPU‑to‑GPU P2P R

hardwarenvidia-developer
24 Jul 2026
Hardware

Ms. Forcing: Efficient Streaming Video Generation with Multi-Scale Patchification and Attention

DGX agent

arXiv:2607.20940v1 Announce Type: new Abstract: Streaming video diffusion models have made substantial progress toward interactive and dynamic world simulation, but the nested autoregressive and denoi

hardwarearxiv-cs-cv
24 Jul 2026
Hardware

Nvidia and SK Group unveil a $500B+ AI initiative that includes an SK Hynix partnership to secure next-gen memory supply for Nvidia and joint development of HBM (Reuters)

DGX agent

Reuters: Nvidia and SK Group unveil a 500B+ AI initiative that includes an SK Hynix partnership to secure next-gen memory supply for Nvidia and joint development of HBM — Nvidia (NVDA.O) and South Kor

hardwaretechmeme
24 Jul 2026
Hardware

NVIDIA-labs OO Agents: Native Python Object-Oriented Agents

DGX agent

arXiv:2607.20709v1 Announce Type: new Abstract: Traditional agent development is split across prompt templates, tool schemas, callback code, and workflow graphs. We present NVIDIA Object-Oriented Agen

hardwarearxiv-cs-ai
24 Jul 2026
Hardware

Nvidia plans to invest $1B in Naver to help finance an AI data center in South Korea, and partners with SK Group to build more than 2 GW of AI data centers (Ian King/Bloomberg)

DGX agent

Ian King / Bloomberg: Nvidia plans to invest 1B in Naver to help finance an AI data center in South Korea, and partners with SK Group to build more than 2 GW of AI data centers — Nvidia Corp. will inv

hardwaretechmeme
24 Jul 2026
Hardware

🚨OpenAI, the company with 'Open' in the name, didn't sign the NVIDIA open weights letter Instead, they're lobbying government to BAN open A…

DGX agent

On July 24 2026, a tweet from NIK @ns123abc reported that OpenAI declined to sign NVIDIA’s “open weights” letter. Instead, the company is reportedly lobbying government authorities to ban open‑AI tech

hardwaregary-marcus--x
24 Jul 2026
Hardware

SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation

DGX agent

arXiv:2607.21553v1 Announce Type: new Abstract: We introduce SANA-Video 2.0, a hybrid video diffusion transformer instantiated at 5B and 14B scales under a unified architecture. Designed to generate h

hardwarearxiv-cs-cv
24 Jul 2026
Hardware

SPORD: A Simulation-Propose-then-OR-Dispose Approach for Supply Chain Planning

DGX agent

arXiv:2607.21354v1 Announce Type: new Abstract: For years, supply chain planning at e-commerce firms has operated as a collection of isolated projects. Each planning task from static network planning

hardwarearxiv-cs-ai
24 Jul 2026
Hardware

Synopsys targets physical AI complexity with co-design and agentic chip workflows

DGX agent

The rapid evolution of intelligent software-defined systems is pushing chip design into a new era of physical AI, one where chip design complexity is outpacing traditional engineering methods. Manufac

hardwaresiliconangle
24 Jul 2026
Hardware

Trainable Log-linear Sparse Attention for Efficient Diffusion Transformers

DGX agent

arXiv:2512.16615v2 Announce Type: replace Abstract: Diffusion Transformers (DiTs) set the state of the art in visual generation, yet their quadratic self-attention cost fundamentally limits scaling to

hardwarearxiv-cs-cv
24 Jul 2026
Hardware

Agentic AI compute reshapes cloud economics as AMD targets GPU and CPU demand surge

DGX agent

Agentic AI compute has become the defining workload of the cloud’s next era, reshaping how hyperscalers, neo-clouds and enterprises architect their infrastructure. As autonomous agents move into produ

hardwaresiliconangle
23 Jul 2026
Hardware

AI infrastructure systems redefine the AMD-Nvidia rivalry as inference reshapes the market

DGX agent

The race to build AI infrastructure systems has moved beyond chip specifications into a battle over entire rack-scale platforms, as inference and agentic workloads redefine what counts as a computer.

hardwaresiliconangle
23 Jul 2026
Hardware

AMD debuts next-generation AI infrastructure for frontier models, agentic workloads and autonomous robots

DGX agent

Advanced Micro Devices Inc. is pushing harder than ever to grab even more market share from Nvidia Corp. in the artificial intelligence chip industry. At its Advancing AI 2026 event today in San Franc

hardwaresiliconangle
23 Jul 2026
Hardware

Anatomy-Aware 3D Mesh Refinement of Pericardium Segmentations on Computed Tomography

DGX agent

arXiv:2607.19210v1 Announce Type: new Abstract: Accurate delineation of the pericardium in a cardiac CT scan is essential for quantifying epicardial adipose tissue, yet it remains one of the most chal

hardwarearxiv-cs-cv
23 Jul 2026
Hardware

At @NVIDIAAI we continue to push open data, techniques and models forward because we know that every organization needs the freedom to build…

DGX agent

At @NVIDIAAI we continue to push open data, techniques and models forward because we know that every organization needs the freedom to build and deploy AI in their own way. We're now the biggest insti

hardwareclem-delangue--x
23 Jul 2026
Hardware

ATSplat: Compact Feed-forward 3D Gaussian Splatting with Adaptive Token Expansion

DGX agent

arXiv:2607.20417v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) achieves high-quality novel-view synthesis by optimizing freely placed primitives in 3D and adaptively densifying them in u

hardwarearxiv-cs-cv
23 Jul 2026
Hardware

Debugging Ray Tracing Applications Using NVIDIA OptiX Toolkit

DGX agent

The NVIDIA OptiX Toolkit (OTK) is a BSD 3‑clause licensed collection of utilities that provides unified, minimal‑macro error checking across OptiX, CUDA runtime, and CUDA driver APIs, generating diagn

hardwarenvidia-developer
23 Jul 2026
Hardware

From Pixel to Prognosis: Convolutional and GLCM Feature Fusion for Automated Four-Class Cataract Severity Classification

DGX agent

arXiv:2607.18349v1 Announce Type: new Abstract: Objective: To develop a low-cost automated cataract severity classification system operating on standard consumer-grade colour photographs of the eye, w

hardwarearxiv-cs-cv
23 Jul 2026
Hardware

GeForce NOW Sets Sail With ‘Path of Exile: Curse of the Allflame’ Joining the Cloud

DGX agent

Lock in and load up the cloud. GFN Thursday brings fresh updates and new adventures, all ready to play without waiting for downloads. Set sail in Path of Exile: Curse of the Allflame and charge in Bat

hardwarenvidia-blog
23 Jul 2026
Hardware

InstantSfM: Towards GPU-Native SfM for the Deep Learning Era

DGX agent

arXiv:2510.13310v3 Announce Type: replace Abstract: Structure-from-Motion (SfM) is a fundamental technique for recovering camera poses and scene structure from multi-view imagery, serving as a critica

hardwarearxiv-cs-cv
23 Jul 2026
Hardware

Leveraging ECRAM for Edge Continual Learning

DGX agent

arXiv:2607.19661v1 Announce Type: cross Abstract: Several edge computing platforms, such as autonomous vehicles and smart sensing devices, need to adapt to dynamic environments in real time by learnin

hardwarearxiv-cs-lg
23 Jul 2026
Hardware

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing

DGX agent

arXiv:2607.19064v2 Announce Type: replace-cross Abstract: Large-scale visual generators are increasingly capable but costly to train, fine-tune, and deploy. We introduce Mage-Flow, a compact 4B-scale

hardwarearxiv-cs-ai
23 Jul 2026
Hardware

Making Single-Cell Data Distillation Auditable: Traceable Real-Cell Coresets via Discrete Min-Max Selection

DGX agent

arXiv:2607.19426v1 Announce Type: cross Abstract: Single-cell datasets are increasingly costly to store, audit, and reuse for model training. Dimensionality reduction and dataset distillation can redu

hardwarearxiv-cs-ai
23 Jul 2026
Hardware

MoA-Structured Decode Attention DNF Derivation, KV-Cache Accumulation, GQA/MQA, and OpenACC Kernel

DGX agent

arXiv:2607.19456v1 Announce Type: cross Abstract: We derive four memory-optimal inference artifacts for transformer attention using the Mathematics of Arrays (MoA), each following directly from the fo

hardwarearxiv-cs-ai
23 Jul 2026
Hardware

Moonshot mudslinging CHART In battle over Chinese AI restrictions, it’s Anthropic & OpenAI versus pretty much everyone else in Silicon Valle…

DGX agent

Amir Efrati (@amir) posted a tweet titled “Moonshot mudslinging CHART” on July 23 2026, arguing that the battle over Chinese AI restrictions pits Anthropic and OpenAI against almost every other Silico

hardwareclem-delangue--x
23 Jul 2026
Hardware

Native Multi-Dimensional Subquadratic Operators via Input Dependent Long Convolutions

DGX agent

arXiv:2607.19378v1 Announce Type: new Abstract: Subquadratic alternatives to attention require compromises when applied to multi-dimensional data: standard convolutions lack global receptive fields an

hardwarearxiv-cs-lg
23 Jul 2026
Hardware

NGPS: GPS-Denied Aerial Geo-Localization and 2.5D Reconstruction via Deep Satellite Image Matching and Multi-Rate Sensor Fusion

DGX agent

arXiv:2607.18936v1 Announce Type: cross Abstract: We present NGPS (Next-Generation Positioning System), a visual geo-localization framework for high-altitude UAVs that provides GPS-free absolute posit

hardwarearxiv-cs-cv
23 Jul 2026
Hardware

Optimal Recalibration of an Online Predictor

DGX agent

arXiv:2607.19689v1 Announce Type: cross Abstract: We study the problem of recalibrating an online predictor [KE17, OKS24]: given an arbitrary 'hint' sequence of forecasts, the learner must output new

hardwarearxiv-cs-lg
23 Jul 2026
Hardware

ReRAM-aware Model Finetuning addressing I-V Non-linearity and Retention Errors

DGX agent

arXiv:2606.17471v3 Announce Type: replace Abstract: Traditional CPU, GPU, and NPU architectures are increasingly limited by the von Neumann bottleneck. While In-Memory Computing (IMC) using ReRAM cros

hardwarearxiv-cs-lg
23 Jul 2026
Hardware

Scaling Time Series Classification via XAI-Driven Data Reduction

DGX agent

arXiv:2607.15774v2 Announce Type: replace-cross Abstract: Explainable AI (XAI) for time series has seen significant algorithmic growth, but its utility in providing measurable performance gains for do

hardwarearxiv-cs-ai
23 Jul 2026
Hardware

Vera Rubin NVL72 vs GB200 NVL72? Inference TCO & Architecture Analysis

DGX agent

Vera Rubin NVL72 is Nvidia’s second‑generation, rack‑scale Oberon architecture that achieves inference gains through extreme co‑design. Early engineering‑sample data from CoreWeave show DeepSeek R1 de

hardwaresemianalysis
23 Jul 2026
Hardware

Decoupling embedding from ingestion means whatever hardware is on hand can do the work. One script, runtime device check: MPS on Apple Silic…

DGX agent

Decoupling embedding from ingestion means whatever hardware is on hand can do the work. One script, runtime device check: MPS on Apple Silicon, CUDA on NVIDIA, CPU fallback otherwise. 10K-20K records

hardwarepinecone--x
22 Jul 2026
← Previous
1…45678…37
Next →