AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
All
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
Hardware

The S-ICDF Dataset: Sionna-Simulated Dynamic Interference Characterization and Direction Finding

DGX agent

arXiv:2607.03411v1 Announce Type: cross Abstract: Jamming and spoofing threaten wireless and satellite navigation by disrupting or manipulating radio frequency (RF) signals, undermining availability,

hardwarearxiv-cs-ai
7 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kerne…

DGX agent

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kernels all working together to make inference faster for users. P

hardwaretogether-ai--x
7 Jul 2026
Hardware

UK-based AI infrastructure startup Nscale secures a $900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific (Mauro Orru/Wall Street Journal)

DGX agent

Mauro Orru / Wall Street Journal: UK-based AI infrastructure startup Nscale secures a 900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific — Nscale raised

hardwaretechmeme
7 Jul 2026
Hardware

Wan-Streamer v0.2: Higher Resolution, Same Latency

DGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microso…

DGX agent

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microsoft, Amazon, Google AI Debt Financing will be over $7T of deb

hardwaredylan-patel--x
6 Jul 2026
Hardware

Enhancing Goodput in Large-Scale LLM Training with Nonuniform Tensor Parallelism

DGX agent

Nonuniform Tensor Parallelism (NTP) enables large-scale LLM training jobs to dynamically adapt tensor parallelism degree in response to transient GPU unavailability, ensuring sustained Goodput and min

hardwarenvidia-developer
6 Jul 2026
Hardware

How Open Models Are Driving AI Research

DGX agent

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open

hardwarenvidia-blog
6 Jul 2026
Hardware

Import AI 464: Fable writes GPU kernels; AI automation; and analog computation

DGX agent

This Import AI newsletter issue covers Fable's development of GPU kernel writing capabilities, likely discussing how AI systems can automatically generate optimized code for graphics processors, along

hardwareimport-ai
6 Jul 2026
Hardware

Join us for a fireside chat on where AI research and infrastructure are headed, led by @tri_dao. Hosted by Together AI, @nvidia and Lyra Lab…

DGX agent

Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro

hardwaretogether-ai--x
6 Jul 2026
Hardware

Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters

DGX agent

Over 7T AI debt by 2029, There can be no Neoclouds without the Trinity. Nvidia's Backstop Economics Explained. AI Debt Needs Quantified. Nvidia's Objective is to Broaden Compute Access, Develop AI Fin

hardwaresemianalysis
6 Jul 2026
Hardware

SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture (Anniek Bao/CNBC)

DGX agent

Anniek Bao / CNBC: SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture — NVIDIA's next marq

hardwaretechmeme
6 Jul 2026
Hardware

Shanghai-based AI chipmaker Biren raises ~$892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO (Ann Cao/South China Morning Post)

DGX agent

Ann Cao / South China Morning Post: Shanghai-based AI chipmaker Biren raises ~892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO — Chinese ar

hardwaretechmeme
6 Jul 2026
Hardware

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs …

DGX agent

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs that also would have killed the company. He realized that th

hardwareemad-mostaque--x
6 Jul 2026
Hardware

183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD student) is one of the be…

DGX agent

183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD student) is one of the best educational videos on kernels out there (think @karpathy

hardwareswyx--x
5 Jul 2026
Hardware

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemming…

DGX agent

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemmings. Ttypical East Coast VCs all think that they're smarter th

hardwareyann-lecun--x
5 Jul 2026
Hardware

First animation from integrated PyTTI C++/CUDA port. I used DD style (@dango233max) CLIP cutouts since that's what I already had in the code…

DGX agent

This post showcases an early animation demonstration from a C++/CUDA implementation of PyTTI (Python Text-to-Image), utilizing CLIP cutout techniques in the style of DD (likely referring to a specific

hardwareemad-mostaque--x
4 Jul 2026
Hardware

God bless America the land of the free and home of the brave 🫡 250

DGX agent

This post likely references the patriotic phrase from the U.S. national anthem, possibly in the context of Dylan Patel's technology industry commentary or analysis. Without access to the specific cont

hardwaredylan-patel--x
4 Jul 2026
Hardware

Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M (Madison McLauchlan/BetaKit)

DGX agent

Madison McLauchlan / BetaKit: Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M — Stathera

hardwaretechmeme
4 Jul 2026
Hardware

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I hav…

DGX agent

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I haven’t yet heard from industry leaders Open source AI is about

hardwareclem-delangue--x
4 Jul 2026
Hardware

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

DGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

DGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

DeadPool: Resilient LLM Training with Hot-Swapping via Zero-Overhead Checkpoint

DGX agent

arXiv:2607.01646v1 Announce Type: new Abstract: State-of-the-art large language model (LLM) training takes tens of thousands of graphics processing units (GPUs) for months and encounters failures acro

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there?

DGX agent

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there? Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference

hardwarefireworks-ai--x
3 Jul 2026
Hardware

EHHN: An Event-driven Heterogeneous Hypergraph Network for Object-Centric Next Activity Prediction

DGX agent

arXiv:2607.01785v1 Announce Type: new Abstract: Next activity prediction helps service-oriented processes anticipate upcoming steps before delays, exceptions, or service-level risks occur. Most existi

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

GPUAlert: A Zero-Instrumentation Process-Boundary Monitor for Diagnosing GPU Training-Job Failures

DGX agent

arXiv:2607.01409v1 Announce Type: cross Abstract: GPU training jobs fail often, roughly two in five on large production clusters, yet the operator typically learns of a failure only by reconnecting ho

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation

DGX agent

arXiv:2607.01803v1 Announce Type: cross Abstract: Recent advances in 3D content generation from text or images have achieved impressive results, yet view inconsistency from 2D generators and the scarc

hardwarearxiv-cs-ro
3 Jul 2026
Hardware

Stable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory Gates

DGX agent

arXiv:2607.02363v1 Announce Type: cross Abstract: Quantum Fast-Weight Programmers (QFWPs) store temporal information in dynamically programmed variational-circuit parameters rather than in nonlinear r

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

WattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs

DGX agent

arXiv:2607.02391v1 Announce Type: cross Abstract: Large Language Model (LLM) inference workloads are a rapidly growing contributor to data center energy consumption. Optimizing these deployments requi

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines t…

DGX agent

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines to serve agentic workloads at trillion token production scale

hardwaretogether-ai--x
3 Jul 2026
Hardware

Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

DGX agent

arXiv:2607.00773v1 Announce Type: new Abstract: Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the ac

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

Condensing Large-Scale Datasets Directly with Minimal Information Loss

DGX agent

arXiv:2607.00916v1 Announce Type: new Abstract: Recent advancements in scaling dataset distillation rely heavily on decoupled information extraction pipelines, comprising SQUEEZE, RECOVER, and RELABEL

hardwarearxiv-cs-cv
2 Jul 2026
Hardware

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

DGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

EMIB-T Roadmap, Custom HBM, HBM4 Packaging Challenges, Microfluidic Cooling, Photonic Interconnects, and More

DGX agent

This SemiAnalysis article discusses advanced semiconductor packaging technologies and innovations presented at or related to ECTC 2026, including Intel's EMIB-T chiplet interconnect roadmap, custom HB

hardwaresemianalysis
2 Jul 2026
Hardware

Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan (New York Times)

DGX agent

New York Times: Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan — President Trump and hi

hardwaretechmeme
2 Jul 2026
Hardware

GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics

DGX agent

arXiv:2607.01203v1 Announce Type: cross Abstract: This paper studies real-time robust optimal control for uncertain nonlinear systems, where linear time-varying (LTV) approximations make planning trac

hardwarearxiv-cs-ai
2 Jul 2026
Hardware

Hardware-Rooted AI Security That Won’t Slow You Down

DGX agent

NVIDIA Confidential Computing addresses data privacy and security concerns for AI workloads by protecting data during inference and engagement with models, offering high-performance protection in loca

hardwarenvidia-developer
2 Jul 2026
Hardware

Joyride Through July With 12 Games Coming to GeForce NOW

DGX agent

Summer is heating up — and GeForce NOW is taking players along for the ride. Start the month with Monopoly: Star Wars Heroes vs. Villains, bringing a galaxy far, far away to the iconic board-game fran

hardwarenvidia-blog
2 Jul 2026
Hardware

Meta Compute: Everyone Wants To Be A Cloud

DGX agent

Meta Compute explores Meta's strategy to position itself as a cloud infrastructure provider, leveraging its massive internal compute investments and AI capabilities to offer services to external custo

hardwaresemianalysis
2 Jul 2026
Hardware

MosaicKV: Serving Long-Context LLM with Dynamic Two-D KV Cache Compression

DGX agent

arXiv:2607.00760v1 Announce Type: new Abstract: Long-context LLM services now sustain prompts with hundreds of thousands to millions of tokens, making the key-value (KV) cache a first-order serving co

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout

DGX agent

As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that generate tokens at scale. This shift requires ac

hardwarenvidia-blog
2 Jul 2026
Hardware

Palantir CEO Alex Karp doesn’t hold back in interview as he rails against AI industry

DGX agent

Palantir Technologies Inc. Chief Executive Alex Karp appeared to go into meltdown mode during an interview with CNBC today where for 20-odd minutes he went off script after being asked to discuss his

hardwaresiliconangle
2 Jul 2026
Hardware

ROSA: A Robotics Foundation Model Serving System for Robot Factories

DGX agent

arXiv:2607.01088v1 Announce Type: new Abstract: Robotics foundation models (RFMs) are making general-purpose robots increasingly practical for factory deployments. While RFM serving systems are centra

hardwarearxiv-cs-ro
2 Jul 2026
Hardware

SNAP-FM: Sparse Nonlinear Accelerated Projection for Physics-Constrained Generative Modeling

DGX agent

arXiv:2607.00095v1 Announce Type: cross Abstract: Generative models have emerged as scalable surrogates for physical simulation, yet they offer no guarantee that their outputs respect the conservation

hardwarearxiv-cs-ai
2 Jul 2026
Hardware

“Within the next 18 months, you will be able to host GLM 5.2 equivalent intelligence on an RTX 5090 GPU.” -Ahmad Osman, AI World’s Fair

DGX agent

Ahmad Osman stated at AI World's Fair that within 18 months, GLM 5.2-equivalent AI intelligence will be deployable locally on a single RTX 5090 GPU, indicating rapid progress toward running advanced l

hardwareswyx--x
2 Jul 2026
Hardware

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Pa…

DGX agent

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Payments, Hissa Fund & Wispr Flow. Proudly presented by GrowthX

hardwarenous-research--x
2 Jul 2026
Hardware

9/ ParallelKernelBench: Benchmarking LLMs on Multi-GPU Kernel Generation Paper: https://www.alphaxiv.org/abs/2606.parallel-kernel-bench

DGX agent

ParallelKernelBench is a benchmarking framework designed to evaluate large language models' ability to generate optimized GPU kernels for multi-GPU computing environments. The benchmark assesses LLMs

hardwaretogether-ai--x
1 Jul 2026
Hardware

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

DGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

hardwarearxiv-cs-ai
1 Jul 2026
Hardware

An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU

DGX agent

arXiv:2603.16428v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) has become essential for domain adaptation, but its memory-intensive property exceeds the capabilitie

hardwarearxiv-cs-ai
1 Jul 2026
← Previous
1…89101112…37
Next →