AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
7 Jul 2026

NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community

HardwareDGX agent

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development can still be gated b

NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads

HardwareDGX agent

NVIDIA Vera is a CPU designed to help AI factories scale agentic AI and reinforcement learning by shortening CPU execution time, increasing task throughput, and enabling smarter, longer-thinking agent

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization

HardwareDGX agent

arXiv:2607.02834v1 Announce Type: new Abstract: Molecular optimization often starts from a pretrained generative model that captures a broad prior over valid molecular structures. At test time, howeve

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the lates…

HardwareDGX agent

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the latest Isaac models and frameworks to the open robotics community.

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

HardwareDGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

RoMa v2: Harder Better Faster Denser Feature Matching

HardwareDGX agent

arXiv:2511.15706v3 Announce Type: replace Abstract: Dense feature matching aims to estimate all correspondences between two images of a 3D scene and has recently been established as the gold standard

Scaling Weisfeiler-Leman Expressiveness Analysis to Massive Graphs with GPUs

HardwareDGX agent

arXiv:2607.02603v1 Announce Type: cross Abstract: The stable coloring of the Weisfeiler-Leman (1-WL) test is a cornerstone of Graph Neural Networks because it provides an upper bound to the expressive

SILO: Simulation-in-the-Loop Sim-to-Real Transfer for Multi-Stage Cable Routing

HardwareDGX agent

arXiv:2607.04616v1 Announce Type: cross Abstract: Linear-deformable manipulation remains challenging due to the complex deformations of objects such as cables and ropes. Prior data-driven approaches,

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

HardwareDGX agent

arXiv:2607.03333v1 Announce Type: cross Abstract: LLM agents are becoming a common interface for research, coding, and question answering, yet their Thought-Action-Observation loop is often serial: th

TabPack: Efficient Hyperparameter Ensembles for Tabular Deep Learning

HardwareDGX agent

arXiv:2607.05380v1 Announce Type: new Abstract: In deep learning for tabular data, efficient ensembles of multilayer perceptrons (MLPs) have recently emerged as effective and practical architectures.

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a w…

HardwareDGX agent

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a welcoming audience for my workshops and presentation. Thanks

The S-ICDF Dataset: Sionna-Simulated Dynamic Interference Characterization and Direction Finding

HardwareDGX agent

arXiv:2607.03411v1 Announce Type: cross Abstract: Jamming and spoofing threaten wireless and satellite navigation by disrupting or manipulating radio frequency (RF) signals, undermining availability,

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kerne…

HardwareDGX agent

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kernels all working together to make inference faster for users. P

UK-based AI infrastructure startup Nscale secures a $900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific (Mauro Orru/Wall Street Journal)

HardwareDGX agent

Mauro Orru / Wall Street Journal: UK-based AI infrastructure startup Nscale secures a 900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific — Nscale raised

Wan-Streamer v0.2: Higher Resolution, Same Latency

HardwareDGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

6 Jul 2026

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microso…

HardwareDGX agent

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microsoft, Amazon, Google AI Debt Financing will be over $7T of deb

Enhancing Goodput in Large-Scale LLM Training with Nonuniform Tensor Parallelism

HardwareDGX agent

Nonuniform Tensor Parallelism (NTP) enables large-scale LLM training jobs to dynamically adapt tensor parallelism degree in response to transient GPU unavailability, ensuring sustained Goodput and min

How Open Models Are Driving AI Research

HardwareDGX agent

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open

Import AI 464: Fable writes GPU kernels; AI automation; and analog computation

HardwareDGX agent

This Import AI newsletter issue covers Fable's development of GPU kernel writing capabilities, likely discussing how AI systems can automatically generate optimized code for graphics processors, along

Join us for a fireside chat on where AI research and infrastructure are headed, led by @tri_dao. Hosted by Together AI, @nvidia and Lyra Lab…

HardwareDGX agent

Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro

Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters

HardwareDGX agent

Over 7T AI debt by 2029, There can be no Neoclouds without the Trinity. Nvidia's Backstop Economics Explained. AI Debt Needs Quantified. Nvidia's Objective is to Broaden Compute Access, Develop AI Fin

SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture (Anniek Bao/CNBC)

HardwareDGX agent

Anniek Bao / CNBC: SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture — NVIDIA's next marq

Shanghai-based AI chipmaker Biren raises ~$892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO (Ann Cao/South China Morning Post)

HardwareDGX agent

Ann Cao / South China Morning Post: Shanghai-based AI chipmaker Biren raises ~892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO — Chinese ar

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs …

HardwareDGX agent

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs that also would have killed the company. He realized that th

5 Jul 2026

183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD student) is one of the be…

HardwareDGX agent

183/365 of GPU Programming This 4.5 hour lesson on CUDA + ThunderKittens by @bfspector (TK co-author, Stanford PhD student) is one of the best educational videos on kernels out there (think @karpathy

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemming…

HardwareDGX agent

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemmings. Ttypical East Coast VCs all think that they're smarter th

4 Jul 2026

First animation from integrated PyTTI C++/CUDA port. I used DD style (@dango233max) CLIP cutouts since that's what I already had in the code…

HardwareDGX agent

This post showcases an early animation demonstration from a C++/CUDA implementation of PyTTI (Python Text-to-Image), utilizing CLIP cutout techniques in the style of DD (likely referring to a specific

God bless America the land of the free and home of the brave 🫡 250

HardwareDGX agent

This post likely references the patriotic phrase from the U.S. national anthem, possibly in the context of Dylan Patel's technology industry commentary or analysis. Without access to the specific cont

Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M (Madison McLauchlan/BetaKit)

HardwareDGX agent

Madison McLauchlan / BetaKit: Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M — Stathera

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I hav…

HardwareDGX agent

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I haven’t yet heard from industry leaders Open source AI is about

3 Jul 2026

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

HardwareDGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

HardwareDGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

DeadPool: Resilient LLM Training with Hot-Swapping via Zero-Overhead Checkpoint

HardwareDGX agent

arXiv:2607.01646v1 Announce Type: new Abstract: State-of-the-art large language model (LLM) training takes tens of thousands of graphics processing units (GPUs) for months and encounters failures acro

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there?

HardwareDGX agent

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there? Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference

EHHN: An Event-driven Heterogeneous Hypergraph Network for Object-Centric Next Activity Prediction

HardwareDGX agent

arXiv:2607.01785v1 Announce Type: new Abstract: Next activity prediction helps service-oriented processes anticipate upcoming steps before delays, exceptions, or service-level risks occur. Most existi

GPUAlert: A Zero-Instrumentation Process-Boundary Monitor for Diagnosing GPU Training-Job Failures

HardwareDGX agent

arXiv:2607.01409v1 Announce Type: cross Abstract: GPU training jobs fail often, roughly two in five on large production clusters, yet the operator typically learns of a failure only by reconnecting ho

PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation

HardwareDGX agent

arXiv:2607.01803v1 Announce Type: cross Abstract: Recent advances in 3D content generation from text or images have achieved impressive results, yet view inconsistency from 2D generators and the scarc

Stable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory Gates

HardwareDGX agent

arXiv:2607.02363v1 Announce Type: cross Abstract: Quantum Fast-Weight Programmers (QFWPs) store temporal information in dynamically programmed variational-circuit parameters rather than in nonlinear r

WattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs

HardwareDGX agent

arXiv:2607.02391v1 Announce Type: cross Abstract: Large Language Model (LLM) inference workloads are a rapidly growing contributor to data center energy consumption. Optimizing these deployments requi

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines t…

HardwareDGX agent

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines to serve agentic workloads at trillion token production scale

2 Jul 2026

Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

HardwareDGX agent

arXiv:2607.00773v1 Announce Type: new Abstract: Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the ac

Condensing Large-Scale Datasets Directly with Minimal Information Loss

HardwareDGX agent

arXiv:2607.00916v1 Announce Type: new Abstract: Recent advancements in scaling dataset distillation rely heavily on decoupled information extraction pipelines, comprising SQUEEZE, RECOVER, and RELABEL

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

HardwareDGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

EMIB-T Roadmap, Custom HBM, HBM4 Packaging Challenges, Microfluidic Cooling, Photonic Interconnects, and More

HardwareDGX agent

This SemiAnalysis article discusses advanced semiconductor packaging technologies and innovations presented at or related to ECTC 2026, including Intel's EMIB-T chiplet interconnect roadmap, custom HB

Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan (New York Times)

HardwareDGX agent

New York Times: Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan — President Trump and hi

GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics

HardwareDGX agent

arXiv:2607.01203v1 Announce Type: cross Abstract: This paper studies real-time robust optimal control for uncertain nonlinear systems, where linear time-varying (LTV) approximations make planning trac

Hardware-Rooted AI Security That Won’t Slow You Down

HardwareDGX agent

NVIDIA Confidential Computing addresses data privacy and security concerns for AI workloads by protecting data during inference and engagement with models, offering high-performance protection in loca

Joyride Through July With 12 Games Coming to GeForce NOW

HardwareDGX agent

Summer is heating up — and GeForce NOW is taking players along for the ride. Start the month with Monopoly: Star Wars Heroes vs. Villains, bringing a galaxy far, far away to the iconic board-game fran

Meta Compute: Everyone Wants To Be A Cloud

HardwareDGX agent

Meta Compute explores Meta's strategy to position itself as a cloud infrastructure provider, leveraging its massive internal compute investments and AI capabilities to offer services to external custo

MosaicKV: Serving Long-Context LLM with Dynamic Two-D KV Cache Compression

HardwareDGX agent

arXiv:2607.00760v1 Announce Type: new Abstract: Long-context LLM services now sustain prompts with hundreds of thousands to millions of tokens, making the key-value (KV) cache a first-order serving co

NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout

HardwareDGX agent

As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that generate tokens at scale. This shift requires ac

Palantir CEO Alex Karp doesn’t hold back in interview as he rails against AI industry

HardwareDGX agent

Palantir Technologies Inc. Chief Executive Alex Karp appeared to go into meltdown mode during an interview with CNBC today where for 20-odd minutes he went off script after being asked to discuss his

ROSA: A Robotics Foundation Model Serving System for Robot Factories

HardwareDGX agent

arXiv:2607.01088v1 Announce Type: new Abstract: Robotics foundation models (RFMs) are making general-purpose robots increasingly practical for factory deployments. While RFM serving systems are centra

SNAP-FM: Sparse Nonlinear Accelerated Projection for Physics-Constrained Generative Modeling

HardwareDGX agent

arXiv:2607.00095v1 Announce Type: cross Abstract: Generative models have emerged as scalable surrogates for physical simulation, yet they offer no guarantee that their outputs respect the conservation

“Within the next 18 months, you will be able to host GLM 5.2 equivalent intelligence on an RTX 5090 GPU.” -Ahmad Osman, AI World’s Fair

HardwareDGX agent

Ahmad Osman stated at AI World's Fair that within 18 months, GLM 5.2-equivalent AI intelligence will be deployable locally on a single RTX 5090 GPU, indicating rapid progress toward running advanced l

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Pa…

HardwareDGX agent

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Payments, Hissa Fund & Wispr Flow. Proudly presented by GrowthX

1 Jul 2026

9/ ParallelKernelBench: Benchmarking LLMs on Multi-GPU Kernel Generation Paper: https://www.alphaxiv.org/abs/2606.parallel-kernel-bench

HardwareDGX agent

ParallelKernelBench is a benchmarking framework designed to evaluate large language models' ability to generate optimized GPU kernels for multi-GPU computing environments. The benchmark assesses LLMs

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

HardwareDGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU

HardwareDGX agent

arXiv:2603.16428v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) has become essential for domain adaptation, but its memory-intensive property exceeds the capabilitie

← Previous
1…678910…29
Next →