AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
Hardware

Lights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy Consumption

DGX agent

arXiv:2607.04553v1 Announce Type: cross Abstract: We present a bidirectional framework for estimating the energy consumption of text-to-video (T2V) and text-to-video-audio (T2VA) models from architect

hardwarearxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community

DGX agent

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development can still be gated b

hardwarenvidia-blog
7 Jul 2026
Hardware

NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads

DGX agent

NVIDIA Vera is a CPU designed to help AI factories scale agentic AI and reinforcement learning by shortening CPU execution time, increasing task throughput, and enabling smarter, longer-thinking agent

hardwarenvidia-developer
7 Jul 2026
Hardware

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

DGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization

DGX agent

arXiv:2607.02834v1 Announce Type: new Abstract: Molecular optimization often starts from a pretrained generative model that captures a broad prior over valid molecular structures. At test time, howeve

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the lates…

DGX agent

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the latest Isaac models and frameworks to the open robotics community.

hardwareclem-delangue--x
7 Jul 2026
Hardware

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

DGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

DGX agent

arXiv:2504.08791v3 Announce Type: replace-cross Abstract: On-device inference offers privacy, offline use, and instant response, but consumer hardware restricts large language models (LLMs) to low thr

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

RoMa v2: Harder Better Faster Denser Feature Matching

DGX agent

arXiv:2511.15706v3 Announce Type: replace Abstract: Dense feature matching aims to estimate all correspondences between two images of a 3D scene and has recently been established as the gold standard

hardwarearxiv-cs-cv
7 Jul 2026
Hardware

SILO: Simulation-in-the-Loop Sim-to-Real Transfer for Multi-Stage Cable Routing

DGX agent

arXiv:2607.04616v1 Announce Type: cross Abstract: Linear-deformable manipulation remains challenging due to the complex deformations of objects such as cables and ropes. Prior data-driven approaches,

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

DGX agent

arXiv:2607.03333v1 Announce Type: cross Abstract: LLM agents are becoming a common interface for research, coding, and question answering, yet their Thought-Action-Observation loop is often serial: th

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

TabPack: Efficient Hyperparameter Ensembles for Tabular Deep Learning

DGX agent

arXiv:2607.05380v1 Announce Type: new Abstract: In deep learning for tabular data, efficient ensembles of multilayer perceptrons (MLPs) have recently emerged as effective and practical architectures.

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a w…

DGX agent

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a welcoming audience for my workshops and presentation. Thanks

hardwareswyx--x
7 Jul 2026
Hardware

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kerne…

DGX agent

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kernels all working together to make inference faster for users. P

hardwaretogether-ai--x
7 Jul 2026
Hardware

UK-based AI infrastructure startup Nscale secures a $900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific (Mauro Orru/Wall Street Journal)

DGX agent

Mauro Orru / Wall Street Journal: UK-based AI infrastructure startup Nscale secures a 900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific — Nscale raised

hardwaretechmeme
7 Jul 2026
Hardware

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microso…

DGX agent

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microsoft, Amazon, Google AI Debt Financing will be over $7T of deb

hardwaredylan-patel--x
6 Jul 2026
Hardware

Enhancing Goodput in Large-Scale LLM Training with Nonuniform Tensor Parallelism

DGX agent

Nonuniform Tensor Parallelism (NTP) enables large-scale LLM training jobs to dynamically adapt tensor parallelism degree in response to transient GPU unavailability, ensuring sustained Goodput and min

hardwarenvidia-developer
6 Jul 2026
Hardware

How Open Models Are Driving AI Research

DGX agent

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open

hardwarenvidia-blog
6 Jul 2026
Hardware

Import AI 464: Fable writes GPU kernels; AI automation; and analog computation

DGX agent

This Import AI newsletter issue covers Fable's development of GPU kernel writing capabilities, likely discussing how AI systems can automatically generate optimized code for graphics processors, along

hardwareimport-ai
6 Jul 2026
Hardware

Join us for a fireside chat on where AI research and infrastructure are headed, led by @tri_dao. Hosted by Together AI, @nvidia and Lyra Lab…

DGX agent

Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro

hardwaretogether-ai--x
6 Jul 2026
Hardware

Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters

DGX agent

Over 7T AI debt by 2029, There can be no Neoclouds without the Trinity. Nvidia's Backstop Economics Explained. AI Debt Needs Quantified. Nvidia's Objective is to Broaden Compute Access, Develop AI Fin

hardwaresemianalysis
6 Jul 2026
Hardware

SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture (Anniek Bao/CNBC)

DGX agent

Anniek Bao / CNBC: SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture — NVIDIA's next marq

hardwaretechmeme
6 Jul 2026
Hardware

Shanghai-based AI chipmaker Biren raises ~$892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO (Ann Cao/South China Morning Post)

DGX agent

Ann Cao / South China Morning Post: Shanghai-based AI chipmaker Biren raises ~892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO — Chinese ar

hardwaretechmeme
6 Jul 2026
Hardware

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs …

DGX agent

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs that also would have killed the company. He realized that th

hardwareemad-mostaque--x
6 Jul 2026
Hardware

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemming…

DGX agent

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemmings. Ttypical East Coast VCs all think that they're smarter th

hardwareyann-lecun--x
5 Jul 2026
Hardware

First animation from integrated PyTTI C++/CUDA port. I used DD style (@dango233max) CLIP cutouts since that's what I already had in the code…

DGX agent

This post showcases an early animation demonstration from a C++/CUDA implementation of PyTTI (Python Text-to-Image), utilizing CLIP cutout techniques in the style of DD (likely referring to a specific

hardwareemad-mostaque--x
4 Jul 2026
Hardware

God bless America the land of the free and home of the brave 🫡 250

DGX agent

This post likely references the patriotic phrase from the U.S. national anthem, possibly in the context of Dylan Patel's technology industry commentary or analysis. Without access to the specific cont

hardwaredylan-patel--x
4 Jul 2026
Hardware

Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M (Madison McLauchlan/BetaKit)

DGX agent

Madison McLauchlan / BetaKit: Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M — Stathera

hardwaretechmeme
4 Jul 2026
Model Releases

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises ju…

DGX agent

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises just woke up to a trap they had been walking into, here's how

model-releasesclem-delangue--x
4 Jul 2026
Hardware

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I hav…

DGX agent

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I haven’t yet heard from industry leaders Open source AI is about

hardwareclem-delangue--x
4 Jul 2026
Hardware

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

DGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

DGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there?

DGX agent

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there? Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference

hardwarefireworks-ai--x
3 Jul 2026
Hardware

EHHN: An Event-driven Heterogeneous Hypergraph Network for Object-Centric Next Activity Prediction

DGX agent

arXiv:2607.01785v1 Announce Type: new Abstract: Next activity prediction helps service-oriented processes anticipate upcoming steps before delays, exceptions, or service-level risks occur. Most existi

hardwarearxiv-cs-lg
3 Jul 2026
Hardware

PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation

DGX agent

arXiv:2607.01803v1 Announce Type: cross Abstract: Recent advances in 3D content generation from text or images have achieved impressive results, yet view inconsistency from 2D generators and the scarc

hardwarearxiv-cs-ro
3 Jul 2026
Hardware

Stable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory Gates

DGX agent

arXiv:2607.02363v1 Announce Type: cross Abstract: Quantum Fast-Weight Programmers (QFWPs) store temporal information in dynamically programmed variational-circuit parameters rather than in nonlinear r

hardwarearxiv-cs-ai
3 Jul 2026
Hardware

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines t…

DGX agent

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines to serve agentic workloads at trillion token production scale

hardwaretogether-ai--x
3 Jul 2026
Hardware

Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

DGX agent

arXiv:2607.00773v1 Announce Type: new Abstract: Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the ac

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

Condensing Large-Scale Datasets Directly with Minimal Information Loss

DGX agent

arXiv:2607.00916v1 Announce Type: new Abstract: Recent advancements in scaling dataset distillation rely heavily on decoupled information extraction pipelines, comprising SQUEEZE, RECOVER, and RELABEL

hardwarearxiv-cs-cv
2 Jul 2026
Hardware

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

DGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

EMIB-T Roadmap, Custom HBM, HBM4 Packaging Challenges, Microfluidic Cooling, Photonic Interconnects, and More

DGX agent

This SemiAnalysis article discusses advanced semiconductor packaging technologies and innovations presented at or related to ECTC 2026, including Intel's EMIB-T chiplet interconnect roadmap, custom HB

hardwaresemianalysis
2 Jul 2026
Hardware

Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan (New York Times)

DGX agent

New York Times: Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan — President Trump and hi

hardwaretechmeme
2 Jul 2026
Hardware

GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics

DGX agent

arXiv:2607.01203v1 Announce Type: cross Abstract: This paper studies real-time robust optimal control for uncertain nonlinear systems, where linear time-varying (LTV) approximations make planning trac

hardwarearxiv-cs-ai
2 Jul 2026
Hardware

Joyride Through July With 12 Games Coming to GeForce NOW

DGX agent

Summer is heating up — and GeForce NOW is taking players along for the ride. Start the month with Monopoly: Star Wars Heroes vs. Villains, bringing a galaxy far, far away to the iconic board-game fran

hardwarenvidia-blog
2 Jul 2026
Hardware

Meta Compute: Everyone Wants To Be A Cloud

DGX agent

Meta Compute explores Meta's strategy to position itself as a cloud infrastructure provider, leveraging its massive internal compute investments and AI capabilities to offer services to external custo

hardwaresemianalysis
2 Jul 2026
Hardware

MosaicKV: Serving Long-Context LLM with Dynamic Two-D KV Cache Compression

DGX agent

arXiv:2607.00760v1 Announce Type: new Abstract: Long-context LLM services now sustain prompts with hundreds of thousands to millions of tokens, making the key-value (KV) cache a first-order serving co

hardwarearxiv-cs-lg
2 Jul 2026
Hardware

NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout

DGX agent

As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that generate tokens at scale. This shift requires ac

hardwarenvidia-blog
2 Jul 2026
Hardware

Palantir CEO Alex Karp doesn’t hold back in interview as he rails against AI industry

DGX agent

Palantir Technologies Inc. Chief Executive Alex Karp appeared to go into meltdown mode during an interview with CNBC today where for 20-odd minutes he went off script after being asked to discuss his

hardwaresiliconangle
2 Jul 2026
← Previous
1…1920212223…94
Next →