AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
7 Jul 2026

RoMa v2: Harder Better Faster Denser Feature Matching

HardwareDGX agent

arXiv:2511.15706v3 Announce Type: replace Abstract: Dense feature matching aims to estimate all correspondences between two images of a 3D scene and has recently been established as the gold standard

SILO: Simulation-in-the-Loop Sim-to-Real Transfer for Multi-Stage Cable Routing

HardwareDGX agent

arXiv:2607.04616v1 Announce Type: cross Abstract: Linear-deformable manipulation remains challenging due to the complex deformations of objects such as cables and ropes. Prior data-driven approaches,

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

HardwareDGX agent

arXiv:2607.03333v1 Announce Type: cross Abstract: LLM agents are becoming a common interface for research, coding, and question answering, yet their Thought-Action-Observation loop is often serial: th

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TabPack: Efficient Hyperparameter Ensembles for Tabular Deep Learning

HardwareDGX agent

arXiv:2607.05380v1 Announce Type: new Abstract: In deep learning for tabular data, efficient ensembles of multilayer perceptrons (MLPs) have recently emerged as effective and practical architectures.

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a w…

HardwareDGX agent

That’s a wrap on AIE from their new HQ! Honored to have been a guest panelist for Nvidia’s Local Al State of the Union, and to have such a welcoming audience for my workshops and presentation. Thanks

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kerne…

HardwareDGX agent

This is what latency optimization looks like below the API 👇 Together ATLAS, NVIDIA Blackwell, CUDA, TensorRT-LLM, Dynamo, and custom kernels all working together to make inference faster for users. P

UK-based AI infrastructure startup Nscale secures a $900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific (Mauro Orru/Wall Street Journal)

HardwareDGX agent

Mauro Orru / Wall Street Journal: UK-based AI infrastructure startup Nscale secures a 900M line of credit to expand its data center buildout across Europe, the US, and the Asia Pacific — Nscale raised

6 Jul 2026

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microso…

HardwareDGX agent

Creative structures are needed to get GPUs in the hands of startups + other companies that aren't Meta, OpenAI, Anthropic, SpaceXAI, Microsoft, Amazon, Google AI Debt Financing will be over $7T of deb

Enhancing Goodput in Large-Scale LLM Training with Nonuniform Tensor Parallelism

HardwareDGX agent

Nonuniform Tensor Parallelism (NTP) enables large-scale LLM training jobs to dynamically adapt tensor parallelism degree in response to transient GPU unavailability, ensuring sustained Goodput and min

How Open Models Are Driving AI Research

HardwareDGX agent

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open

Import AI 464: Fable writes GPU kernels; AI automation; and analog computation

HardwareDGX agent

This Import AI newsletter issue covers Fable's development of GPU kernel writing capabilities, likely discussing how AI systems can automatically generate optimized code for graphics processors, along

Join us for a fireside chat on where AI research and infrastructure are headed, led by @tri_dao. Hosted by Together AI, @nvidia and Lyra Lab…

HardwareDGX agent

Together AI, in partnership with NVIDIA and Lyra Lab, hosted a fireside chat led by Tri Dao discussing current trends and future directions in AI research and infrastructure development. The event bro

Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters

HardwareDGX agent

Over 7T AI debt by 2029, There can be no Neoclouds without the Trinity. Nvidia's Backstop Economics Explained. AI Debt Needs Quantified. Nvidia's Objective is to Broaden Compute Access, Develop AI Fin

SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture (Anniek Bao/CNBC)

HardwareDGX agent

Anniek Bao / CNBC: SemiAnalysis: Nvidia delays its next-gen AI rack system Kyber NVL144 by 12+ months to 2028 due to PCB manufacturing issues, and cancels its NVL72x2 architecture — NVIDIA's next marq

Shanghai-based AI chipmaker Biren raises ~$892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO (Ann Cao/South China Morning Post)

HardwareDGX agent

Ann Cao / South China Morning Post: Shanghai-based AI chipmaker Biren raises ~892.5M in a new share sale to boost GPU production; Biren's stock is up 150%+ since its January Hong Kong IPO — Chinese ar

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs …

HardwareDGX agent

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs that also would have killed the company. He realized that th

5 Jul 2026

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemming…

HardwareDGX agent

Groq Founder @JonathanRoss321 explains why all the West Coast VCs missed on investing in Groq: “Typical West Coast VCs are more like lemmings. Ttypical East Coast VCs all think that they're smarter th

4 Jul 2026

First animation from integrated PyTTI C++/CUDA port. I used DD style (@dango233max) CLIP cutouts since that's what I already had in the code…

HardwareDGX agent

This post showcases an early animation demonstration from a C++/CUDA implementation of PyTTI (Python Text-to-Image), utilizing CLIP cutout techniques in the style of DD (likely referring to a specific

God bless America the land of the free and home of the brave 🫡 250

HardwareDGX agent

This post likely references the patriotic phrase from the U.S. national anthem, possibly in the context of Dylan Patel's technology industry commentary or analysis. Without access to the specific cont

Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M (Madison McLauchlan/BetaKit)

HardwareDGX agent

Madison McLauchlan / BetaKit: Montreal-based Stathera, a maker of MEMS-based silicon timing components for chips, raised a 55M Series B led by Maverick Silicon, taking total funding to 75M — Stathera

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises ju…

Model ReleasesDGX agent

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises just woke up to a trap they had been walking into, here's how

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I hav…

HardwareDGX agent

Wow… this is amazing This latest episode of the @theallinpod is a complete referendum of closed source AI (ie. Anthropic/OAI) in a way I haven’t yet heard from industry leaders Open source AI is about

3 Jul 2026

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

HardwareDGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

HardwareDGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there?

HardwareDGX agent

Don't forget, we'll be in Paris for @RaiseSummit. Will we see you there? Going to be in Paris for @RaiseSummit? Join us and @nvidia to hear from leaders at both companies on the future of AI inference

EHHN: An Event-driven Heterogeneous Hypergraph Network for Object-Centric Next Activity Prediction

HardwareDGX agent

arXiv:2607.01785v1 Announce Type: new Abstract: Next activity prediction helps service-oriented processes anticipate upcoming steps before delays, exceptions, or service-level risks occur. Most existi

PixGS: Pixel-Space Diffusion for Direct 3D Gaussian Splat Generation

HardwareDGX agent

arXiv:2607.01803v1 Announce Type: cross Abstract: Recent advances in 3D content generation from text or images have achieved impressive results, yet view inconsistency from 2D generators and the scarc

Stable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory Gates

HardwareDGX agent

arXiv:2607.02363v1 Announce Type: cross Abstract: Quantum Fast-Weight Programmers (QFWPs) store temporal information in dynamically programmed variational-circuit parameters rather than in nonlinear r

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines t…

HardwareDGX agent

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines to serve agentic workloads at trillion token production scale

2 Jul 2026

Accelerating Discrete Diffusion Models with Parallel-In-Time Sampling

HardwareDGX agent

arXiv:2607.00773v1 Announce Type: new Abstract: Discrete diffusion models are widely used for learning and generating discrete distributions. As the generation process is inherently sequential, the ac

Condensing Large-Scale Datasets Directly with Minimal Information Loss

HardwareDGX agent

arXiv:2607.00916v1 Announce Type: new Abstract: Recent advancements in scaling dataset distillation rely heavily on decoupled information extraction pipelines, comprising SQUEEZE, RECOVER, and RELABEL

Efficient Compression of Structured and Unstructured Volumes via Learned 3D Gaussian Representation

HardwareDGX agent

arXiv:2607.01164v1 Announce Type: new Abstract: Recent work has shown that implicit neural representations (INRs) can be trained to effectively compress structured and unstructured volume data, allowi

EMIB-T Roadmap, Custom HBM, HBM4 Packaging Challenges, Microfluidic Cooling, Photonic Interconnects, and More

HardwareDGX agent

This SemiAnalysis article discusses advanced semiconductor packaging technologies and innovations presented at or related to ECTC 2026, including Intel's EMIB-T chiplet interconnect roadmap, custom HB

Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan (New York Times)

HardwareDGX agent

New York Times: Filings: President Trump purchased up to $5M each in Broadcom, Meta, Amazon, Apple, Microsoft, and Nvidia stocks on July 23, the same day as his AI action plan — President Trump and hi

GPU-Parallel Linearization Error Bounds for Real-Time Robust Optimal Control of Nonlinear and Neural Network Dynamics

HardwareDGX agent

arXiv:2607.01203v1 Announce Type: cross Abstract: This paper studies real-time robust optimal control for uncertain nonlinear systems, where linear time-varying (LTV) approximations make planning trac

Joyride Through July With 12 Games Coming to GeForce NOW

HardwareDGX agent

Summer is heating up — and GeForce NOW is taking players along for the ride. Start the month with Monopoly: Star Wars Heroes vs. Villains, bringing a galaxy far, far away to the iconic board-game fran

Meta Compute: Everyone Wants To Be A Cloud

HardwareDGX agent

Meta Compute explores Meta's strategy to position itself as a cloud infrastructure provider, leveraging its massive internal compute investments and AI capabilities to offer services to external custo

MosaicKV: Serving Long-Context LLM with Dynamic Two-D KV Cache Compression

HardwareDGX agent

arXiv:2607.00760v1 Announce Type: new Abstract: Long-context LLM services now sustain prompts with hundreds of thousands to millions of tokens, making the key-value (KV) cache a first-order serving co

NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout

HardwareDGX agent

As AI moves from model development to production inference, compute demand is accelerating and shifting toward continuously operating AI factories that generate tokens at scale. This shift requires ac

Palantir CEO Alex Karp doesn’t hold back in interview as he rails against AI industry

HardwareDGX agent

Palantir Technologies Inc. Chief Executive Alex Karp appeared to go into meltdown mode during an interview with CNBC today where for 20-odd minutes he went off script after being asked to discuss his

ROSA: A Robotics Foundation Model Serving System for Robot Factories

HardwareDGX agent

arXiv:2607.01088v1 Announce Type: new Abstract: Robotics foundation models (RFMs) are making general-purpose robots increasingly practical for factory deployments. While RFM serving systems are centra

SNAP-FM: Sparse Nonlinear Accelerated Projection for Physics-Constrained Generative Modeling

HardwareDGX agent

arXiv:2607.00095v1 Announce Type: cross Abstract: Generative models have emerged as scalable surrogates for physical simulation, yet they offer no guarantee that their outputs respect the conservation

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Pa…

HardwareDGX agent

🌍 World’s Largest Hermes Buildathon 10 cities. 5 countries. $750K credits Backed by OpenAI, Convex, Cloudflare, ElevenLabs, Linkup, Dodo Payments, Hissa Fund & Wispr Flow. Proudly presented by GrowthX

1 Jul 2026

9/ ParallelKernelBench: Benchmarking LLMs on Multi-GPU Kernel Generation Paper: https://www.alphaxiv.org/abs/2606.parallel-kernel-bench

HardwareDGX agent

ParallelKernelBench is a benchmarking framework designed to evaluate large language models' ability to generate optimized GPU kernels for multi-GPU computing environments. The benchmark assesses LLMs

An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU

HardwareDGX agent

arXiv:2603.16428v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) has become essential for domain adaptation, but its memory-intensive property exceeds the capabilitie

How we keep GPUs reliable across Databricks AI

HardwareDGX agent

Databricks implements reliability measures and monitoring systems to ensure consistent GPU performance and availability across its AI platform infrastructure. The article likely covers their approache

Lambda’s keynote at the ALVR workshop co-located with ACL 2026

HardwareDGX agent

Lambda Labs presented a keynote address at the ALVR (Augmented Language and Vision Research) workshop, which was held in conjunction with ACL 2026, a major computational linguistics conference. The pr

Learning Video Dynamics with Predictive Differentiable Rendering

HardwareDGX agent

arXiv:2606.31050v1 Announce Type: cross Abstract: How to accurately predict a high-fidelity future world? While the visual world is inherently continuous, existing deterministic video prediction model

Look who stopped by today! Great time chatting with @dylan522p.

HardwareDGX agent

Dylan Patel, a notable tech industry figure, was visited by the post author (Sassine Ghazi), and they had a positive conversation together. The post appears to be a casual social media update sharing

NVIDIA and Partners Build in America, for America

HardwareDGX agent

NVIDIA and its partners are investing in American manufacturing, supply chains, energy grids and skilled workforces so the U.S. can produce the infrastructure needed for better healthcare, breakthroug

Our research team has 9 papers at ICML next week! Spanning the full stack from frontier agents to GPU kernels, we're excited to share what o…

HardwareDGX agent

Our research team has 9 papers at ICML next week! Spanning the full stack from frontier agents to GPU kernels, we're excited to share what our researchers and collaborators have been working on. If yo

Raja Koduri’s Oxmiq Labs raises $35M to lower the design cost of custom AI silicon

HardwareDGX agent

Artificial intelligence chipmaking startup Oxmiq Labs Inc. says it wants to become the next Arm Holdings Plc. after raising 35 million in an early-stage A funding, bringing its total amount raised to

SeKV: Resolution-Adaptive KV Cache with Hierarchical Semantic Memory for Long-Context LLM Inference

HardwareDGX agent

arXiv:2606.31145v1 Announce Type: new Abstract: Large language models increasingly operate over long contexts, where the KV cache becomes a dominant memory bottleneck: its size grows linearly with seq

Taiwanese authorities detain two Super Micro staff and an Albatron manager after a raid of Super Micro's local offices this week over Nvidia shipments to China (Bloomberg)

HardwareDGX agent

Bloomberg: Taiwanese authorities detain two Super Micro staff and an Albatron manager after a raid of Super Micro's local offices this week over Nvidia shipments to China — Taiwanese prosecutors detai

Together AI raises $800M to grow its AI-optimized public cloud

HardwareDGX agent

Together AI Inc., the operator of a cloud platform optimized to run open-source artificial intelligence models, has raised 800 million from investors. The startup stated in its funding announcement to

Verkada takes Nvidia investment to expand its physical AI platform

HardwareDGX agent

Physical security company Verkada Inc. has taken an investment from Nvidia Corp. and signed a technical partnership with the chipmaker, the two said today, in a deal meant to speed up the artificial i

When boomer companies get high Anthropic bill, they set spend limits When I see a nearly million dollar Anthropic monthly bill, my first rea…

HardwareDGX agent

When boomer companies get high Anthropic bill, they set spend limits When I see a nearly million dollar Anthropic monthly bill, my first reaction is to complain about use of shitty Haiku models Imagin

30 Jun 2026

A Trainable-by-Parts Operator Learning Framework: Bridging DeepONet and Karhunen-Loeve Expansions for Large-Scale Applications

HardwareDGX agent

arXiv:2606.28519v1 Announce Type: new Abstract: Training operator-learning models for large-scale problems governed by partial differential equations (PDEs) is challenging due to the curse of dimensio

AI inference startup Etched raised 800M from investors including Jane Street and a TSMC-linked venture firm, and says it has signed sales contracts worth 1B (Dina Bass/Bloomberg)

HardwareDGX agent

Dina Bass / Bloomberg: AI inference startup Etched raised 800M from investors including Jane Street and a TSMC-linked venture firm, and says it has signed sales contracts worth 1B — Nvidia rival says

Building tech in the world’s secret R&D hub

HardwareDGX agent

Apple. Anthropic. Disney Research. Google. Meta. Microsoft. NVIDIA. OpenAI. Few places outside Silicon Valley can claim R&D hubs from all of these companies. Fewer still are concentrated in a city of

← Previous
1…1516171819…75
Next →