AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
13 Apr 2026

Nvidia hasn’t budged in six months. Coreweave is down 21% Oracle is down 50% Microsoft is down 25% Like it or not, the bubble has already st…

HardwareDGX agent

Gary Marcus, an AI skeptic and cognitive scientist, posted on X highlighting diverging stock performance in the AI infrastructure sector, noting that while Nvidia has remained relatively stable, CoreW

Ornn Compute Price Index: renting one Nvidia Blackwell GPU for an hour now costs 4.08, up 48% from 2.75 two months ago, driven by agentic AI demand (Wall Street Journal)

HardwareDGX agent

Wall Street Journal: Ornn Compute Price Index: renting one Nvidia Blackwell GPU for an hour now costs 4.08, up 48% from 2.75 two months ago, driven by agentic AI demand — AI companies are rationing of

Shares of Dell and HP jump after a report said Nvidia 'has been in negotiations for over a year to buy a large company and it will reshape the PC landscape' (Dina Bass/Bloomberg)


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
HardwareDGX agent

Dina Bass / Bloomberg: Shares of Dell and HP jump after a report said Nvidia “has been in negotiations for over a year to buy a large company and it will reshape the PC landscape” — Shares of Dell Tec

TensorHub: Scalable and Elastic Weight Transfer for LLM RL Training

HardwareDGX agent

arXiv:2604.09107v1 Announce Type: cross Abstract: Modern LLM reinforcement learning (RL) workloads require a highly efficient weight transfer system to scale training across heterogeneous computationa

TurboOCR: 270–1200 img/s OCR with Paddle + TensorRT (C++/CUDA, FP16) [P]

HardwareDGX agent

TurboOCR is a high-performance OCR project that combines PaddleOCR with NVIDIA TensorRT, implemented in C++ and CUDA, achieving throughput of 270–1,200 images per second using FP16 half-precision infe

U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster

HardwareDGX agent

arXiv:2604.09041v1 Announce Type: cross Abstract: AI-based weather forecasting now rivals traditional physics-based ensembles, but state-of-the-art (SOTA) models rely on specialized architectures and

12 Apr 2026

MiniMax M2.7 Advances Scalable Agentic Workflows on NVIDIA Platforms for Complex AI Applications

HardwareDGX agent

MiniMax M2.7 is an enhancement of the MiniMax M2.5 model, built as a 230B-parameter Mixture-of-Experts (MoE) model with 10B active parameters per token and a 200K context length, designed for agentic

11 Apr 2026

NVIDIA’s New AI Shouldn’t Work…But It Does

HardwareDGX agent

This is a Two Minute Papers episode by Dr. Károly Zsolnai-Fehér reviewing a counterintuitive NVIDIA AI research result — likely covering a technique that defies conventional expectations yet delivers

The Infinity Man

HardwareDGX agent

'The Infinity Man' is likely a profile or biographical piece published in The Chip Letter, a Substack newsletter focused on the history of semiconductors and computing, examining a key figure in the c

10 Apr 2026

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

HardwareDGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

🫡 @aiDotEngineer in London was extremely good. Absolutely incredible speaker line up, great curation, escaped the Silicon Valley bubble mas…

HardwareDGX agent

"**AI Engineer Europe (London) — Conference Recap**

Behind the Analysis with Google Cloud and Team USA: Architecting AI infrastructure for U.S. Winter Olympians

HardwareDGX agent

In freeskiing and snowboarding, traditional video replay shows you what happened during a complex aerial maneuver, but it fails to explain the physics of how it was possible. At the speed of the sport

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference

HardwareDGX agent

arXiv:2604.06036v3 Announce Type: replace-cross Abstract: Video streaming analytics is a crucial workload for vision-language model serving, but the high cost of multimodal inference limits scalabilit

DynLP: Parallel Dynamic Batch Update for Label Propagation in Semi-Supervised Learning

HardwareDGX agent

arXiv:2604.06596v1 Announce Type: cross Abstract: Semi-supervised learning aims to infer class labels using only a small fraction of labeled data. In graph-based semi-supervised learning, this is typi

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache

HardwareDGX agent

arXiv:2604.06370v1 Announce Type: cross Abstract: The serving paradigm of large language models (LLMs) is rapidly shifting towards complex multi-agent workflows where specialized agents collaborate ov

Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start

HardwareDGX agent

arXiv:2604.06664v1 Announce Type: cross Abstract: Modern LLM service providers increasingly rely on autoscaling and parallelism reconfiguration to respond to rapidly changing workloads, but cold-start

Full State-Space Visualisation of the 8-Puzzle: Feasibility, Design, and Educational Use

HardwareDGX agent

arXiv:2604.06186v1 Announce Type: cross Abstract: Search algorithms are a foundational topic in artificial intelligence education, yet even simple domains can generate large state spaces that challeng

Graph Neural ODE Digital Twins for Control-Oriented Reactor Thermal-Hydraulic Forecasting Under Partial Observability

HardwareDGX agent

arXiv:2604.07292v1 Announce Type: new Abstract: Real-time supervisory control of advanced reactors requires accurate forecasting of plant-wide thermal-hydraulic states, including locations where physi

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

HardwareDGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

LAsset: An LLM-assisted Security Asset Identification Framework for System-on-Chip (SoC) Verification

HardwareDGX agent

arXiv:2601.02624v2 Announce Type: replace-cross Abstract: The growing complexity of modern system-on-chip (SoC) and IP designs is making security assurance difficult day by day. One of the fundamental

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of …

HardwareDGX agent

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of Open Source, will show you how to build with it. Live worksh

Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning

HardwareDGX agent

arXiv:2604.07345v1 Announce Type: cross Abstract: The rapid growth of generative artificial intelligence (AI) has introduced unprecedented computational demands, driving significant increases in the e

National Robotics Week — Latest Physical AI Research, Breakthroughs and Resources

HardwareDGX agent

This National Robotics Week, NVIDIA is highlighting the breakthroughs that are bringing AI into the physical world — as well as the growing wave of robots transforming industries, from agricultural an

NestPipe: Large-Scale Recommendation Training on 1,500+ Accelerators via Nested Pipelining

HardwareDGX agent

arXiv:2604.06956v1 Announce Type: cross Abstract: Modern recommendation models have increased to trillions of parameters. As cluster scales expand to O(1k), distributed training bottlenecks shift from

Object-Centric Stereo Ranging for Autonomous Driving: From Dense Disparity to Census-Based Template Matching

HardwareDGX agent

arXiv:2604.07980v1 Announce Type: new Abstract: Accurate depth estimation is critical for autonomous driving perception systems, particularly for long range vehicle detection on highways. Traditional

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and …

HardwareDGX agent

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and removing unnecessary constraints so developers can truly exp

OpenPRC: A Unified Open-Source Framework for Physics-to-Task Evaluation in Physical Reservoir Computing

HardwareDGX agent

arXiv:2604.07423v1 Announce Type: new Abstract: Physical Reservoir Computing (PRC) leverages the intrinsic nonlinear dynamics of physical substrates, mechanical, optical, spintronic, and beyond, as fi

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing

HardwareDGX agent

arXiv:2512.06774v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a leading representation for high-fidelity 3D assets, yet protecting these assets via digital watermarking r

RISC-V chip design startup SiFive nabs $400M investment

HardwareDGX agent

SiFive Inc., a startup that sells chip designs based on the open-source RISC-V architecture, has raised $400 million in funding at a $3.65 billion valuation. Atreides Management led the Series G round

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

HardwareDGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

Splats under Pressure: Exploring Performance-Energy Trade-offs in Real-Time 3D Gaussian Splatting under Constrained GPU Budgets

HardwareDGX agent

arXiv:2604.07177v1 Announce Type: cross Abstract: We investigate the feasibility of real-time 3D Gaussian Splatting (3DGS) rasterisation on edge clients with varying Gaussian splat counts and GPU comp

That’s a wrap on HumanX. Custom comics, hats, a happy hour with @getmetronome & @nvidia, and two sessions on what actually matters for AI-na…

HardwareDGX agent

Together Computer (Together AI) participated in HumanX 2026, a major AI conference held April 6–9 in San Francisco, where they hosted activations including custom comics, branded hats, and a happy ...

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us curren…

HardwareDGX agent

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us currently. SemiAnalysis employees consume 1.86B tokens per month.

The power of JAX

HardwareDGX agent

The power of JAX Introducing gyaradax 🐉: A JAX solver for local flux-tube gyrokinetics with custom CUDA kernels for acceleration. This entire code was vibecoded by @ggalletti_ and me in a month. Valid

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training

HardwareDGX agent

arXiv:2603.04385v3 Announce Type: replace Abstract: Feed-forward transformer models have driven rapid progress in 3D vision, but state-of-the-art methods such as VGGT and pi^3 have a computational

9 Apr 2026

(1/5) FP4 hardware is here, but 4-bit attention still kills model quality, blocking true end-to-end FP4 serving. To fix that, we propose Att…

HardwareDGX agent

(1/5) FP4 hardware is here, but 4-bit attention still kills model quality, blocking true end-to-end FP4 serving. To fix that, we propose Attn-QAT, the first systematic study of quantization-aware trai

Cut Checkpoint Costs with About 30 Lines of Python and NVIDIA nvCOMP

HardwareDGX agent

Training LLMs requires periodic checkpoints — full snapshots of model weights, optimizer states, and gradients — whose storage costs can reach $200,000/month for a 405B model on 128 NVIDIA DGX B200...

How Estée Lauder Companies uses Cloud Run worker pools for its pull-based agentic workloads

HardwareDGX agent

Cloud Run has long provided developers with a straightforward, opinionated platform for running code. You can easily deploy request-driven web applications using Cloud Run services, or execute run-to-

How to Accelerate Protein Structure Prediction at Proteome-Scale

HardwareDGX agent

NVIDIA, Google DeepMind, EMBL-EBI, and Seoul National University collaborated to extend the AlphaFold Protein Structure Database (AFDB) beyond monomeric structures to proteome-scale quaternary stru...

New course: Efficient Inference with SGLang: Text and Image Generation, built in partnership with LMSys @lmsysorg and RadixArk @radixark, an…

HardwareDGX agent

New course: Efficient Inference with SGLang: Text and Image Generation, built in partnership with LMSys @lmsysorg and RadixArk @radixark, and taught by Richard Chen @richardczl, a Member of Technical

Nvidia published DWDP (Distributed Weight-Data Parallelism), a new inference parallelism strategy focused on prefill. It sounds slightly ins…

HardwareDGX agent

Nvidia published DWDP (Distributed Weight-Data Parallelism), a new inference parallelism strategy focused on prefill. It sounds slightly insane until you remember the target machine is GB200 NVL72. Th

Running Large-Scale GPU Workloads on Kubernetes with Slurm

HardwareDGX agent

NVIDIA's open-source project **Slinky** (developed by SchedMD, now part of NVIDIA) enables organizations to run full Slurm clusters directly on Kubernetes infrastructure by managing the complete li...

Strength and Destiny Collide: ‘Samson: A Tyndalston Story’ Arrives in the Cloud

HardwareDGX agent

A timeless story of grit, faith and rebellion takes center stage as Samson: A Tyndalston Story joins the GeForce NOW library today. The highly anticipated release from Liquid Swords can now be streame

The sad thing is that Ro is the guy preaching for socialism while he is the most active insider trader in Congress. He is a terrible represe…

HardwareDGX agent

The sad thing is that Ro is the guy preaching for socialism while he is the most active insider trader in Congress. He is a terrible representative of Silicon Valley. Nancy Pelosi take a seat. There i

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at…

HardwareDGX agent

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at least some strategy, since the robot first has to spread th

8 Apr 2026

Integrate Physical AI Capabilities into Existing Apps with NVIDIA Omniverse Libraries

HardwareDGX agent

NVIDIA has introduced a modular, library-based architecture for Omniverse, exposing core components—RTX rendering, PhysX-based simulation, and data storage pipelines—as standalone, headless-first C...

New Vultr Identity and Access Management Upgrades

HardwareDGX agent

Vultr has upgraded its Identity and Access Management (IAM) framework by replacing legacy ACLs with fine-grained, policy-driven access controls, enabling organizations to assign roles, manage permi...

This is literally why I wrote Taming Silicon Valley.

HardwareDGX agent

This is literally why I wrote Taming Silicon Valley. Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government funded

We're seeing even more autonomous AI coworkers. The new MLE agent on the market is Disarray. In Kaggle competitions, Disarray: - won 28 meda…

HardwareDGX agent

We're seeing even more autonomous AI coworkers. The new MLE agent on the market is Disarray. In Kaggle competitions, Disarray: - won 28 medals across diverse domains (vision, NLP, tabular data) - plac

7 Apr 2026

NVIDIA’s New AI Just Changed Everything

HardwareDGX agent

I was unable to directly access the YouTube video at the provided URL, and my search did not return results specifically identifying that video's content. YouTube videos cannot be fetched or read a...

← Previous
1…272829
Next →