AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,477 results
13 Apr 2026

Shares of Dell and HP jump after a report said Nvidia 'has been in negotiations for over a year to buy a large company and it will reshape the PC landscape' (Dina Bass/Bloomberg)

HardwareDGX agent

Dina Bass / Bloomberg: Shares of Dell and HP jump after a report said Nvidia “has been in negotiations for over a year to buy a large company and it will reshape the PC landscape” — Shares of Dell Tec

TensorHub: Scalable and Elastic Weight Transfer for LLM RL Training

HardwareDGX agent

arXiv:2604.09107v1 Announce Type: cross Abstract: Modern LLM reinforcement learning (RL) workloads require a highly efficient weight transfer system to scale training across heterogeneous computationa

TurboOCR: 270–1200 img/s OCR with Paddle + TensorRT (C++/CUDA, FP16) [P]

HardwareDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

TurboOCR is a high-performance OCR project that combines PaddleOCR with NVIDIA TensorRT, implemented in C++ and CUDA, achieving throughput of 270–1,200 images per second using FP16 half-precision infe

U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster

HardwareDGX agent

arXiv:2604.09041v1 Announce Type: cross Abstract: AI-based weather forecasting now rivals traditional physics-based ensembles, but state-of-the-art (SOTA) models rely on specialized architectures and

12 Apr 2026

MiniMax M2.7 Advances Scalable Agentic Workflows on NVIDIA Platforms for Complex AI Applications

HardwareDGX agent

MiniMax M2.7 is an enhancement of the MiniMax M2.5 model, built as a 230B-parameter Mixture-of-Experts (MoE) model with 10B active parameters per token and a 200K context length, designed for agentic

11 Apr 2026

NVIDIA’s New AI Shouldn’t Work…But It Does

HardwareDGX agent

This is a Two Minute Papers episode by Dr. Károly Zsolnai-Fehér reviewing a counterintuitive NVIDIA AI research result — likely covering a technique that defies conventional expectations yet delivers

The Infinity Man

HardwareDGX agent

'The Infinity Man' is likely a profile or biographical piece published in The Chip Letter, a Substack newsletter focused on the history of semiconductors and computing, examining a key figure in the c

What's model should I run?

Local AiDGX agent

A Reddit discussion from the r/ollama community where a user seeks advice on which AI language model to run locally using Ollama. Responses likely include hardware-based recommendations (such as RAM a

10 Apr 2026

A Lightweight Library for Energy-Based Joint-Embedding Predictive Architectures

HardwareDGX agent

arXiv:2602.03604v3 Announce Type: replace-cross Abstract: We present EB-JEPA, an open-source library for learning representations and world models using Joint-Embedding Predictive Architectures (JEPAs

AgiPIX: Bridging Simulation and Reality in Indoor Aerial Inspection

Model ReleasesDGX agent

arXiv:2604.08009v1 Announce Type: new Abstract: Autonomous indoor flight for critical asset inspection presents fundamental challenges in perception, planning, control, and learning. Despite rapid pro

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference

HardwareDGX agent

arXiv:2604.06036v3 Announce Type: replace-cross Abstract: Video streaming analytics is a crucial workload for vision-language model serving, but the high cost of multimodal inference limits scalabilit

DynLP: Parallel Dynamic Batch Update for Label Propagation in Semi-Supervised Learning

HardwareDGX agent

arXiv:2604.06596v1 Announce Type: cross Abstract: Semi-supervised learning aims to infer class labels using only a small fraction of labeled data. In graph-based semi-supervised learning, this is typi

ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache

HardwareDGX agent

arXiv:2604.06370v1 Announce Type: cross Abstract: The serving paradigm of large language models (LLMs) is rapidly shifting towards complex multi-agent workflows where specialized agents collaborate ov

Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start

HardwareDGX agent

arXiv:2604.06664v1 Announce Type: cross Abstract: Modern LLM service providers increasingly rely on autoscaling and parallelism reconfiguration to respond to rapidly changing workloads, but cold-start

Full State-Space Visualisation of the 8-Puzzle: Feasibility, Design, and Educational Use

HardwareDGX agent

arXiv:2604.06186v1 Announce Type: cross Abstract: Search algorithms are a foundational topic in artificial intelligence education, yet even simple domains can generate large state spaces that challeng

Graph Neural ODE Digital Twins for Control-Oriented Reactor Thermal-Hydraulic Forecasting Under Partial Observability

HardwareDGX agent

arXiv:2604.07292v1 Announce Type: new Abstract: Real-time supervisory control of advanced reactors requires accurate forecasting of plant-wide thermal-hydraulic states, including locations where physi

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

HardwareDGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of …

HardwareDGX agent

LiteParse hit 4K+ GitHub stars in 3 weeks. ~500 pages in 2 seconds. No GPU. No API keys. 50+ file formats. Now @LoganMarkewich, our Head of Open Source, will show you how to build with it. Live worksh

Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning

HardwareDGX agent

arXiv:2604.07345v1 Announce Type: cross Abstract: The rapid growth of generative artificial intelligence (AI) has introduced unprecedented computational demands, driving significant increases in the e

National Robotics Week — Latest Physical AI Research, Breakthroughs and Resources

HardwareDGX agent

This National Robotics Week, NVIDIA is highlighting the breakthroughs that are bringing AI into the physical world — as well as the growing wave of robots transforming industries, from agricultural an

NestPipe: Large-Scale Recommendation Training on 1,500+ Accelerators via Nested Pipelining

HardwareDGX agent

arXiv:2604.06956v1 Announce Type: cross Abstract: Modern recommendation models have increased to trillions of parameters. As cluster scales expand to O(1k), distributed training bottlenecks shift from

Object-Centric Stereo Ranging for Autonomous Driving: From Dense Disparity to Census-Based Template Matching

HardwareDGX agent

arXiv:2604.07980v1 Announce Type: new Abstract: Accurate depth estimation is critical for autonomous driving perception systems, particularly for long range vehicle detection on highways. Traditional

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and …

HardwareDGX agent

Often discuss my three-level vision for opening GLM to the community: First, we focus on accessibility by lowering the barrier to entry and removing unnecessary constraints so developers can truly exp

OpenPRC: A Unified Open-Source Framework for Physics-to-Task Evaluation in Physical Reservoir Computing

HardwareDGX agent

arXiv:2604.07423v1 Announce Type: new Abstract: Physical Reservoir Computing (PRC) leverages the intrinsic nonlinear dynamics of physical substrates, mechanical, optical, spintronic, and beyond, as fi

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing

HardwareDGX agent

arXiv:2512.06774v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a leading representation for high-fidelity 3D assets, yet protecting these assets via digital watermarking r

RISC-V chip design startup SiFive nabs $400M investment

HardwareDGX agent

SiFive Inc., a startup that sells chip designs based on the open-source RISC-V architecture, has raised $400 million in funding at a $3.65 billion valuation. Atreides Management led the Series G round

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

HardwareDGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

Splats under Pressure: Exploring Performance-Energy Trade-offs in Real-Time 3D Gaussian Splatting under Constrained GPU Budgets

HardwareDGX agent

arXiv:2604.07177v1 Announce Type: cross Abstract: We investigate the feasibility of real-time 3D Gaussian Splatting (3DGS) rasterisation on edge clients with varying Gaussian splat counts and GPU comp

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us curren…

HardwareDGX agent

The Information wrote that Meta consumes 60.2T tokens per month, or ~750M per employee. Meta is getting absolutely token-mogged by us currently. SemiAnalysis employees consume 1.86B tokens per month.

The power of JAX

HardwareDGX agent

The power of JAX Introducing gyaradax 🐉: A JAX solver for local flux-tube gyrokinetics with custom CUDA kernels for acceleration. This entire code was vibecoded by @ggalletti_ and me in a month. Valid

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training

HardwareDGX agent

arXiv:2603.04385v3 Announce Type: replace Abstract: Feed-forward transformer models have driven rapid progress in 3D vision, but state-of-the-art methods such as VGGT and pi^3 have a computational

9 Apr 2026

How Estée Lauder Companies uses Cloud Run worker pools for its pull-based agentic workloads

HardwareDGX agent

Cloud Run has long provided developers with a straightforward, opinionated platform for running code. You can easily deploy request-driven web applications using Cloud Run services, or execute run-to-

New course: Efficient Inference with SGLang: Text and Image Generation, built in partnership with LMSys @lmsysorg and RadixArk @radixark, an…

HardwareDGX agent

New course: Efficient Inference with SGLang: Text and Image Generation, built in partnership with LMSys @lmsysorg and RadixArk @radixark, and taught by Richard Chen @richardczl, a Member of Technical

Nvidia published DWDP (Distributed Weight-Data Parallelism), a new inference parallelism strategy focused on prefill. It sounds slightly ins…

HardwareDGX agent

Nvidia published DWDP (Distributed Weight-Data Parallelism), a new inference parallelism strategy focused on prefill. It sounds slightly insane until you remember the target machine is GB200 NVL72. Th

Strength and Destiny Collide: ‘Samson: A Tyndalston Story’ Arrives in the Cloud

HardwareDGX agent

A timeless story of grit, faith and rebellion takes center stage as Samson: A Tyndalston Story joins the GeForce NOW library today. The highly anticipated release from Liquid Swords can now be streame

The sad thing is that Ro is the guy preaching for socialism while he is the most active insider trader in Congress. He is a terrible represe…

HardwareDGX agent

The sad thing is that Ro is the guy preaching for socialism while he is the most active insider trader in Congress. He is a terrible representative of Silicon Valley. Nancy Pelosi take a seat. There i

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at…

HardwareDGX agent

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at least some strategy, since the robot first has to spread th

8 Apr 2026

This is literally why I wrote Taming Silicon Valley.

HardwareDGX agent

This is literally why I wrote Taming Silicon Valley. Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: 'AGI will not emerge from a government funded

We're seeing even more autonomous AI coworkers. The new MLE agent on the market is Disarray. In Kaggle competitions, Disarray: - won 28 meda…

HardwareDGX agent

We're seeing even more autonomous AI coworkers. The new MLE agent on the market is Disarray. In Kaggle competitions, Disarray: - won 28 medals across diverse domains (vision, NLP, tabular data) - plac

13 Aug 2026

ggml-cpu/ops: vectorize flash-attention V-cache F16 to F32 conversion by jinzihao · Pull Request #26947 · ggml-org/llama.cpp

Model ReleasesDGX agent

Overview ggml_cpu_fp16_to_fp32 leverages hardware F16C intrinsics (AVX-512, AVX2, etc.), faster than the software-only ggml_fp16_to_fp32_row, bringing 17-31% gain in prompt processing rate for a small

Hybrid-LUT: Channel-Aware Hybrid Lookup Table and Filtering for Efficient Image Denoising

ApplicationsDGX agent

arXiv:2608.11646v1 Announce Type: new Abstract: Lookup table (LUT)-based image denoising methods have attracted increasing attention due to their high efficiency and hardware-friendly properties. Howe

LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Language Model Training

SafetyDGX agent

arXiv:2608.11919v1 Announce Type: new Abstract: Training large language models on limited hardware is increasingly a scheduling problem across GPU compute, host memory, PCIe transfer, and storage band

12 Aug 2026

Google launches five new Pixel devices, array of Gemini Intelligence features

Model ReleasesDGX agent

Google LLC today expanded its Pixel consumer hardware line with a foldable handset, three smartphones and a smartwatch. The devices will ship with an upgraded version of the company’s Gemini Intellige

11 Aug 2026

Epically Powerful: An open-source software and mechatronics infrastructure for wearable robotic systems

TutorialsDGX agent

arXiv:2511.05033v2 Announce Type: replace Abstract: Epically Powerful is an open-source robotics infrastructure that streamlines the underlying framework of wearable robotic systems - managing communi

Janus: An Algorithm-Evaluator Co-Evolution Framework for LLM-Driven Discovery under Expensive Evaluation Budgets

ResearchDGX agent

arXiv:2608.08189v1 Announce Type: new Abstract: LLM-driven program discovery relies on rapid evaluator feedback, but many scientific and engineering tasks require high-fidelity simulations, hardware e

QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture

Model ReleasesDGX agent

arXiv:2510.22087v3 Announce Type: replace-cross Abstract: The field of computer architecture, which bridges high-level software abstractions and low-level hardware implementations, remains absent from

Reconfigurable Structural Robotic Assembly: Interlocking 3D Aggregations with Self-Aligning Compound Nested Lattice Modules

SafetyDGX agent

arXiv:2608.07576v1 Announce Type: new Abstract: Robotic construction systems often treat the material system and the robot as separate design problems, locating intelligence primarily in hardware, sen

10 Aug 2026

DeepSeek V4 Flash 0731 is the ‘killer app’ that is going to sell A LOT of DGX Sparks

Model ReleasesDGX agent

Having a ‘Killer Application’ that everyone wants to use helps sell hardware, plain and simple. DeepSeek V4 Flash 0731 isn’t an app of course, but I think it’s going to be the major catalyst for getti

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

Model ReleasesDGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

Learning Fault-Tolerant Locomotion with Adaptive Gait Timing

Model ReleasesDGX agent

arXiv:2608.07328v1 Announce Type: cross Abstract: Hardware failures require legged robots to rapidly reorganize coordination and gait timing to maintain stability and mobility. This is particularly ch

8 Aug 2026

Extremely slow DSpark draft model performance (1-2 t/s) with DeepSeek-V4-Flash on llama-server compared to MTP?

Model ReleasesDGX agent

Hey everyone, I could use some advice on setting up speculative decoding correctly with llama-server. My Hardware: GPUs: RTX 4090 + RTX 6000 Pro (120GB total VRAM) RAM: 32GB I am currently testing the

6 Aug 2026

CheckOne: Lightweight Fault Detection and Mitigation for Vision Transformers

SafetyDGX agent

arXiv:2608.04035v1 Announce Type: cross Abstract: The wide adoption of Vision Transformers (ViTs) in safety-critical applications raises reliability concerns related to hardware faults. Algorithm-Base

LLM-based Vulnerability Discovery in Business Process Documentation

ResearchDGX agent

arXiv:2608.04271v1 Announce Type: cross Abstract: Just like software and hardware, business processes are susceptible to vulnerabilities that can lead to product quality issues, delays, and increased

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

ResearchDGX agent

The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling their use across diverse hardware and software plat

Report: OpenAI’s upcoming AI speaker will be shaped like a donut and cost around $300

IndustryDGX agent

More details have emerged from Bloomberg about OpenAI Group PBC’s long-awaited hardware device, which was previously described as an artificial intelligence-enabled speaker that will be a “physical ma

Representational separation between unitary and channel quantum generative models via shared classical randomness at shallow depth

Local AiDGX agent

arXiv:2608.05110v1 Announce Type: cross Abstract: Near-term quantum hardware limits circuit depth and often imposes geometrically local connectivity for quantum generative models, restricting the outp

5 Aug 2026

A Survey on Design Methodologies for Accelerating Deep Learning on Heterogeneous Architectures

ResearchDGX agent

arXiv:2311.17815v3 Announce Type: replace-cross Abstract: Given their increasing size and complexity, the need for efficient execution of deep neural networks has become increasingly pressing in the d

Deepseek V4 Flash just hit Colibri, does anyone have numbers?

Model ReleasesDGX agent

I'm mosty interested in 128-192GB VRAM with 128-256GB RAM to spare, so SSD streaming is basically not even necessary. Seems only FP4 is supported, so older hardware will likely be slow - no Unsloth GG

Morphology-Aware Implicit Super-Resolution Network for Pathological Images

Model ReleasesDGX agent

arXiv:2608.03664v1 Announce Type: new Abstract: Accurate diagnosis in Digital Pathology (DP) relies on high-resolution whole-slide images, yet clinical deployment is often limited by hardware costs. S

Scaling agentic AI: How UiPath built its high-performance GPU platform on AI Hypercomputer

Model ReleasesDGX agent

As a market leader in enterprise agentic automation and business orchestration, UiPath is helping to pioneer an industry shift toward agentic AI. With it, the company is deploying autonomous agents to

← Previous
1…3435363738…75
Next →