AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,477 results
21 Apr 2026

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, an…

HardwareDGX agent

I tested @huggingface ml-intern, given the prompt 'Fine-tune a Segment Anything Model (SAM) on a useful medical dataset. Train the model, and provide a comprehensive tutorial in a Jupyter Notebook fil

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

HardwareDGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
HardwareDGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

LoRaQ: Optimized Low Rank Approximation for 4-bit Quantization

ResearchDGX agent

arXiv:2604.18117v1 Announce Type: new Abstract: Post-training quantization (PTQ) is essential for deploying large diffusion transformers on resource-constrained hardware, but aggressive 4-bit quantiza

Muscle-inspired magnetic actuators that push, pull, crawl, and grasp

HardwareDGX agent

arXiv:2604.18090v1 Announce Type: new Abstract: Functional magnetic composites capable of large deformation, load bearing, and multifunctional motion are essential for next-generation adaptive soft ro

Neptune: Advanced ML Operator Fusion for Locality and Parallelism on GPUs

HardwareDGX agent

arXiv:2510.08726v2 Announce Type: replace-cross Abstract: Operator fusion has become a key optimization for deep learning, which combines multiple deep learning operators to improve data reuse and red

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning

HardwareDGX agent

arXiv:2509.18169v3 Announce Type: replace-cross Abstract: Tasks on complex systems require high-precision numerical computation to support decisions, but current large language models (LLMs) cannot in

POLAR: Online Learning for LoRA Adapter Caching and Routing in Edge LLM Serving

HardwareDGX agent

arXiv:2604.16583v1 Announce Type: new Abstract: Edge deployment of large language models (LLMs) increasingly relies on libraries of lightweight LoRA adapters, yet GPU/DRAM can keep only a small reside

Probabilistic Programs of Thought

HardwareDGX agent

arXiv:2604.17290v1 Announce Type: new Abstract: LLMs are widely used for code generation and mathematical reasoning tasks where they are required to generate structured output. They either need to rea

SLO-Guard: Crash-Aware, Budget-Consistent Autotuning for SLO-Constrained LLM Serving

HardwareDGX agent

arXiv:2604.17627v1 Announce Type: new Abstract: Serving large language models under latency service-level objectives (SLOs) is a configuration-heavy systems problem with an unusually failure-prone sea

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’…

HardwareDGX agent

SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge work AI. The combination of Cursor’s leading product and distribution to expert software engine

20 Apr 2026

Apple says Tim Cook, as executive chairman, 'will assist with certain aspects of the company, including engaging with policymakers around the world' (Marcus Mendes/9to5Mac)

IndustryDGX agent

Marcus Mendes / 9to5Mac: Apple says Tim Cook, as executive chairman, “will assist with certain aspects of the company, including engaging with policymakers around the world” — Apple confirmed today th

AutoFlows++: Hierarchical Message Flow Mining for System on Chip Designs

HardwareDGX agent

arXiv:2604.15359v1 Announce Type: cross Abstract: Understanding communication behavior in modern system-on-chip (SoC) designs is critical for functional verification, performance analysis, and post-si

Autonomous AI at Scale: Adobe Agents Unlock Breakthrough Creative Intelligence With NVIDIA and WPP

HardwareDGX agent

AI agents are transforming how work gets done across all industries, accelerating everything from content creation to decision-making. NVIDIA’s expanded strategic collaborations with Adobe and WPP are

Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~$2.6B in its IPO, the largest in the city so far this year (Eunice Xu/South China Morning Post)

HardwareDGX agent

Eunice Xu / South China Morning Post: Chinese PCB maker Victory Giant surges ~57% in its Hong Kong trading debut after raising ~2.6B in its IPO, the largest in the city so far this year — Nvidia suppl

Lossless Compression via Chained Lightweight Neural Predictors with Information Inheritance

HardwareDGX agent

arXiv:2604.15472v1 Announce Type: cross Abstract: This paper is dedicated to lossless data compression with probability estimation using neural networks. First, we propose a probability estimation arc

Mitigating Indirect AGENTS.md Injection Attacks in Agentic Environments

HardwareDGX agent

NVIDIA researchers discovered a vulnerability in AI coding assistants where malicious software dependencies can inject harmful instructions into AGENTS.md configuration files, allowing attackers to re

NVIDIA and Partners Showcase the Future of AI-Driven Manufacturing at Hannover Messe 2026

HardwareDGX agent

Manufacturing is at an inflection point. Across every major industrial economy, the pressure to do more with less — due to faster design cycles, leaner operations and strain on skilled labor pools — i

PyLO: Towards Accessible Learned Optimizers in PyTorch

HardwareDGX agent

arXiv:2506.10315v3 Announce Type: replace Abstract: Learned optimizers have been an active research topic over the past decade, with increasing progress toward practical, general-purpose optimizers th

SK hynix says it has begun mass production of the 192GB SOCAMM2, a next-gen LPDDR5X low-power DRAM module designed particularly for Nvidia's Vera Rubin (The Korea Herald)

HardwareDGX agent

The Korea Herald: SK hynix says it has begun mass production of the 192GB SOCAMM2, a next-gen LPDDR5X low-power DRAM module designed particularly for Nvidia's Vera Rubin — SK hynix Inc. said Monday it

The threat of analytic flexibility in using large language models to simulate human data

HardwareDGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, …

HardwareDGX agent

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, multi-modal hierarchical caching, and prefill-decode disaggr

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed…

HardwareDGX agent

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed with building bigger LLMs. Trillions of parameters. Billion

19 Apr 2026

A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research (Abram Brown/The Information)

HardwareDGX agent

Abram Brown / The Information: A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research — For Dylan

A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M in 2026 (Emily Shugerman/The San Francisco ...)

HardwareDGX agent

Emily Shugerman / The San Francisco Standard: A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M i

Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models (Qianer Liu/The Information)

HardwareDGX agent

Qianer Liu / The Information: Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models — Google is in talk

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge back…

HardwareDGX agent

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge backlash against AI, he cobbles together this cope - to try to p

18 Apr 2026

Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection (Reuters)

HardwareDGX agent

Reuters: Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection — Hesai (2525.HK), China's leadi

17 Apr 2026

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on…

HardwareDGX agent

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on @latentspacepod with Alessio Fanelli (@FanaHOVA) and @swyx

A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the interview (Zvi Mowshowitz/Don't Worry About the Vase)

HardwareDGX agent

Zvi Mowshowitz / Don't Worry About the Vase: A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the inter

Accelerate Clean, Modular, Nuclear Reactor Design with AI Physics

HardwareDGX agent

NVIDIA released PhysicsNeMo, an open-source AI framework designed to speed up nuclear reactor design through physics-based simulations. The framework uses machine learning-based multiphysics emulators

AdaSplash-2: Faster Differentiable Sparse Attention

HardwareDGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle

HardwareDGX agent

Quantum computing keeps gaining momentum even though practically it’s years away from wide commercialization, and World Quantum Day April 14 provided an excuse for a lot of announcements. Among them:

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

HardwareDGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

cuRoboV2: Dynamics-Aware Motion Generation with Depth-Fused Distance Fields for High-DoF Robots

HardwareDGX agent

arXiv:2603.05493v2 Announce Type: replace Abstract: Effective robot autonomy requires motion generation that is safe, feasible, and reactive. Current methods are fragmented: fast planners output physi

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

HardwareDGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

DEEP-GAP: Deep-learning Evaluation of Execution Parallelism in GPU Architectural Performance

HardwareDGX agent

arXiv:2604.14552v1 Announce Type: cross Abstract: Modern datacenters increasingly rely on low-power, single-slot inference accelerators to balance performance, energy efficiency, and rack density cons

Dell and Nvidia turn AI infrastructure into the new power center of enterprise tech

HardwareDGX agent

As artificial intelligence scales, AI infrastructure is becoming the deciding factor in whether enterprise AI delivers real value or stalls. AI is shifting from experimentation to unified systems wher

Ernie Image Turbo is not bad at all (Using INT8 quant and Gemini for prompt enhancement, RTX 30 series GPU with low vram)

Model ReleasesDGX agent

Ernie Image Turbo is a text-to-image generation model that can run efficiently on consumer-grade hardware like RTX 30 series GPUs with limited VRAM by using INT8 quantization. The post discusses techn

Friends, housemates, officemates. Maybe lovers?

HardwareDGX agent

This post likely explores the various relationship dynamics and social connections people develop in shared living and working spaces, with a speculative or humorous tone suggested by the 'Maybe lover

Hangzhou-based Manycore shares rose 187% early in its Hong Kong debut after raising $156M in its IPO; it is pivoting to selling AI training data to robot makers (Bloomberg)

HardwareDGX agent

Bloomberg: Hangzhou-based Manycore shares rose 187% early in its Hong Kong debut after raising $156M in its IPO; it is pivoting to selling AI training data to robot makers — Manycore Tech Inc. built a

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding

HardwareDGX agent

arXiv:2601.14724v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated significant improvement in offline video understanding. Howe

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

HardwareDGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

Nautilus: An Auto-Scheduling Tensor Compiler for Efficient Tiled GPU Kernels

HardwareDGX agent

arXiv:2604.14825v1 Announce Type: cross Abstract: We present Nautilus, a novel tensor compiler that moves toward fully automated math-to-kernel optimization. Nautilus compiles a high-level algebraic s

Oh look! Anthropic's entire 'we are delaying Mythos' narrative was marketing hogwash. Kudos to FT for confirming what was obvious. Anthropic…

HardwareDGX agent

Oh look! Anthropic's entire 'we are delaying Mythos' narrative was marketing hogwash. Kudos to FT for confirming what was obvious. Anthropic simply doesn't have the compute. FT: 'Multiple people with

Recent advances push Big Tech closer to the Q-Day danger zone

IndustryDGX agent

Q-Day—the moment when quantum computers can break widely used cryptography—may be approaching faster than expected due to recent algorithmic and hardware advances. Advances in algorithms and design ar

Sources: Cursor is in advanced talks to raise about 2B co-led by a16z at a pre-money valuation of more than 50B, with Nvidia participating (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Cursor is in advanced talks to raise about 2B co-led by a16z at a pre-money valuation of more than 50B, with Nvidia participating — Cursor, a leading artificial intelligence startu

Sources: Recursive Superintelligence, a four-month-old start-up developing self-teaching AI and founded by ex-DeepMind and OpenAI engineers, has raised $500M+ (Financial Times)

HardwareDGX agent

Financial Times: Sources: Recursive Superintelligence, a four-month-old start-up developing self-teaching AI and founded by ex-DeepMind and OpenAI engineers, has raised 500M+ — Group founded by former

Thermodynamic Diffusion Inference with Minimal Digital Conditioning

HardwareDGX agent

arXiv:2604.14332v1 Announce Type: new Abstract: Diffusion-model inference and overdamped Langevin dynamics are formally identical. A physical substrate that encodes the score function therefore equili

Unweight: how we compressed an LLM 22% without sacrificing quality

HardwareDGX agent

Running LLMs across Cloudflare’s network requires us to be smarter and more efficient about GPU memory bandwidth. That’s why we developed Unweight, a lossless inference-time compression system that ac

16 Apr 2026

A Lightweight, Transferable, and Self-Adaptive Framework for Intelligent DC Arc-Fault Detection in Photovoltaic Systems

Local AiDGX agent

arXiv:2603.25749v2 Announce Type: replace-cross Abstract: Arc-fault circuit interrupters (AFCIs) are essential for mitigating fire hazards in residential photovoltaic (PV) systems, yet achieving relia

Completely agree with Jensen! Not exporting AI out of fear is classic case where the cure would be 100x worse than the disease. You slow dow…

HardwareDGX agent

Completely agree with Jensen! Not exporting AI out of fear is classic case where the cure would be 100x worse than the disease. You slow down innovation, progress and US technology and economic leader

Cross-Layer Co-Optimized LSTM Accelerator for Real-Time Gait Analysis

ApplicationsDGX agent

arXiv:2604.13543v1 Announce Type: cross Abstract: Long Short-Term Memory (LSTM) neural networks have penetrated healthcare applications where real-time requirements and edge computing capabilities are

Event Tensor: A Unified Abstraction for Compiling Dynamic Megakernel

HardwareDGX agent

arXiv:2604.13327v1 Announce Type: cross Abstract: Modern GPU workloads, especially large language model (LLM) inference, suffer from kernel launch overheads and coarse synchronization that limit inter

Fluids You Can Trust: Property-Preserving Operator Learning for Incompressible Flows

HardwareDGX agent

arXiv:2602.15472v4 Announce Type: replace-cross Abstract: We present a novel property-preserving kernel-based operator learning method for incompressible flows governed by the incompressible Navier--S

Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model

HardwareDGX agent

arXiv:2603.28554v2 Announce Type: replace Abstract: Visual document understanding typically requires separate retrieval and generation models, doubling memory and system complexity. We present Hydra,

Intel is so fucking back Long road ahead but @LipBuTan1 is the man for the job He is making many long term principled decisions Despite my l…

HardwareDGX agent

Intel is so fucking back Long road ahead but @LipBuTan1 is the man for the job He is making many long term principled decisions Despite my long term optimism, I believe numbers won't be great in the s

Intel refreshes non-Ultra Core CPUs with new silicon for the first time

HardwareDGX agent

Intel has revealed its new Core Series 3 mobile processor (codenamed Wildcat Lake), all fabricated on Intel 18A . The Core Series 3 is a value-focused counterpart to Panther Lake, which was revealed a

MaMe & MaRe: Matrix-Based Token Merging and Restoration for Efficient Visual Perception and Synthesis

HardwareDGX agent

arXiv:2604.13432v1 Announce Type: new Abstract: Token compression is crucial for mitigating the quadratic complexity of self-attention mechanisms in Vision Transformers (ViTs), which often involve num

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

HardwareDGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

← Previous
1…3233343536…75
Next →