AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
Hardware

Rethinking Network Topologies for Cost-Effective Mixture-of-Experts LLM Serving

DGX agent

arXiv:2605.00254v1 Announce Type: cross Abstract: Mixture-of-experts (MoE) architectures have turned LLM serving into a cluster-scale workload in which communication consumes a considerable portion of

hardwarearxiv-cs-ai
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

Silicon Valley bets $200M on AI data centers floating in the ocean

DGX agent

Panthalassa, a US startup, is developing floating data centers powered by ocean waves and cooled by seawater to address AI infrastructure's growing energy demands. Peter Thiel led a $140 million inves

hardwarears-technica
5 May 2026
Hardware

SplitZip: Ultra Fast Lossless KV Compression for Disaggregated LLM Serving

DGX agent

arXiv:2605.01708v1 Announce Type: cross Abstract: Contemporary systems serving large language models (LLMs) have adopted prefill-decode disaggregation to better load-balance between the compute-bound

hardwarearxiv-cs-lg
5 May 2026
Hardware

Stochastic Sparse Attention for Memory-Bound Inference

DGX agent

arXiv:2605.01910v1 Announce Type: new Abstract: Autoregressive decoding becomes bandwidth-limited at long contexts, as generating each token requires reading all n_k key and value vectors from KV cach

hardwarearxiv-cs-lg
5 May 2026
Hardware

The extit{Silicon Society} Cookbook: Design Space of LLM-based Social Simulations

DGX agent

arXiv:2605.00197v1 Announce Type: cross Abstract: Studies attempting to simulate human behavior with extit{Silicon Societies} grow in numbers while LLM-only social networks have started appearing outs

hardwarearxiv-cs-ai
5 May 2026
Hardware

The Quantization Trap: Breaking Linear Scaling Laws in Multi-Hop Reasoning

DGX agent

arXiv:2602.13595v2 Announce Type: replace Abstract: Neural scaling laws provide a predictable recipe for AI advancement: reducing numerical precision should linearly improve computational efficiency a

hardwarearxiv-cs-ai
5 May 2026
Hardware

ViM-Q: Scalable Algorithm-Hardware Co-Design for Vision Mamba Model Inference on FPGA

DGX agent

arXiv:2605.01935v1 Announce Type: cross Abstract: Vision Mamba (ViM) models offer a compelling efficiency advantage over Transformers by leveraging the linear complexity of State Space Models (SSMs),

hardwarearxiv-cs-cv
5 May 2026
Hardware

We cut model cold-start times from minutes to seconds. 60x faster, saving thousands of GPU-minutes and hundreds of terabytes of transfers ev…

DGX agent

We cut model cold-start times from minutes to seconds. 60x faster, saving thousands of GPU-minutes and hundreds of terabytes of transfers every day. It turns out GPUs already holding weights are faste

hardwarecristobal-valenzuela--x
5 May 2026
Hardware

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization

DGX agent

arXiv:2605.02262v1 Announce Type: cross Abstract: Recently, video language models (VLMs) have been applied in various fields. However, the visual token sequence of the VLM is too long, which may cause

hardwarearxiv-cs-cl
5 May 2026
Hardware

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments

DGX agent

arXiv:2605.00650v1 Announce Type: new Abstract: Fine-tuning LLMs is necessary for various dedicated downstream tasks, but classic backpropagation-based fine-tuning methods require substantial GPU memo

hardwarearxiv-cs-lg
4 May 2026
Hardware

Boosting multimodal inference performance by >10% with a single Python dictionary

DGX agent

This Modal blog post describes a performance optimization technique for multimodal AI inference that achieves over 10% improvement through a simple Python dictionary-based approach. The article likely

hardwaremodal-blog
4 May 2026
Hardware

Cerebras seeks a valuation of up to 26.62B in its US IPO, aiming to raise 3.5B by selling 28M shares at 115 to 125 apiece in its second attempt to go public (Reuters)

DGX agent

Reuters: Cerebras seeks a valuation of up to 26.62B in its US IPO, aiming to raise 3.5B by selling 28M shares at 115 to 125 apiece in its second attempt to go public — Nvidia-rival Cerebras is seeking

hardwaretechmeme
4 May 2026
Hardware

Down payment on a EUV machine!

DGX agent

Dylan Patel likely announced or commented on a significant financial commitment or investment milestone related to an EUV (Extreme Ultraviolet) lithography machine, which are critical semiconductor ma

hardwaredylan-patel--x
4 May 2026
Hardware

DPU or GPU for Accelerating Neural Networks Inference -- Why not both? Split CNN Inference

DGX agent

arXiv:2605.00174v1 Announce Type: cross Abstract: Video and image streaming on edge devices requires low latency. To address this, Neural Networks (NNs) are widely used, and prior work mainly focuses

hardwarearxiv-cs-cv
4 May 2026
Hardware

Lattice Semiconductor agrees to acquire AMI for $1.65B in cash and stock; Georgia-based AMI provides firmware and infrastructure manageability for cloud and AI (Mike Rogoway/Oregonian)

DGX agent

Mike Rogoway / Oregonian: Lattice Semiconductor agrees to acquire AMI for 1.65B in cash and stock; Georgia-based AMI provides firmware and infrastructure manageability for cloud and AI — Hillsboro-bas

hardwaretechmeme
4 May 2026
Hardware

Legislators and experts criticize the EU's €20B sovereign compute data center plan, questioning whether there is demand and the plan's reliance on Nvidia GPUs (Pieter Haeck/Politico)

DGX agent

Pieter Haeck / Politico: Legislators and experts criticize the EU's €20B sovereign compute data center plan, questioning whether there is demand and the plan's reliance on Nvidia GPUs — Critics argue

hardwaretechmeme
4 May 2026
Hardware

llamacpp on Apple Silicon, once configured correctly is really rock solid! You can throw anything at it and it will answer. Impressive!

DGX agent

Llamacpp, when properly configured on Apple Silicon hardware, demonstrates robust performance and reliability for running language models. The tool can handle varied input requests effectively, making

hardwareclem-delangue--x
4 May 2026
Hardware

MiniVLA-Nav v1: A Multi-Scene Simulation Dataset for Language-Conditioned Robot Navigation

DGX agent

arXiv:2605.00397v1 Announce Type: new Abstract: We present MiniVLA-Nav v1, a simulation dataset for Language-Conditioned Object Approach (LCOA) navigation: given a short natural-language instruction,

hardwarearxiv-cs-ro
4 May 2026
Hardware

Most AI teams treat compute as a commodity. It's not.

DGX agent

AI compute resources should not be treated as interchangeable commodities, as different hardware configurations, providers, and architectures significantly impact training costs, inference latency, an

hardwarelambda-labs
4 May 2026
Hardware

Optimize Supply Chain Decision Systems Using NVIDIA cuOpt Agent Skills

DGX agent

NVIDIA cuOpt Agent Skills integrate LLM reasoning with GPU-accelerated solvers to enable AI agents to translate natural language supply chain problems into optimized mathematical models, with skills e

hardwarenvidia-developer
4 May 2026
Hardware

The case against OpenAI is getting markedly stronger now that Musk is off the stand. Why? Musk’s lawyer is interrogating OpenAI founder Greg…

DGX agent

The case against OpenAI is getting markedly stronger now that Musk is off the stand. Why? Musk’s lawyer is interrogating OpenAI founder Greg Brockman, making clear that OpenAI sold its mission as a no

hardwaregary-marcus--x
4 May 2026
Hardware

Analysis: Asian suppliers account for ~90% of Nvidia's production costs, up from 65% in 2025, as latest wave of collaborations shifts from chips to physical AI (Abhishek Vishnoi/Bloomberg)

DGX agent

Abhishek Vishnoi / Bloomberg: Analysis: Asian suppliers account for ~90% of Nvidia's production costs, up from 65% in 2025, as latest wave of collaborations shifts from chips to physical AI — The list

hardwaretechmeme
3 May 2026
Hardware

Are babies and relationships measured on the same time scale? After 18-24 months you start measuring in years.

DGX agent

This post draws a comparison between how parents measure their babies' development and relationship milestones, noting that both shift from month-based to year-based metrics around the 18-24 month mar

hardwaredylan-patel--x
3 May 2026
Hardware

NVIDIA's New AI Turns One Photo Into A World That Never Breaks

DGX agent

NVIDIA's Instant Splat is a breakthrough AI technique that synthesizes missing information from a few images to create immersive 3D models of virtual worlds. The method produces high-quality 3D models

hardwaretwo-minute-papers
3 May 2026
Hardware

People hate on SF men for group think, but every girl in SF has Tabi's now It's an epidemic

DGX agent

This post humorously critiques conformity in San Francisco by observing that while men in the tech scene are often stereotyped as groupthink followers, women in SF have similarly adopted Tabis (Japane

hardwaredylan-patel--x
3 May 2026
Hardware

The number of jobs in the future is endless because the problems to solve are endless. Jobs multiply as we get more complex. No AI or human …

DGX agent

The number of jobs in the future is endless because the problems to solve are endless. Jobs multiply as we get more complex. No AI or human can solve all problems and all the work to do in the Univers

hardwareyann-lecun--x
3 May 2026
Hardware

torch-nvenc-compress: GPU NVENC silicon as a PCIe bandwidth multiplier — PCA + pure-ctypes Video Codec SDK wrapper. Parallel-path overlap measured at 67% of theoretical max on a real GEMM + encode workload. [P]

DGX agent

torch-nvenc-compress is a Python library that leverages GPU NVENC (NVIDIA's hardware video encoding) to optimize PCIe bandwidth utilization by compressing data during transfer. The project implements

hardwarer-machinelearning
3 May 2026
Hardware

Whenever I make a group chat green, I now get to proudly proclaim it as goblin mode

DGX agent

The post humorously references Apple's iMessage feature that displays group chats in green when they include non-iPhone users, joking that activating 'goblin mode' (a playful term for behaving mischie

hardwaredylan-patel--x
3 May 2026
Hardware

CUDA V.13?

DGX agent

Ollama's MLX engine runs on NVIDIA GPUs via CUDA v13 on Windows and Linux. Users have reported that Ollama crashes on RTX 3060 with cuda_v13 in versions 0.13.0 through 0.15.6, while deleting the cuda_

hardwarer-ollama
2 May 2026
Hardware

Intel Inside the Micro Revolution: 8008 Origins

DGX agent

This article explores the origins and development of Intel's 8008 microprocessor, a pioneering chip that played a crucial role in the early microcomputer revolution of the 1970s. It likely examines th

hardwarethe-chip-letter
2 May 2026
Hardware

Sources: Anthropic is in early talks to buy AI inference chips from UK-based Fractile when they become available in 2027 (The Information)

DGX agent

The Information: Sources: Anthropic is in early talks to buy AI inference chips from UK-based Fractile when they become available in 2027 — As Anthropic's sales explode, straining the servers it uses,

hardwaretechmeme
2 May 2026
Hardware

AI Value Capture - The Shift To Model Labs

DGX agent

This article examines how value in the AI industry is shifting away from traditional chip manufacturers toward model development labs and companies that build large language models. It likely analyzes

hardwaresemianalysis
1 May 2026
Hardware

AI Value Capture - The Shift To Model Labs Vera Rubin VR NVL72: V for Value - Rubin delivers a step jump in performance per TCO. ROI accruin…

DGX agent

AI Value Capture - The Shift To Model Labs Vera Rubin VR NVL72: V for Value - Rubin delivers a step jump in performance per TCO. ROI accruing to users, Neoclouds, Hyperscalers, AI Labs, Memory Vendors

hardwaredylan-patel--x
1 May 2026
Hardware

All the Doomers and hawks are lining up behind this distillation 'attack' farce because they want to see open source banned. It's really as …

DGX agent

All the Doomers and hawks are lining up behind this distillation 'attack' farce because they want to see open source banned. It's really as simple as that. They want to take away your right to choose,

hardwareyann-lecun--x
1 May 2026
Hardware

Benchmarking Deep Learning Models for Object Detection on Edge Computing Devices

DGX agent

arXiv:2409.16808v1 Announce Type: cross Abstract: Modern applications, such as autonomous vehicles, require deploying deep learning algorithms on resource-constrained edge devices for real-time image

hardwarearxiv-cs-lg
1 May 2026
Hardware

By popular request, YOLO-mode has landed on ml-intern https://smolagents-ml-intern.hf.space This enables the intern to execute long-running …

DGX agent

By popular request, YOLO-mode has landed on ml-intern https://smolagents-ml-intern.hf.space This enables the intern to execute long-running tasks like launching parallel ablations to determine the opt

hardwareclem-delangue--x
1 May 2026
Hardware

EdgeFM: Efficient Edge Inference for Vision-Language Models

DGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

hardwarearxiv-cs-cv
1 May 2026
Hardware

Efficient Training on Multiple Consumer GPUs with RoundPipe

DGX agent

arXiv:2604.27085v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) on consumer-grade GPUs is highly cost-effective, yet constrained by limited GPU memory and slow PCIe intercon

hardwarearxiv-cs-ai
1 May 2026
Hardware

First Silicon Valley captured the US Government, and now it is capturing “independent” media.

DGX agent

First Silicon Valley captured the US Government, and now it is capturing “independent” media. 🟡 Semafor is launching a new initiative: Silicon Valley & The World, bringing together leaders building ar

hardwaregary-marcus--x
1 May 2026
Hardware

FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems

DGX agent

arXiv:2604.28156v1 Announce Type: cross Abstract: We present FlexiTac, a low-cost, open-source, and scalable piezoresistive tactile sensing solution designed for robotic end-effectors. FlexiTac is a p

hardwarearxiv-cs-ai
1 May 2026
Hardware

GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2

DGX agent

arXiv:2509.16248v3 Announce Type: replace-cross Abstract: This paper presents GRAPHMEND, a high-level compiler technique that eliminates FX graph breaks in PyTorch 2 programs. Although PyTorch 2 intro

hardwarearxiv-cs-lg
1 May 2026
Hardware

Marketing Architects Scales Real-Time CTV Advertising with Vultr Infrastructure

DGX agent

Discover how Marketing Architects scaled real-time CTV advertising with Vultr, doubling throughput, reducing latency by 50%, and cutting cloud costs by up to 50% with high-performance Kubernetes and c

hardwarevultr
1 May 2026
Hardware

MARS: Efficient, Adaptive Co-Scheduling for Heterogeneous Agentic Systems

DGX agent

arXiv:2604.26963v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as the execution core of autonomous agents rather than as standalone text generators. Agentic w

hardwarearxiv-cs-lg
1 May 2026
Hardware

More money for worse work that you have to fix, good stuff this AI thing, thanks Nvidia.

DGX agent

Gary Marcus critiques AI systems (particularly Nvidia's offerings) for producing lower-quality outputs that require significant correction and refinement despite increased computational investment. Th

hardwaregary-marcus--x
1 May 2026
Hardware

Pentagon inks AI procurement deals with seven companies, leaves out Anthropic

DGX agent

The U.S. Defense Department today announced that it has inked artificial intelligence procurement contracts with seven tech firms. The group includes Amazon Web Service Inc., Google LLC, Microsoft Cor

hardwaresiliconangle
1 May 2026
Hardware

Pentagon strikes classified AI deals with OpenAI, Google, and Nvidia — but not Anthropic

DGX agent

The Pentagon has struck deals with OpenAI, Google, Microsoft, Amazon, Nvidia, Elon Musk's xAI, and the startup Reflection, allowing the agency to use their AI tools in classified settings, according t

hardwarethe-verge-ai
1 May 2026
Hardware

Predictive Multi-Tier Memory Management for KV Cache in Large-Scale GPU Inference

DGX agent

arXiv:2604.26968v1 Announce Type: cross Abstract: Key-value (KV) cache memory management is the primary bottleneck limiting throughput and cost-efficiency in large-scale GPU inference serving. Current

hardwarearxiv-cs-ai
1 May 2026
Hardware

Sources: Huawei expects AI chip revenue to hit ~12B in 2026, up 60% from 7.5B in 2025, as orders for its Ascend 950PR chip surge and Nvidia stalls in China (Zijing Wu/Financial Times)

DGX agent

Zijing Wu / Financial Times: Sources: Huawei expects AI chip revenue to hit ~12B in 2026, up 60% from 7.5B in 2025, as orders for its Ascend 950PR chip surge and Nvidia stalls in China — Chinese tech

hardwaretechmeme
1 May 2026
← Previous
1…2728293031…37
Next →