AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
20 Apr 2026

The threat of analytic flexibility in using large language models to simulate human data

HardwareDGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

HardwareDGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, …

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

We are excited to have @baseten as a day 0 launch partner for Kimi K2.6! Their inference stack brings KV-aware routing, NVFP4 on Blackwell, multi-modal hierarchical caching, and prefill-decode disaggr

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was th…

HardwareDGX agent

We’re launching Kimi K2.6 on Fireworks as a Day-0 launch partner! K2.5 was the base for standout models like @Cursor’s Composer 2 and was the most popular model on our training platform. K2.6 on Firew

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed…

HardwareDGX agent

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed with building bigger LLMs. Trillions of parameters. Billion

19 Apr 2026

A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research (Abram Brown/The Information)

HardwareDGX agent

Abram Brown / The Information: A look at Dylan Patel's SemiAnalysis, an AI newsletter and research firm that expects $100M+ in 2026 revenue from subscriptions and AI supply chain research — For Dylan

A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M in 2026 (Emily Shugerman/The San Francisco ...)

HardwareDGX agent

Emily Shugerman / The San Francisco Standard: A profile of Maria Davidson, who heads California Renewal, a pro-business political group backed by Silicon Valley power players, seeking to raise $100M i

Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models (Qianer Liu/The Information)

HardwareDGX agent

Qianer Liu / The Information: Sources: Google is in talks with Marvell Technology to develop a memory processing unit that works alongside TPUs, and a new TPU for running AI models — Google is in talk

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge back…

HardwareDGX agent

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge backlash against AI, he cobbles together this cope - to try to p

18 Apr 2026

Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection (Reuters)

HardwareDGX agent

Reuters: Chinese lidar maker Hesai, a primary supplier for Nvidia's ADAS, announces EXT lidar, calling it the industry's first to integrate spatial and color detection — Hesai (2525.HK), China's leadi

17 Apr 2026

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on…

HardwareDGX agent

1/ Old Nvidia chips are getting MORE valuable with age, not less. That has never happened in computing history. Marc Andreessen (@pmarca) on @latentspacepod with Alessio Fanelli (@FanaHOVA) and @swyx

A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the interview (Zvi Mowshowitz/Don't Worry About the Vase)

HardwareDGX agent

Zvi Mowshowitz / Don't Worry About the Vase: A deep dive into Dwarkesh Patel's interview with Jensen Huang, including Huang's takes on Nvidia's moat and chip sales to China, and reactions to the inter

Accelerate Clean, Modular, Nuclear Reactor Design with AI Physics

HardwareDGX agent

NVIDIA released PhysicsNeMo, an open-source AI framework designed to speed up nuclear reactor design through physics-based simulations. The framework uses machine learning-based multiphysics emulators

AdaSplash-2: Faster Differentiable Sparse Attention

HardwareDGX agent

arXiv:2604.15180v1 Announce Type: cross Abstract: Sparse attention has been proposed as a way to alleviate the quadratic cost of transformers, a central bottleneck in long-context training. A promisin

AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle

HardwareDGX agent

Quantum computing keeps gaining momentum even though practically it’s years away from wide commercialization, and World Quantum Day April 14 provided an excuse for a lot of announcements. Among them:

At SemiAnalysis, we're quite tired of the minimalist style webapps and landing pages that have become so commonplace lately. Today, we're in…

HardwareDGX agent

At SemiAnalysis, we're quite tired of the minimalist style webapps and landing pages that have become so commonplace lately. Today, we're introducing Minecraft mode on inferencex dot com so that you c

Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy

HardwareDGX agent

arXiv:2511.10909v2 Announce Type: replace-cross Abstract: Modern AI accelerators rely on matrix multiply-accumulate units (MMAUs), such as NVIDIA Tensor Cores and AMD Matrix Cores, to accelerate deep

cuRoboV2: Dynamics-Aware Motion Generation with Depth-Fused Distance Fields for High-DoF Robots

HardwareDGX agent

arXiv:2603.05493v2 Announce Type: replace Abstract: Effective robot autonomy requires motion generation that is safe, feasible, and reactive. Current methods are fragmented: fast planners output physi

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

HardwareDGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

DEEP-GAP: Deep-learning Evaluation of Execution Parallelism in GPU Architectural Performance

HardwareDGX agent

arXiv:2604.14552v1 Announce Type: cross Abstract: Modern datacenters increasingly rely on low-power, single-slot inference accelerators to balance performance, energy efficiency, and rack density cons

Dell and Nvidia turn AI infrastructure into the new power center of enterprise tech

HardwareDGX agent

As artificial intelligence scales, AI infrastructure is becoming the deciding factor in whether enterprise AI delivers real value or stalls. AI is shifting from experimentation to unified systems wher

Friends, housemates, officemates. Maybe lovers?

HardwareDGX agent

This post likely explores the various relationship dynamics and social connections people develop in shared living and working spaces, with a speculative or humorous tone suggested by the 'Maybe lover

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability

HardwareDGX agent

arXiv:2604.13048v1 Announce Type: cross Abstract: Modern cloud-native platforms expose thousands of time series metrics through systems like Prometheus, yet formulating correct queries in domain-speci

Hangzhou-based Manycore shares rose 187% early in its Hong Kong debut after raising $156M in its IPO; it is pivoting to selling AI training data to robot makers (Bloomberg)

HardwareDGX agent

Bloomberg: Hangzhou-based Manycore shares rose 187% early in its Hong Kong debut after raising $156M in its IPO; it is pivoting to selling AI training data to robot makers — Manycore Tech Inc. built a

HERMES: KV Cache as Hierarchical Memory for Efficient Streaming Video Understanding

HardwareDGX agent

arXiv:2601.14724v3 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated significant improvement in offline video understanding. Howe

LayerScope: Predictive Cross-Layer Scheduling for Efficient Multi-Batch MoE Inference on Legacy Servers

HardwareDGX agent

arXiv:2509.23638v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models face memory and PCIe latency bottlenecks when deployed on commodity hardware. Offloading expert weights to CPU memor

Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3

HardwareDGX agent

arXiv:2603.27844v2 Announce Type: replace Abstract: Majority voting over multiple LLM attempts improves mathematical reasoning, but correlated errors limit the effective sample size. A natural fix is

Nautilus: An Auto-Scheduling Tensor Compiler for Efficient Tiled GPU Kernels

HardwareDGX agent

arXiv:2604.14825v1 Announce Type: cross Abstract: We present Nautilus, a novel tensor compiler that moves toward fully automated math-to-kernel optimization. Nautilus compiles a high-level algebraic s

Oh look! Anthropic's entire 'we are delaying Mythos' narrative was marketing hogwash. Kudos to FT for confirming what was obvious. Anthropic…

HardwareDGX agent

Oh look! Anthropic's entire 'we are delaying Mythos' narrative was marketing hogwash. Kudos to FT for confirming what was obvious. Anthropic simply doesn't have the compute. FT: 'Multiple people with

Reference-Free Sampling-Based Model Predictive Control

HardwareDGX agent

arXiv:2511.19204v3 Announce Type: replace Abstract: We present a sampling-based model predictive control (MPC) framework that enables emergent locomotion without relying on handcrafted gait patterns o

Sources: Cursor is in advanced talks to raise about 2B co-led by a16z at a pre-money valuation of more than 50B, with Nvidia participating (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Cursor is in advanced talks to raise about 2B co-led by a16z at a pre-money valuation of more than 50B, with Nvidia participating — Cursor, a leading artificial intelligence startu

Sources: Recursive Superintelligence, a four-month-old start-up developing self-teaching AI and founded by ex-DeepMind and OpenAI engineers, has raised $500M+ (Financial Times)

HardwareDGX agent

Financial Times: Sources: Recursive Superintelligence, a four-month-old start-up developing self-teaching AI and founded by ex-DeepMind and OpenAI engineers, has raised 500M+ — Group founded by former

Thermodynamic Diffusion Inference with Minimal Digital Conditioning

HardwareDGX agent

arXiv:2604.14332v1 Announce Type: new Abstract: Diffusion-model inference and overdamped Langevin dynamics are formally identical. A physical substrate that encodes the score function therefore equili

Unweight: how we compressed an LLM 22% without sacrificing quality

HardwareDGX agent

Running LLMs across Cloudflare’s network requires us to be smarter and more efficient about GPU memory bandwidth. That’s why we developed Unweight, a lossless inference-time compression system that ac

16 Apr 2026

Automatic Charge State Tuning of 300 mm FDSOI Quantum Dots Using Neural Network Segmentation of Charge Stability Diagram

HardwareDGX agent

arXiv:2604.13662v1 Announce Type: cross Abstract: Tuning of gate-defined semiconductor quantum dots (QDs) is a major bottleneck for scaling spin qubit technologies. We present a deep learning (DL) dri

Completely agree with Jensen! Not exporting AI out of fear is classic case where the cure would be 100x worse than the disease. You slow dow…

HardwareDGX agent

Completely agree with Jensen! Not exporting AI out of fear is classic case where the cure would be 100x worse than the disease. You slow down innovation, progress and US technology and economic leader

Event Tensor: A Unified Abstraction for Compiling Dynamic Megakernel

HardwareDGX agent

arXiv:2604.13327v1 Announce Type: cross Abstract: Modern GPU workloads, especially large language model (LLM) inference, suffer from kernel launch overheads and coarse synchronization that limit inter

Fluids You Can Trust: Property-Preserving Operator Learning for Incompressible Flows

HardwareDGX agent

arXiv:2602.15472v4 Announce Type: replace-cross Abstract: We present a novel property-preserving kernel-based operator learning method for incompressible flows governed by the incompressible Navier--S

Hydra: Unifying Document Retrieval and Generation in a Single Vision-Language Model

HardwareDGX agent

arXiv:2603.28554v2 Announce Type: replace Abstract: Visual document understanding typically requires separate retrieval and generation models, doubling memory and system complexity. We present Hydra,

Intel is so fucking back Long road ahead but @LipBuTan1 is the man for the job He is making many long term principled decisions Despite my l…

HardwareDGX agent

Intel is so fucking back Long road ahead but @LipBuTan1 is the man for the job He is making many long term principled decisions Despite my long term optimism, I believe numbers won't be great in the s

Intel refreshes non-Ultra Core CPUs with new silicon for the first time

HardwareDGX agent

Intel has revealed its new Core Series 3 mobile processor (codenamed Wildcat Lake), all fabricated on Intel 18A . The Core Series 3 is a value-focused counterpart to Panther Lake, which was revealed a

MaMe & MaRe: Matrix-Based Token Merging and Restoration for Efficient Visual Perception and Synthesis

HardwareDGX agent

arXiv:2604.13432v1 Announce Type: new Abstract: Token compression is crucial for mitigating the quadratic complexity of self-attention mechanisms in Vision Transformers (ViTs), which often involve num

Scalable Spatiotemporal Inference with Biased Scan Attention Transformer Neural Processes

HardwareDGX agent

arXiv:2506.09163v2 Announce Type: replace Abstract: Neural Processes (NPs) are a rapidly evolving class of models designed to directly model the posterior predictive distribution of stochastic process

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

HardwareDGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

Sources: OpenAI agrees to pay Cerebras $20B+ to use its server chips, double the amount previously associated with the deal, and may receive equity in Cerebras (The Information)

HardwareDGX agent

The Information: Sources: OpenAI agrees to pay Cerebras $20B+ to use its server chips, double the amount previously associated with the deal, and may receive equity in Cerebras — OpenAI recently struc

The Multi-GPU Era: Why Heterogeneous Compute Is Becoming Enterprise Standard

HardwareDGX agent

Heterogeneous computing, which combines multiple types of GPUs and processors, is becoming standard in enterprise environments to handle diverse workloads more efficiently than traditional homogeneous

to be clear, NVIDIA is NOT a car

HardwareDGX agent

Dylan Patel of SemiAnalysis posted a clarification on X distinguishing NVIDIA from automotive companies, likely in response to comparisons or confusion surrounding NVIDIA's growing presence in the aut

15 Apr 2026

AMD Instinct™ GPU Preemptible Instances at Vultr

HardwareDGX agent

Vultr offers preemptible cloud instances powered by AMD Instinct GPUs, providing a cost-effective option for running GPU-accelerated workloads such as AI/ML training, inference, and high-performance c

Bandwidth Overage Alerts

HardwareDGX agent

Stay informed with new bandwidth usage alerts on Vultr.com as your instances near their transfer limits. For details, visit your Vultr Console's settings section. Give us your feedback on Twitter @Vul

Best part are retrieval and search details: >Agent queries differ from human queries & what good results means changes too >Parallel queries…

HardwareDGX agent

Best part are retrieval and search details: >Agent queries differ from human queries & what good results means changes too >Parallel queries and ranking are both tools to the same outcome >Top-K preci

Comparing hardware on InferenceX just got a massive upgrade. Our new changelog update lets you compare every change for any gpu all on the s…

HardwareDGX agent

Comparing hardware on InferenceX just got a massive upgrade. Our new changelog update lets you compare every change for any gpu all on the same chart. Just select any model and GPU config, then select

Data Security Compliance Made Simple with Vultr

HardwareDGX agent

Discover how Vultr prioritizes data security and compliance, ensuring your data is protected across 32 cloud data center regions worldwide. Explore comprehensive compliance reports and robust certific

DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos

HardwareDGX agent

arXiv:2604.12270v1 Announce Type: new Abstract: Stereo video inpainting, which aims to fill the occluded regions of warped videos with visually coherent content while maintaining temporal consistency,

Fast AI Model Partition for Split Learning over Edge Networks

HardwareDGX agent

arXiv:2507.01041v4 Announce Type: replace-cross Abstract: Split learning (SL) is a distributed learning paradigm that can enable computation-intensive artificial intelligence (AI) applications by part

finally: @simonlast + @sarahmsachs on Latent Space! Notion has rebuilt Notion AI 5 times. This is the first time Simon has told the entire s…

HardwareDGX agent

finally: @simonlast + @sarahmsachs on Latent Space! Notion has rebuilt Notion AI 5 times. This is the first time Simon has told the entire story. I've been trying to do this interview for ~3 years. We

Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data

HardwareDGX agent

arXiv:2503.10676v2 Announce Type: replace-cross Abstract: We study the efficacy of fine-tuning Large Language Models (LLMs) for the specific task of report (government archives, news, intelligence rep

Get Organized with Cloud Server Tagging

HardwareDGX agent

Introducing instance tagging for better organization and easier server management on Vultr. Try tagging now and stay tuned for upcoming features like scheduled backups! Share your feedback on Twitter

GPU stays sometimes at 100% usage even when done replying. Is it normal?

HardwareDGX agent

This r/ollama post addresses a commonly reported behavior where Ollama's GPU usage remains at or near 100% even after a model has finished generating a response. Certain models appear to 'hang' after

I chatted with @ysmulki about MatX, chip design and where silicon designed for LLMs is headed (8:17) Tightly coupling SRAM and HBM on one ch…

HardwareDGX agent

I chatted with @ysmulki about MatX, chip design and where silicon designed for LLMs is headed (8:17) Tightly coupling SRAM and HBM on one chip (14:03) More MoE FLOPS, smaller KV cache load (16:08) Num

Important Update

HardwareDGX agent

Vultr informs customers about a security incident involving unauthorized access to a third-party vendor's marketing data affecting limited customer information. No Vultr systems, customer subscription

← Previous
1…2526272829
Next →