AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
28 Apr 2026

Reward Models Are Secretly Value Functions: Temporally Coherent Reward Modeling

HardwareDGX agent

arXiv:2604.22981v1 Announce Type: new Abstract: Reward models in RLHF are trained to score only the final token of a response - a choice that discards rich signal from every intermediate position and

Scalable Hyperparameter-Divergent Ensemble Training with Automatic Learning Rate Exploration for Large Models

HardwareDGX agent

arXiv:2604.24708v1 Announce Type: cross Abstract: Training large neural networks with data-parallel stochastic gradient descent allocates N GPU replicas to compute effectively identical updates -- a p

Self-Rewarding Vision-Language Model via Reasoning Decomposition

HardwareDGX agent

arXiv:2508.19652v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) often suffer from visual hallucinations: generating things that are not consistent with visual inputs and language sho


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Thank god these drones run on the American tech stack

HardwareDGX agent

Dylan Patel discusses the advantages of drones utilizing American technology infrastructure and components, likely highlighting benefits such as reliability, security, supply chain independence, and t

27 Apr 2026

Beyond Land Surface Temperature: Explainable Spatial Machine Learning Reveals Urban Morphology Effects on Human-Centric Heat Stress

HardwareDGX agent

arXiv:2604.22433v1 Announce Type: new Abstract: Heat exposure connects the built environment and public health, directly shaping the livability and sustainability of urban areas. Understanding the spa

@dylan522p You’re their president

HardwareDGX agent

This appears to be a post by Dylan Patel (@dylan522p) on X (Twitter), though the context of who 'their' refers to and what organization or group he is president of is not clear from the title alone. T

Elon’s influence was monumental to OpenAI. The reality is there simply wouldn’t have been the AI world we have today. It is not perfect. But…

HardwareDGX agent

Elon’s influence was monumental to OpenAI. The reality is there simply wouldn’t have been the AI world we have today. It is not perfect. But would have been orders of magnitude worse. This is the way

Incentivizing Neuro-symbolic Language-based Reasoning in VLMs via Reinforcement Learning

HardwareDGX agent

arXiv:2604.22062v1 Announce Type: new Abstract: There are 7,407 languages in the world. But, what about the languages that are not there in the world? Are humans so narrow minded that we don't care ab

Keras Kinetic has a new alpha release: v0.0.2! Including a new docs website: http://kinetic.readthedocs.io Kinetic is my favorite new releas…

HardwareDGX agent

Keras Kinetic has a new alpha release: v0.0.2! Including a new docs website: http://kinetic.readthedocs.io Kinetic is my favorite new release from the Keras team: a super simple Modal-like API to run

Kernel Contracts: A Specification Language for ML Kernel Correctness Across Heterogeneous Silicon

HardwareDGX agent

arXiv:2604.22032v1 Announce Type: new Abstract: Every ML kernel ships with an implicit contract about what it computes. People rarely write the contract down. When two kernels disagree -- when a matmu

Motivating Next-Gen Accelerators with Flexible (N:M) Activation Sparsity via Benchmarking Lightweight Post-Training Sparsification Approaches

HardwareDGX agent

arXiv:2509.22166v4 Announce Type: replace-cross Abstract: The demand for efficient large language model (LLM) inference has intensified the focus on sparsification techniques. While semi-structured (N

Multimodal Neural Operators for Real-Time Biomechanical Modelling of Traumatic Brain Injury

HardwareDGX agent

arXiv:2510.03248v3 Announce Type: replace-cross Abstract: Background: Traumatic brain injury modeling requires integrating volumetric neuroimaging, demographic parameters, and acquisition metadata. Fi

SIE3D: Single-Image Expressive 3D Avatar Generation via Semantic Embedding and Perceptual Expression Loss

HardwareDGX agent

arXiv:2509.24004v2 Announce Type: replace Abstract: Generating high-fidelity 3D head avatars from a single image is challenging, as current methods lack fine-grained, intuitive control over expression

Soft Anisotropic Diagrams for Differentiable Image Representation

HardwareDGX agent

arXiv:2604.21984v1 Announce Type: new Abstract: We introduce Soft Anisotropic Diagrams (SAD), an explicit and differentiable image representation parameterized by a set of adaptive sites in the image

SpikingBrain2.0: Brain-Inspired Foundation Models for Efficient Long-Context and Cross-Platform Inference

HardwareDGX agent

arXiv:2604.22575v1 Announce Type: new Abstract: Scaling context length is reshaping large-model development, yet full-attention Transformers suffer from prohibitive computation and inference bottlenec

The third leg of AI’s infrastructure race isn’t silicon or power. It’s capital

HardwareDGX agent

The artificial intelligence arms race has spent the last two years obsessed with a duopoly of constraints: the desperate hunt for Nvidia Corp. silicon and the grueling wait for grid-scale megawatts. I

Toward Robust and Efficient ML-Based GPU Caching for Modern Inference

HardwareDGX agent

arXiv:2509.20979v2 Announce Type: replace Abstract: In modern GPU inference, cache efficiency remains a major bottleneck, and heuristic policies such as extsc{LRU} can perform far worse than the offli

26 Apr 2026

3 hr notice I had over 50 people come to the coffee shop incl a guy who brought his GPU for me to sign. Crazy!

HardwareDGX agent

3 hr notice I had over 50 people come to the coffee shop incl a guy who brought his GPU for me to sign. Crazy! Media If anyone wants to meet in New York City today, I'll be at Dark Matter Coffee, NYC,

If anyone wants to meet in New York City today, I'll be at Dark Matter Coffee, NYC, from 11:00AM to 12:30! https://maps.app.goo.gl/jsEibm89k…

HardwareDGX agent

Dylan Patel announced his availability to meet in person at Dark Matter Coffee in New York City on a specific date between 11:00 AM and 12:30 PM. The post included a Google Maps link to the coffee sho

25 Apr 2026

33% of S&P 500 companies provided security perks for execs in 2025: Jensen Huang's security cost rose from 690K in 2023 to 3.5M in 2025, Zuckerberg spent $22M (Eli Rosenberg/The Information)

HardwareDGX agent

Eli Rosenberg / The Information: 33% of S&P 500 companies provided security perks for execs in 2025: Jensen Huang's security cost rose from 690K in 2023 to 3.5M in 2025, Zuckerberg spent $22M — Nvidia

A look at 'Stanford inside Stanford', where VCs pursue 18- and 19-year-old students, offering mentorship and funding in a bid to convert promise into profit (Theo Baker/The Atlantic)

HardwareDGX agent

Theo Baker / The Atlantic: A look at “Stanford inside Stanford”, where VCs pursue 18- and 19-year-old students, offering mentorship and funding in a bid to convert promise into profit — Silicon Valley

confidence interval 98%, binomial distribution

HardwareDGX agent

A 98% confidence interval for a binomial distribution is a range of values that has a 98% probability of containing the true population proportion, calculated using binomial probability methods or nor

DAVIS, APRIL 25, 2026 — InferenceX has added DeepSeekv4 for @vllm_project 's day 0 support for GB200 disagg! Great work to @flowpow123 @roge…

HardwareDGX agent

DAVIS, APRIL 25, 2026 — InferenceX has added DeepSeekv4 for @vllm_project 's day 0 support for GB200 disagg! Great work to @flowpow123 @rogerw0108 @NVIDIAAIDev @inferact for the fast support and engin

Google’s AI agent platform takes pole position but work remains

HardwareDGX agent

Enterprises are rapidly moving from an artificial intelligence that answers questions and generates content to one that performs tasks and takes actions. According to Google Cloud Chief Executive Thom

Intel's server chip volume up sequentially for 3 straight quarters but still less than 52% vs prior peak in 1Q20-1Q26 time period. However, …

HardwareDGX agent

Intel's server chip volume up sequentially for 3 straight quarters but still less than 52% vs prior peak in 1Q20-1Q26 time period. However, ASPs up 14% vs prior peak in the same period. DCAI revenue a

Launching pyptx — a Python DSL for writing NVIDIA PTX kernels. One PTX instruction = one Python call. Write pure PTX in Python. Direct Hoppe…

HardwareDGX agent

Launching pyptx — a Python DSL for writing NVIDIA PTX kernels. One PTX instruction = one Python call. Write pure PTX in Python. Direct Hopper + Blackwell support: wgmma, TMA, tcgen05, mbarriers. JAX +

NVIDIA’s New AI Changed Robotics Forever

HardwareDGX agent

NVIDIA's Isaac GR00T open models enable robots to understand natural language instructions and perform complex, multistep tasks using vision language action reasoning. Cosmos 3 is the first world foun

24 Apr 2026

Association Is Not Similarity: Learning Corpus-Specific Associations for Multi-Hop Retrieval

HardwareDGX agent

arXiv:2604.20850v1 Announce Type: cross Abstract: Dense retrieval systems rank passages by embedding similarity to a query, but multi-hop questions require passages that are associatively related thro

DDN and Google Cloud are redefining AI storage infrastructure for the agentic era

HardwareDGX agent

As enterprises pivot from AI experimentation to production deployment, AI storage infrastructure is emerging as the critical bottleneck determining whether massive chip investments deliver real return

GraphLeap: Decoupling Graph Construction and Convolution for Vision GNN Acceleration on FPGA

HardwareDGX agent

arXiv:2604.21290v1 Announce Type: new Abstract: Vision Graph Neural Networks (ViGs) represent an image as a graph of patch tokens, enabling adaptive, feature-driven neighborhoods. Unlike CNNs with fix

GSpaRC: Gaussian Splatting for Real-time Reconstruction of RF Channels

HardwareDGX agent

arXiv:2511.22793v2 Announce Type: replace Abstract: Channel state information (CSI) is essential for adaptive beamforming and maintaining robust links in wireless communication systems. However, acqui

Nvidia stock jumps 4.3% to close at a record for the first time since Oct., pushing Nvidia's market cap past $5T, as a rally in Intel pushed chipmakers higher (Jordan Novet/CNBC)

HardwareDGX agent

Jordan Novet / CNBC: Nvidia stock jumps 4.3% to close at a record for the first time since Oct., pushing Nvidia's market cap past 5T, as a rally in Intel pushed chipmakers higher — Nvidia shares close

Stream2LLM: Overlap Context Streaming and Prefill for Reduced Time-to-First-Token (TTFT)

HardwareDGX agent

arXiv:2604.16395v2 Announce Type: replace-cross Abstract: Context retrieval systems for LLM inference face a critical challenge: high retrieval latency creates a fundamental tension between waiting fo

🔴 Trump vient de publier sur Truth Social un monologue de 1 500 mots sur le droit du sol, l'ACLU, les Indiens, les Chinois et la 'Savage Na…

HardwareDGX agent

🔴 Trump vient de publier sur Truth Social un monologue de 1 500 mots sur le droit du sol, l'ACLU, les Indiens, les Chinois et la 'Savage Nation'. Ce texte inhabituellement long, inhabituellement cohér

23 Apr 2026

A Delta-Aware Orchestration Framework for Scalable Multi-Agent Edge Computing

HardwareDGX agent

arXiv:2604.20129v1 Announce Type: new Abstract: The Synergistic Collapse occurs when scaling beyond 100 agents causes superlinear performance degradation that individual optimizations cannot prevent.

Algorithm and Hardware Co-Design for Efficient Complex-Valued Uncertainty Estimation

HardwareDGX agent

arXiv:2604.19993v1 Announce Type: cross Abstract: Complex-Valued Neural Networks (CVNNs) have significant advantages in handling tasks that involve complex numbers. However, existing CVNNs are unable

Amortized Vine Copulas for High-Dimensional Density and Information Estimation

HardwareDGX agent

arXiv:2604.20568v1 Announce Type: new Abstract: Modeling high-dimensional dependencies while keeping likelihoods tractable remains challenging. Classical vine-copula pipelines are interpretable but ca

BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching

HardwareDGX agent

arXiv:2412.03594v3 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly play an important role in a wide range of information processing and management tasks in industry. M

Every conversation I have with @dylan522p, I'm really just trying to understand the supply and demand of tokens. This is a unique episode in…

HardwareDGX agent

Every conversation I have with @dylan522p, I'm really just trying to understand the supply and demand of tokens. This is a unique episode in that it's entirely dedicated to talking about both sides of

From GPUs to AI factories: Inside the Nvidia-Google Cloud superstack

HardwareDGX agent

Nvidia Corp. and Google LLC used the search giant’s annual Cloud Next event to deepen their long-running partnership, creating a full-stack “artificial intelligence factory” that integrates Google’s A

Gaussians on a Diet: High-Quality Memory-Bounded 3D Gaussian Splatting Training

HardwareDGX agent

arXiv:2604.20046v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) has revolutionized novel view synthesis with high-quality rendering through continuous aggregations of millions of 3D Gauss

LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel

HardwareDGX agent

arXiv:2604.20368v1 Announce Type: cross Abstract: The quadratic complexity of softmax attention presents a major obstacle for scaling Transformers to high-resolution vision tasks. Existing linear atte

lean in

HardwareDGX agent

'Lean In' likely refers to a post by Dylan Patel from SemiAnalysis discussing lean manufacturing, operational efficiency, or business strategy concepts, though the specific content cannot be verified

Making Sense of the Early Universe

HardwareDGX agent

UC Santa Cruz astronomers using GPU-accelerated simulations have discovered rotating disk galaxies resembling the Milky Way appearing far earlier in the universe than expected, breaking records for th

OpenCLAW-P2P v6.0: Resilient Multi-Layer Persistence, Live Reference Verification, and Production-Scale Evaluation of Decentralized AI Peer Review

HardwareDGX agent

arXiv:2604.19792v1 Announce Type: new Abstract: This paper presents OpenCLAW-P2P v6.0, a comprehensive evolution of the decentralized collective-intelligence platform in which autonomous AI agents pub

people who type with capslock are kinda fakers when i am crusing with caps, I HOLD DOWN SHIFT KEY WITH MY RIGHT PINKIE EXTREMELY HARD! THE E…

HardwareDGX agent

people who type with capslock are kinda fakers when i am crusing with caps, I HOLD DOWN SHIFT KEY WITH MY RIGHT PINKIE EXTREMELY HARD! THE EMOTION IS FLOWING THROUGH TO MY TEXT i often imagine others

Stream-CQSA: Avoiding Out-of-Memory in Attention Computation via Flexible Workload Scheduling

HardwareDGX agent

arXiv:2604.20819v1 Announce Type: new Abstract: The scalability of long-context large language models is fundamentally limited by the quadratic memory cost of exact self-attention, which often leads t

Tag, You’re It: GeForce NOW Levels Up Game Discovery With Xbox Game Pass and Ubisoft+ Labels

HardwareDGX agent

GeForce NOW is doubling down on what matters most: gamers. This week’s upgrades bring smarter libraries, making it easier than ever for gamers to turn a PC collection into a cloud-powered flex. It sta

Temporally Extended Mixture-of-Experts Models

HardwareDGX agent

arXiv:2604.20156v1 Announce Type: new Abstract: Mixture-of-Experts models, now popular for scaling capacity at fixed inference speed, switch experts at nearly every token. Once a model outgrows availa

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind…

HardwareDGX agent

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind paper on test-time compute scaling. Here's what the agent d

US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases (Alexandra Alper/Reuters)

HardwareDGX agent

Alexandra Alper / Reuters: US Commerce Secretary Howard Lutnick says Nvidia has yet to sell H200 chips to Chinese companies and that the Chinese government has not approved such purchases — Nvidia's (

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a gre…

HardwareDGX agent

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a great business strategy 🙈 So now we have 2 free opus sessions/d

We tried a new thing with NVIDIA to roll out Codex across a whole company and it was awesome to see it work. Let us know if you'd like to do…

HardwareDGX agent

Sam Altman posted about OpenAI's successful pilot program with NVIDIA to deploy Codex (an AI code generation model) across an entire company, highlighting positive results from the enterprise rollout.

22 Apr 2026

AdaGScale: Viewpoint-Adaptive Gaussian Scaling in 3D Gaussian Splatting to Reduce Gaussian-Tile Pairs

HardwareDGX agent

arXiv:2604.18980v1 Announce Type: new Abstract: Reducing the number of Gaussian-tile pairs is one of the most promising approaches to improve 3D Gaussian Splatting (3D-GS) rendering speed on GPUs. How

Advancing Emerging Optimizers for Accelerated LLM Training with NVIDIA Megatron

HardwareDGX agent

This blog post discusses higher-order optimization algorithms such as Shampoo that have been applied in neural network training for at least a decade. The article explores how these emerging optimizer

ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants

HardwareDGX agent

arXiv:2604.18616v1 Announce Type: cross Abstract: LLM-based coding agents can generate functionally correct GPU kernels, yet their performance remains far below hand-optimized libraries on critical co

Compute is king. I’m excited about the next phase of growth for Cursor, as their close partner.

HardwareDGX agent

Compute is king. I’m excited about the next phase of growth for Cursor, as their close partner. SpaceXAI and @cursor_ai are now working closely together to create the world’s best coding and knowledge

@dwarkesh_sp @dylan522p No catfishing here, we're pixelmaxxing!

HardwareDGX agent

This post appears to be a humorous take on AI optimization and technology trends, likely referencing 'pixel-maxxing' (maximizing pixel-level performance or visual quality) in contrast to 'catfishing'

Earlier this month, the inaugural Runway AI Summit brought together over 700 leaders across media, entertainment, gaming and advertising. Th…

HardwareDGX agent

Earlier this month, the inaugural Runway AI Summit brought together over 700 leaders across media, entertainment, gaming and advertising. The day was filled with firesides, panels and keynotes from le

Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs

HardwareDGX agent

arXiv:2604.18788v1 Announce Type: new Abstract: Apple Neural Engine (ANE) is a dedicated neural processing unit (NPU) present in every Apple Silicon chip. Mixture-of-Experts (MoE) LLMs improve inferen

← Previous
1…2324252627…29
Next →