AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
10 May 2026

Nvidia, AI factories and the transition to accelerated computing

HardwareDGX agent

The market is trying to price a transition it hasn’t fully internalized. It sees Nvidia Corp.’s market cap with a five-handle and assumes the valuation is too high to grow further. We believe that’s t

‘Your Career Starts at the Beginning of the AI Revolution,’ NVIDIA CEO Tells Graduates

HardwareDGX agent

“You are entering the world at an extraordinary moment,” NVIDIA founder and CEO Jensen Huang told graduates as he delivered the keynote address at Carnegie Mellon University’s 128th commencement cerem

9 May 2026

Yann LeCun closed $1.03B for AMI Labs on March 10. Three days later, this paper dropped from his NYU collaborators. 15M parameters. Single G…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
HardwareDGX agent

Yann LeCun closed 1.03B for AMI Labs on March 10. Three days later, this paper dropped from his NYU collaborators. 15M parameters. Single GPU. A few hours of training. LeWorldModel is the first JEPA t

8 May 2026

AI data center firm IREN’s stock soars after it strikes $2.1B deal with Nvidia

HardwareDGX agent

Neocloud company IREN Ltd. has secured a 2.1 billion commitment from the chipmaker Nvidia Corp. as part of a new data center partnership aimed at artificial intelligence workloads. The partners plan t

Deploy and inference any model from HuggingFace

HardwareDGX agent

Learn how to deploy any Hugging Face model in one session using Goose and Together's Dedicated Container Inference. Skip the setup complexity — one prompt gets your model running in a production-grade

Improving Bash Generation in Small Language Models with Grammar-Constrained Decoding

HardwareDGX agent

Grammar-constrained decoding modifies language model generation by applying grammar constraints at each step to block structurally invalid tokens , ensuring syntactically correct Bash command generati

OPINION: Ever since private equity bought the @Veritasium YouTube channel, the content just hasn't been as good. Used to be a GREAT channel,…

HardwareDGX agent

OPINION: Ever since private equity bought the @Veritasium YouTube channel, the content just hasn't been as good. Used to be a GREAT channel, maybe the best Science channel on YouTube, and now? SAD! Ve

Tickets are now available for the 4th annual AI Film Festival: June 11 at Alice Tully Hall at Lincoln Center in NYC, and June 18 at The Broa…

HardwareDGX agent

Tickets are now available for the 4th annual AI Film Festival: June 11 at Alice Tully Hall at Lincoln Center in NYC, and June 18 at The Broad Stage in LA. It’s the biggest and most important celebrati

7 May 2026

A Physics-Aware Framework for Short-Term GPU Power Forecasting of AI Data Centers

HardwareDGX agent

arXiv:2605.04074v1 Announce Type: new Abstract: AI data centers experience rapid fluctuations in power demand due to the heterogeneity of computational tasks that they have to support. For example, th

A Queueing-Theoretic Framework for Stability Analysis of LLM Inference with KV Cache Memory Constraints

HardwareDGX agent

arXiv:2605.04595v1 Announce Type: new Abstract: The rapid adoption of large language models (LLMs) has created significant challenges for efficient inference at scale. Unlike traditional workloads, LL

b9050

Local AiDGX agent

B9050 is a release build of llama.cpp, an open-source C/C++ project that enables LLM inference with minimal setup and high performance on diverse hardware platforms. The project provides LLM inference

b9061

Local AiDGX agent

Release b9061 is a build version of llama.cpp, a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance across various hardware platforms . As a spec

b9063

Local AiDGX agent

Release b9063 is a build of llama.cpp, a project for LLM inference in C/C++. llama.cpp enables efficient large language model execution on consumer hardware through optimized implementations and quant

Budget-aware Auto Optimizer Configurator

HardwareDGX agent

arXiv:2605.04711v1 Announce Type: cross Abstract: Optimizer states occupy massive GPU memory in large-scale model training. However, gradients in different network blocks exhibit distinct behaviors, s

CuBridge: An LLM-Based Framework for Understanding and Reconstructing High-Performance Attention Kernels

HardwareDGX agent

arXiv:2605.05023v1 Announce Type: new Abstract: Efficient CUDA implementations of attention mechanisms are critical to modern deep learning systems, yet supporting diverse and evolving attention varia

Deep Wave Network for Modeling Multi-Scale Physical Dynamics

HardwareDGX agent

arXiv:2605.04198v1 Announce Type: new Abstract: Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-scien

DualTCN: A Physics-Constrained Temporal Convolutional Network for 2 Time-Domain Marine CSEM Inversion

HardwareDGX agent

arXiv:2605.04997v1 Announce Type: new Abstract: DualTCN is the first deep-learning framework for inverting time-domain marine controlled-source electromagnetic (MCSEM) transient data. Moving away from

Hot take on Elon’s surprise decision to rent 30 megawatts of compute to Anthropic: 1. It’s a tacit concession that xAI is not all that close…

HardwareDGX agent

Hot take on Elon’s surprise decision to rent 30 megawatts of compute to Anthropic: 1. It’s a tacit concession that xAI is not all that close to AGI (despite what he suggested last year). 2. It’s more

.@huggingface's agentic robotics app store for Reachy Mini is a big step toward more accessible physical AI. 🙌 Excited to see NVIDIA Isaac …

HardwareDGX agent

.@huggingface's agentic robotics app store for Reachy Mini is a big step toward more accessible physical AI. 🙌 Excited to see NVIDIA Isaac GR00T N integrated with Hugging Face LeRobot, helping develop

I wish you could filter resteruant reviews and rankings on Google maps searches for by race Gotta authentic ethnic food max Too much overly …

HardwareDGX agent

This post discusses a request for filtering functionality on Google Maps that would allow users to search for restaurants by ethnic cuisine type or authenticity level, suggesting a desire for better t

Lambda closes $1 billion senior secured credit facility to meet gigawatt-scale AI infrastructure demand

HardwareDGX agent

Lambda Labs secured a $1 billion senior secured credit facility to fund the expansion of its AI infrastructure and meet growing demand for gigawatt-scale computational resources. This financing enable

Linked and Loaded: Gaijin Single Sign-On Now Available on GeForce NOW

HardwareDGX agent

Less typing, more tanking. Faster logins mean more time in the gaming action — and this week provides GeForce NOW members with a smoother path straight into the battlefield. Cloud gaming is all about

Massively Parallel Exact Inference for Hawkes Processes

HardwareDGX agent

arXiv:2604.01342v2 Announce Type: replace Abstract: Multivariate Hawkes processes are a widely used class of self-exciting point processes, but maximum likelihood estimation naively scales as O(N^2) i

Moonshot, the Chinese AI startup behind Kimi chatbot, raised ~2B at a 20B+ valuation led by Meituan's venture arm; Moonshot's ARR topped $200M in April 2026 (Zheping Huang/Bloomberg)

HardwareDGX agent

Zheping Huang / Bloomberg: Moonshot, the Chinese AI startup behind Kimi chatbot, raised ~2B at a 20B+ valuation led by Meituan's venture arm; Moonshot's ARR topped 200M in April 2026 — Moonshot AI has

Nvidia and data center operator IREN announce a deal to deploy up to 5 GW of AI infrastructure; Nvidia can invest $2.1B into IREN; IREN jumps 9%+ after hours (Jonathan Vanian/CNBC)

HardwareDGX agent

Jonathan Vanian / CNBC: Nvidia and data center operator IREN announce a deal to deploy up to 5 GW of AI infrastructure; Nvidia can invest $2.1B into IREN; IREN jumps 9%+ after hours — IREN shares surg

OptiLookUp: An Optical ROM-Based Loop up Table Engine for Photonic Accelerators

HardwareDGX agent

arXiv:2605.03241v1 Announce Type: cross Abstract: Read-only memory (ROM) provides deterministic access to predefined data mappings. Extending ROM concepts to the optical domain enables high-bandwidth,

Powering the Next American Century: US Energy Secretary Chris Wright and NVIDIA’s Ian Buck on the Genesis Mission

HardwareDGX agent

AI will help build the energy it needs. That’s the case U.S. Energy Secretary Chris Wright and NVIDIA Vice President of Hyperscale and High-Performance Computing Ian Buck made Thursday morning at the

Quadrature-TreeSHAP: Depth-Independent TreeSHAP and Shapley Interactions

HardwareDGX agent

arXiv:2605.04497v1 Announce Type: new Abstract: Shapley values are a standard tool for explaining predictions of tree ensembles, with Path-Dependent SHAP being the most widely used variant. Despite su

Quantum Motion raises $160M to build faster quantum chips

HardwareDGX agent

Quantum computer maker Quantum Motion Ltd. today announced that it has raised 160 million in funding to enhance its silicon-based qubit technology. The Series C round was led by DCVC and Kembara. It c

Real-Time Performance Monitoring and Faster Debugging with NCCL Inspector and Prometheus

HardwareDGX agent

NCCL Inspector is a profiling plugin that provides detailed, per-communicator, per-collective performance and metadata logging, designed to help users analyze and debug NCCL collective operations by g

Secure short-term GPU capacity for ML workloads with EC2 Capacity Blocks for ML and SageMaker training plans

HardwareDGX agent

In this post, you will learn how to secure reserved GPU capacity for short-term workloads using Amazon Elastic Compute Cloud (Amazon EC2) Capacity Blocks for ML and Amazon SageMaker training plans. Th

SemiConLens: Visual Analytics for 2D Semiconductor Discovery

HardwareDGX agent

arXiv:2605.04067v1 Announce Type: cross Abstract: The past few years have witnessed vibrant efforts in discovering new two-dimensional (2D) semiconductor materials from both academia and the industry,

Vultr Archival Object Storage Now Available

HardwareDGX agent

Vultr has launched Archival Object Storage, a new storage service offering cost-effective long-term data retention and backup solutions. This service is designed for infrequently accessed data with lo

We are so used to seeing chip company marketing teams exaggerate specs that it is refreshing to see them understate specs for a change. Here…

HardwareDGX agent

We are so used to seeing chip company marketing teams exaggerate specs that it is refreshing to see them understate specs for a change. Here's one example from Cerebras's website, where they understat

When you're sad you can either become and emotional eater or an emotional lifter. The latter is so much better.

HardwareDGX agent

The post discusses two contrasting responses to sadness: emotional eating (using food as a coping mechanism) versus emotional lifting (likely referring to exercise or physical activity as a healthier

6 May 2026

A great conversation between Noah Kravitz from the @nvidia team + @hwchase17.

HardwareDGX agent

A great conversation between Noah Kravitz from the @nvidia team + @hwchase17. “Every enterprise needs a claw strategy.” How did @LangChain go from a weekend project to 1B+ downloads in 3 years? We sat

AI supercomputers need a new kind of network to stay in sync at massive scale. OpenAI’s @markjhandley and @poyntingatgreg join @AndrewMayne …

HardwareDGX agent

AI supercomputers need a new kind of network to stay in sync at massive scale. OpenAI’s @markjhandley and @poyntingatgreg join @AndrewMayne to discuss what it takes to move data across record numbers

[AINews] Silicon Valley gets Serious about Services

HardwareDGX agent

Silicon Valley companies are increasingly shifting focus toward AI service offerings and applications rather than solely developing foundational models, reflecting a maturing market where practical de

b9047

Local AiDGX agent

b9047 is a build release of llama.cpp, an open-source C/C++ library for running large language model inference on consumer hardware. The release includes pre-compiled binaries for multiple platforms i

Bengaluru-based Pronto, an on-demand home-help service, raised a 20M Series B extension from Lachy Groom at a 200M valuation, up from $100M in March (Jagmeet Singh/TechCrunch)

HardwareDGX agent

Jagmeet Singh / TechCrunch: Bengaluru-based Pronto, an on-demand home-help service, raised a 20M Series B extension from Lachy Groom at a 200M valuation, up from $100M in March — Lachy Groom, one of S

DARTH-PUM: A Hybrid Processing-Using-Memory Architecture

ResearchDGX agent

arXiv:2602.16075v2 Announce Type: replace-cross Abstract: Analog processing-using-memory (PUM; a.k.a. in-memory computing) makes use of electrical interactions inside memory arrays to perform bulk mat

Exciting to work with @googledevs . Dflash is one of the most powerful technique developed here at UCSD by @zhijianliu_ and @jianchen1799 an…

HardwareDGX agent

Exciting to work with @googledevs . Dflash is one of the most powerful technique developed here at UCSD by @zhijianliu_ and @jianchen1799 and glad that our students and collaborators help port them in

I mogged @jxnlco at a pushup competiton

HardwareDGX agent

Dylan Patel posted about winning a pushup competition against a user named @jxnlco on X (formerly Twitter). The post uses casual internet slang ('mogged,' meaning decisively outperformed) to describe

Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute

HardwareDGX agent

arXiv:2504.17816v3 Announce Type: replace Abstract: Subject-driven video generation (SDV-Gen) aims to produce videos of a specific subject by adapting a pretrained video model, enabling personalized a

MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier

HardwareDGX agent

arXiv:2603.03756v3 Announce Type: replace-cross Abstract: While large language models (LLMs) show promise in scientific discovery, existing research focuses on inference or feedback-driven training, l

NVIDIA Spectrum-X — the Open, AI-Native Ethernet Fabric — Sets the Standard for Gigascale AI, Now With MRC

HardwareDGX agent

The race to build the world’s most powerful AI factories demands networking that keeps pace with the ambitions of AI itself. NVIDIA Spectrum-X Ethernet scale-out infrastructure stands at the forefront

Nvidia’s MRC: When ‘just Ethernet’ isn’t enough for gigascale AI

HardwareDGX agent

Nvidia Corp.‘s latest networking innovations meet the needs of a new kind of network that supports the unique demands of artificial intelligence factories. Ethernet is no longer a generic plumbing cho

One Sequence to Segment Them All: Efficient Data Augmentation for CT and MRI Cross-Domain 3D Spine Segmentation

HardwareDGX agent

arXiv:2605.03098v1 Announce Type: new Abstract: Deep learning-based medical image segmentation is increasingly used to support clinical diagnosis and develop new treatment strategies. However, model p

Robust Visual SLAM for UAV Navigation in GPS-Denied and Degraded Environments: A Multi-Paradigm Evaluation and Deployment Study

HardwareDGX agent

arXiv:2605.03678v1 Announce Type: new Abstract: Reliable localization in GPS-denied, visually degraded environments is critical for autonomous UAV opera- tions. This paper presents a systematic compar

Test-Time Training with KV Binding Is Secretly Linear Attention

HardwareDGX agent

arXiv:2602.21204v3 Announce Type: replace-cross Abstract: Test-time training (TTT) with KV binding as sequence modeling layer is commonly interpreted as a form of online meta-learning that memorizes a

The GB300 is the best AI computer

HardwareDGX agent

The GB300 is the best AI computer Two frontier labs. One accelerated computing platform. Congrats to @SpaceX and @AnthropicAI on the new compute partnership, powered by 220,000+ NVIDIA GPUs inside Col

The Polar Express: Optimal Matrix Sign Methods and Their Application to the Muon Algorithm

HardwareDGX agent

arXiv:2505.16932v5 Announce Type: replace-cross Abstract: Computing the polar decomposition and the related matrix sign function has been a well-studied problem in numerical analysis for decades. Rece

👇 This is an incredible admission by Jensen Huang. Effectively he is saying that AI, for all the hype (some from himself) wasn’t really use…

HardwareDGX agent

👇 This is an incredible admission by Jensen Huang. Effectively he is saying that AI, for all the hype (some from himself) wasn’t really useful until late 2025. Let that sink in. This means practically

We’ve partnered with @AMD, @Broadcom, @Intel, @Microsoft, and @NVIDIA, to release Multipath Reliable Connection (MRC), a new open networking…

HardwareDGX agent

We’ve partnered with @AMD, @Broadcom, @Intel, @Microsoft, and @NVIDIA, to release Multipath Reliable Connection (MRC), a new open networking protocol that helps large AI training clusters run faster a

ZeRO-Prefill: Zero Redundancy Overheads in MoE Prefill Serving

HardwareDGX agent

arXiv:2605.02960v1 Announce Type: new Abstract: Production LLM workloads increasingly serve discriminative tasks, such as classification, recommendation, and verification, whose answers are read from

5 May 2026

AgriKD: Cross-Architecture Knowledge Distillation for Efficient Leaf Disease Classification

HardwareDGX agent

arXiv:2605.01355v1 Announce Type: new Abstract: Automated leaf disease classification is critical for early disease detection in resource-constrained field environments. Vision Transformers (ViTs) pro

Blitzy raises 200M at 1.4B valuation to deploy thousands of coding agents in parallel

HardwareDGX agent

Autonomous software development startup Blitzy Inc. said today it has raised 200 million in new funding on a valuation of 1.4 billion to expand its enterprise coding platform. The company was founded

Can AI Debias the News? LLM Interventions Improve Cross-Partisan Receptivity but LLMs Overestimate Their Own Effectiveness

HardwareDGX agent

arXiv:2605.01006v1 Announce Type: new Abstract: Partisan news media erode cross-partisan trust, but large language models (LLMs) offer a potential means of debiasing such content at scale. Across two

Deepinfra lands $107M in funding to build out its dedicated inference cloud for open-source models

HardwareDGX agent

Dedicated inference cloud startup Deepinfra Inc. is looking to expand its global capacity after raising 107 million in a Series B round of funding led by 500 Global and Georges Harik, who was one of G

DELTA: Dynamic Layer-Aware Token Attention for Efficient Long-Context Reasoning

HardwareDGX agent

arXiv:2510.09883v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve state-of-the-art performance on challenging benchmarks by generating long chains of intermediate steps, but th

← Previous
1…2829303132…75
Next →