AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
26 May 2026

Dropbox founder Drew Houston is stepping down as CEO after 19 years to become executive chairman, replaced by Ashraf Alkarmi, who is SVP and GM of Dropbox Core (Jonathan Vanian/CNBC)

HardwareDGX agent

Jonathan Vanian / CNBC: Dropbox founder Drew Houston is stepping down as CEO after 19 years to become executive chairman, replaced by Ashraf Alkarmi, who is SVP and GM of Dropbox Core — Drew Houston f

H^{2}MT: Semantic Hierarchy-Aware Hierarchical Memory Transformer

HardwareDGX agent

arXiv:2605.24930v1 Announce Type: new Abstract: Transformer-based LLMs achieve strong results on many language tasks; however, long inputs remain challenging because context windows are finite, and pr

How we evolved Google’s global and data center networks for the AI era

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Over the last 25 years of building Google’s global network, we’ve navigated major architectural eras — from the Internet, to streaming, and the cloud. Today, we are squarely in the midst of a fourth:

HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos

SafetyDGX agent

arXiv:2605.24934v1 Announce Type: cross Abstract: Human egocentric video captures rich manipulation demonstrations without any robot hardware, yet transferring these skills to robots remains challengi

I have been very impressed by @SemiAnalysis_ . I think of myself as a wide ranging systems engineer, looking for value at every level from t…

HardwareDGX agent

I have been very impressed by @SemiAnalysis_ . I think of myself as a wide ranging systems engineer, looking for value at every level from the chip specs to the user interface, but SA exposes me to ad

Initial benchmarks show Nvidia's Vera CPU, which features 88 in-house-designed Olympus cores, packs a heavy-hitting punch, beating Intel's and AMD's x86_64 CPUs (Michael Larabel/Phoronix)

HardwareDGX agent

Michael Larabel / Phoronix: Initial benchmarks show Nvidia's Vera CPU, which features 88 in-house-designed Olympus cores, packs a heavy-hitting punch, beating Intel's and AMD's x86_64 CPUs — NVIDIA's

Inside the 800VDC Revolution – Part 1

HardwareDGX agent

This article examines the emerging 800V direct current (800VDC) electrical architecture in electric vehicles, which represents a significant advancement over traditional 400V systems by enabling faste

Inside the 800VDC Revolution – Part 1 Four-Phase 800VDC Transition, Power Rack Economics, SST, Equipment Content/MW Build, Supplier Implicat…

HardwareDGX agent

Inside the 800VDC Revolution – Part 1 Four-Phase 800VDC Transition, Power Rack Economics, SST, Equipment Content/MW Build, Supplier Implications https://newsletter.semianalysis.com/p/inside-the-800vdc

Learning, Solving and Optimizing PDEs with TensorGalerkin: an efficient high-performance Galerkin assembly algorithm

HardwareDGX agent

arXiv:2602.05052v3 Announce Type: replace Abstract: We present a unified algorithmic framework for the numerical solution, constrained optimization, and physics-informed learning of PDEs with a variat

MathOptAI.jl: Embed trained machine learning predictors into JuMP models

HardwareDGX agent

arXiv:2507.03159v2 Announce Type: replace Abstract: We present exttt{MathOptAI.jl}, an open-source Julia library for embedding trained machine learning predictors into a JuMP model. exttt{MathOptAI.jl

Nvidia officially retires its GeForce Control Panel app after 20 years, following the porting of all of its major features to the Nvidia app (Tom Warren/The Verge)

HardwareDGX agent

Tom Warren / The Verge: Nvidia officially retires its GeForce Control Panel app after 20 years, following the porting of all of its major features to the Nvidia app — Nvidia has now ported across all

OpenRouter raises $113M to bring order to enterprise AI inference routing

HardwareDGX agent

Artificial intelligence inference routing startup OpenRouter Inc. today announced it raised 113 million in new funding led by CapitalG, Alphabet Inc.’s independent growth fund, to push the envelope fo

Optimizing Multidimensional Scaling in Gini Metric Spaces

HardwareDGX agent

arXiv:2605.25124v1 Announce Type: new Abstract: The Gini Multidimensional Scaling (Gini MDS) framework extends the Euclidean multidimensional scaling. We introduce a Gini pseudo-distance based on valu

Paris 2.0: A Decentralized Diffusion Model for Video Generation

HardwareDGX agent

arXiv:2605.26064v1 Announce Type: cross Abstract: We present Paris 2.0, the first video generation model pre-trained through decentralized computation. Its training recipe builds upon Paris 1.0 (arXiv

Profiling-Driven Adaptive Distributed Transformer Inference on Embedded Edge Deployment

Local AiDGX agent

arXiv:2605.25682v1 Announce Type: cross Abstract: Distributing Transformer inference across embedded edge devices can alleviate individual memory and compute constraints, yet practical benefits on rea

Rambus targets agentic AI workloads with faster client memory chipset

HardwareDGX agent

Rambus Inc. today announced a complete DDR5 9600 client memory module chipset designed to push PC memory speeds to 9,600 megatransfers per second, targeting the bandwidth and capacity demands of agent

recommended reading. i too am very done with people anthropomorphizing a bunch of matrices on a GPU cluster, especially if the same people d…

HardwareDGX agent

recommended reading. i too am very done with people anthropomorphizing a bunch of matrices on a GPU cluster, especially if the same people do not give two fucks about actual human beings. More musings

Run Key Genomics and Protein Folding Workloads Faster with NVIDIA RTX PRO 4500 Blackwell

HardwareDGX agent

The NVIDIA RTX PRO 4500 accelerates genomic sequencing and drug discovery by slashing processing times. This professional desktop GPU features 32GB of ultra-fast GDDR7 memory and Blackwell architectur

We aren’t going to do this again so quickly, are we? Rising demand results in higher costs. Higher costs result in lower demand. It is almos…

HardwareDGX agent

We aren’t going to do this again so quickly, are we? Rising demand results in higher costs. Higher costs result in lower demand. It is almost like some sort of equilibrium is being achieved. But there

25 May 2026

Anthropic is really a mixed bag; they rightly support moderate regulation, but they invite wildly anthropomorphic language about LLMs, and s…

HardwareDGX agent

Anthropic is really a mixed bag; they rightly support moderate regulation, but they invite wildly anthropomorphic language about LLMs, and steal IP just like their peers. Anthropic are the bad guys. W

China’s Huawei unveils new sanctions-busting chip architecture that replaces Moore’s Law

HardwareDGX agent

Chinese electronics giant Huawei Technologies Co. Ltd. today unveiled a new chip design framework that it says will help it to close the gap in the semiconductor industry with global leaders like Taiw

Convex Low-resource Accent-Robust Language Detection in Speech Recognition

HardwareDGX agent

arXiv:2605.23235v1 Announce Type: new Abstract: Globalization and multiculturalism continue to produce increasingly diverse speech varieties. Yet current spoken dialogue systems frequently fail on und

Dell and Nvidia target the data chaos standing between AI pilots and real results

HardwareDGX agent

As organizations push to convert AI ambition into defensible business results, many are discovering that infrastructure alone is not enough. Instead, the real challenge lies in reconciling decades of

DFSAttn: Dynamic Fine-grained Sparse Attention for Efficient Video Generation

HardwareDGX agent

arXiv:2605.23445v1 Announce Type: new Abstract: Diffusion transformers have achieved remarkable success in high-quality video generation, yet their reliance on spatiotemporal 3D full attention incurs

Graph Learning via Logic-Based Weisfeiler-Leman Variants and Tabularization

HardwareDGX agent

arXiv:2508.10651v3 Announce Type: replace Abstract: We present a novel approach for graph classification based on tabularizing graph data via new variants of the Weisfeiler-Leman algorithm and then ap

If I wrote a short, cheap, electronic sequel to Taming Silicon Valley, perhaps with a title like “Getting to an AI that is good for all of h…

HardwareDGX agent

If I wrote a short, cheap, electronic sequel to Taming Silicon Valley, perhaps with a title like “Getting to an AI that is good for all of humanity” (perhaps donating proceeds) would you read it and g

MiniCPM5-1B is an impressive release in the 1B class! @OpenBMB https://huggingface.co/collections/openbmb/minicpm5 ✨ 1B - Apache 2.0 ✨ Hybri…

HardwareDGX agent

MiniCPM5-1B is an impressive release in the 1B class! @OpenBMB https://huggingface.co/collections/openbmb/minicpm5 ✨ 1B - Apache 2.0 ✨ Hybrid reasoning with Think / No-Think modes ✨ 128K context ✨ Run

PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion

HardwareDGX agent

arXiv:2605.23902v1 Announce Type: new Abstract: Most practical high-resolution text-to-image systems, including latent diffusion and autoregressive models, perform generation in a compact latent space

Sources: Meta, Google, and Amazon execs met Vatican officials on April 29, as part of a quiet lobbying push ahead of Pope Leo XIV's first AI encyclical (Océane Herrero/Politico)

HardwareDGX agent

Océane Herrero / Politico: Sources: Meta, Google, and Amazon execs met Vatican officials on April 29, as part of a quiet lobbying push ahead of Pope Leo XIV's first AI encyclical — As Leo XIV prepares

Taxes screw over both the companies and the states they’re in. Taxes didn’t build those big companies. They show up after the success. They …

HardwareDGX agent

Taxes screw over both the companies and the states they’re in. Taxes didn’t build those big companies. They show up after the success. They sure as hell didn’t create Silicon Valley. The sky-high taxe

Testing ZIT and Flux-1 with 'NVIDIA PiD — Pixel Diffusion Decoder'

HardwareDGX agent

NVIDIA's Pixel Diffusion Decoder (PiD) is an open-source decoder that replaces VAE decoders in image generation pipelines without retraining, producing sharper fine details and textures through a lear

ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attention

HardwareDGX agent

arXiv:2605.23081v1 Announce Type: new Abstract: Efficient attention algorithms are critical to mitigate the quadratic cost of attention in long-context workloads. Prior work utilises block-scaled quan

XWind: A Cross-site Router for Large Language Model Inference Serving at Renewable Energy Farms

HardwareDGX agent

arXiv:2605.23348v1 Announce Type: cross Abstract: AI power demand is growing at an unprecedented rate while power grids are often ailing and struggle to keep up. Grid expansion comes with high capital

24 May 2026

Our inference stack, optimized for Blackwells, with a novel attention kernel and many new optimizations has started rolling out! It's alread…

HardwareDGX agent

Our inference stack, optimized for Blackwells, with a novel attention kernel and many new optimizations has started rolling out! It's already charting on Artificial Analysis, eg: #1 speed and latency

23 May 2026

Dell says it has 5,000 clients for its AI Factory, a product line of servers with Nvidia chips, software, and services, including 1,000 new clients last quarter (Dina Bass/Bloomberg)

HardwareDGX agent

Dina Bass / Bloomberg: Dell says it has 5,000 clients for its AI Factory, a product line of servers with Nvidia chips, software, and services, including 1,000 new clients last quarter — Dell Technolog

FlashSinkhorn: IO-Aware Entropic Optimal Transport on GPU

HardwareDGX agent

arXiv:2602.03067v3 Announce Type: replace Abstract: Entropic optimal transport (EOT) via Sinkhorn iterations is widely used in modern machine learning, yet GPU solvers remain inefficient at scale. Ten

From Sequential Nodes to GPU Batches: Parallel Branch and Bound for Optimal k-Sparse GLMs

HardwareDGX agent

arXiv:2605.22188v1 Announce Type: new Abstract: GPUs have significantly accelerated first-order methods for large-scale optimization, especially in continuous optimization. However, this success has n

ImplicitTerrainV2: Wavelet-Guided Spatially Adaptive Neural Terrain Representation

HardwareDGX agent

arXiv:2605.22556v1 Announce Type: new Abstract: Digital elevation models (DEMs) underpin terrain analysis in Geographic Information Systems (GIS), but in their common raster form, they rely on interpo

Jensen Huang urged Super Micro to tighten up compliance after Taiwan detained three people for allegedly trying to export servers with Nvidia chips to China (Debby Wu/Bloomberg)

HardwareDGX agent

Debby Wu / Bloomberg: Jensen Huang urged Super Micro to tighten up compliance after Taiwan detained three people for allegedly trying to export servers with Nvidia chips to China — Nvidia Corp. Chief

La plupart des grandes boîtes sont des organisations zombies. Voici pourquoi. Dans League of Legends, ton rang n'est pas un titre. C'est une…

HardwareDGX agent

La plupart des grandes boîtes sont des organisations zombies. Voici pourquoi. Dans League of Legends, ton rang n'est pas un titre. C'est une mesure continue. Tu es Master parce que tu joues comme un M

LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations

HardwareDGX agent

arXiv:2602.01935v2 Announce Type: replace Abstract: LLM-guided compiler optimization has recently shown promise, but existing approaches rely on a single large LLM throughout search, making them expen

Reading Task Failure Off the Activations: A Sparse-Feature Audit of GPT-2 Small on Indirect Object Identification

HardwareDGX agent

arXiv:2605.22719v1 Announce Type: new Abstract: We report a small, reproducible audit of which sparse-autoencoder (SAE) features of GPT-2 small fire differently on failed versus successful trials of t

someone else seeing what i am seeing:

HardwareDGX agent

someone else seeing what i am seeing: 🚨 MICHAEL BURRY JUST WARNED THE ENTIRE AI BOOM MAY BE BUILT ON TEMPORARY DEMAND. He published a post today calling Nvidia 'the North Star, Orion, the whole Milky

WarmServe: Enabling One-for-Many GPU Prewarming for Multi-LLM Serving

HardwareDGX agent

arXiv:2512.09472v2 Announce Type: replace-cross Abstract: Deploying multiple models within shared GPU clusters is a key strategy to improve resource efficiency in large language model (LLM) serving. E

22 May 2026

Don't Collapse Your Features: Why CenterLoss Hurts OOD Detection and Multi-Scale Mahalanobis Wins

HardwareDGX agent

arXiv:2605.21493v1 Announce Type: cross Abstract: The ability to detect out-of-distribution (OOD) inputs is fundamental to safe deployment of machine learning systems. Yet, current methods often rely

EasyVFX: Frequency-Driven Decoupling for Resource-Efficient VFX Generation

HardwareDGX agent

arXiv:2605.22051v1 Announce Type: new Abstract: Generating high-fidelity visual effects (VFX) typically demands massive datasets and prohibitive computational power due to the intricate coupling of sp

Happiest birthday to @dylan522p. Welcome to 30 🎂

HardwareDGX agent

Dylan Patel, founder of SemiAnalysis, celebrated his 30th birthday. The post is a personal milestone announcement shared on X (formerly Twitter). This appears to be a casual birthday greeting rather t

Hark raises $700M+ to build ‘personalized intelligence’ devices

HardwareDGX agent

Hark Inc., a startup developing artificial intelligence devices for the consumer market, today announced that it has raised more than 700 million in funding. Parkway Venture Capital led the Series A r

Mega AI IPOs incoming, Google’s agentic blitz and Nvidia’s next big business

HardwareDGX agent

Now we know for sure: This will be the year of monster initial public offerings. Elon Musk’s SpaceX filed for an IPO this week, aiming to raise a record $80 billion or more, and OpenAI was expected to

PALS: Power-Aware LLM Serving for Mixture-of-Experts Models

HardwareDGX agent

arXiv:2605.21427v1 Announce Type: new Abstract: Large language model (LLM) inference has become a dominant workload in modern data centers, driving significant GPU utilization and energy consumption.

PSA: Just added a thousand H100s and H200s to Together on-demand GPU clusters and Dedicated Endpoints: http://api.together.ai/clusters

HardwareDGX agent

Together AI announced the addition of 1,000 H100 and H200 GPUs to their on-demand GPU clusters and dedicated endpoints, expanding their available compute resources for users accessing services through

WorldKV: Efficient World Memory with World Retrieval and Compression

HardwareDGX agent

arXiv:2605.22718v1 Announce Type: new Abstract: Autoregressive video diffusion models have enabled real-time, action-conditioned world generation. However, sustaining a persistent world, where revisit

21 May 2026

Accelerating Video Inverse Problem Solvers with Autoregressive Diffusion Models

HardwareDGX agent

arXiv:2605.20624v1 Announce Type: new Abstract: Diffusion models provide powerful priors for zero-shot video inverse problems, but their real-time deployment is hindered by two inefficiencies: high in

Building Token‑Metered AI Services on Telco AI Factories

HardwareDGX agent

Telcos are building sovereign AI factories based on the NVIDIA Cloud Partner reference architecture to provide secure, in-country AI infrastructure with controls and performance suited for enterprise

Depth Completion in Unseen Field Robotics Environments Using Extremely Sparse Depth Measurements

HardwareDGX agent

arXiv:2602.03209v2 Announce Type: replace Abstract: Autonomous field robots operating in unstructured environments require robust perception to ensure safe and reliable operations. Recent advances in

Fine-tuning used to mean a team, a GPU cluster, and weeks of iteration. Now it's just a CLI command, ~10 min of GPU time, a few cents of com…

HardwareDGX agent

Fine-tuning used to mean a team, a GPU cluster, and weeks of iteration. Now it's just a CLI command, ~10 min of GPU time, a few cents of compute. You walk away owning the weights. Open models off the

Five takeaways from Nvidia’s earnings, and what they mean for the AI industry

HardwareDGX agent

Nvidia Corp.‘s latest results are remarkable even by its own high bar. Revenue set new records, driven by a 90%-plus surge in data center demand and what management described as “parabolic” demand as

Frontier: Towards Comprehensive and Accurate LLM Inference Simulation

HardwareDGX agent

arXiv:2605.21312v1 Announce Type: cross Abstract: Modern LLM serving is no longer homogeneous or monolithic. Production systems now combine disaggregated execution, complex parallelism, runtime optimi

HyperBones: Realtime Bone-driven Neural Garment Simulation with Hypernetwork Conditioning

HardwareDGX agent

arXiv:2605.20460v1 Announce Type: cross Abstract: Recent advances in garment simulation have brought high-quality results closer to real-time performance. Physics-based simulators can produce accurate

I got tired of API limits, so I hooked up OpenClaw to an unlimited Qwen3.6:35b backend on a full H100 for $1.6/hr (Demo)

HardwareDGX agent

This post describes setting up OpenClaw with a self-hosted Qwen 3.6:35b language model backend running on an H100 GPU for approximately $1.60 per hour, eliminating API rate limits. The user shares the

← Previous
1…2425262728…75
Next →