AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,734 results
27 May 2026

Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 (Nikkei Asia)

HardwareDGX agent

Nikkei Asia: Jensen Huang says Nvidia is spending 100B-150B annually on its Taiwan supply chain, up from 10B-15B four years ago, and will boost its Taiwan staff to 4,000 — TAIPEI — Nvidia is now spend

JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search

HardwareDGX agent

arXiv:2605.26636v1 Announce Type: cross Abstract: We introduce JetViT, a novel family of hybrid-architecture Vision Transformer (ViT) models that match the accuracy of state-of-the-art full-attention

Morphling: Fast, Fused, and Flexible GNN Training at Scale

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.01678v5 Announce Type: replace Abstract: Graph Neural Networks (GNNs) present a fundamental hardware challenge by fusing irregular, memory-bound graph traversals with regular, compute-inten

Multi-Robot Box Transport over Different Surfaces with Decentralized Role-based Proportional Control

HardwareDGX agent

arXiv:2605.26430v1 Announce Type: new Abstract: Collaborative transport of objects via pushing by multiple robots has many applications, ranging from construction and warehouse environments to post di

NightSight: Passive Computation for Navigation in Dark Using Events

HardwareDGX agent

arXiv:2605.26330v1 Announce Type: new Abstract: Small aerial robots are particularly well-suited for search and rescue in confined and hazardous environments due to their agility, low cost, and abilit

Nvidia bets $150B on Taiwan as Trump's plan to make US an AI hub backfires

HardwareDGX agent

Nvidia CEO Jensen Huang announced the chip company plans to invest approximately 150 billion annually in Taiwan, terming it the 'epicentre' of the AI revolution . This represents a dramatic increase f

NVIDIA Blackwell Sets STAC-AI Record for LLM Inference in Finance

HardwareDGX agent

NVIDIA's Blackwell GB200 NVL72 architecture achieved the fastest-ever results on the STAC-AI benchmark for financial LLM inference, delivering up to 3.2x single-GPU performance improvements over the p

NVIDIA Dynamo Snapshot: Fast Startup for Inference Workloads on Kubernetes

HardwareDGX agent

NVIDIA Dynamo Snapshot introduces ModelExpress for 7x faster startup via checkpoint restore and weight streaming with NVIDIA NVLink and NIXL , enabling efficient model initialization for inference wor

Nvidia kills Windows XP-era Control Panel 'after 20 years of dedicated service'

HardwareDGX agent

Nvidia has officially retired the NVIDIA Control Panel after 20 years of service, shifting all functionality to the newer Nvidia App. The legacy Control Panel is no longer included in default GeForce

Qrita: High-performance Top-k and Top-p using Pivot-based Truncation and Selection

HardwareDGX agent

arXiv:2602.01518v2 Announce Type: replace Abstract: Despite their importance in model sampling, efficient implementation of Top-k and Top-p algorithms for large vocabularies remains a significant chal

RT-Lynx: Putting the GEMM Sparsity In a Right Way for Diffusion Models

HardwareDGX agent

arXiv:2605.26632v1 Announce Type: new Abstract: Diffusion Transformers (DiT) achieve strong performance in image generation but incur substantial inference costs. While prior work has reduced this cos

SIA: Self Improving AI with Harness & Weight Updates

HardwareDGX agent

arXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l

SIKA-GP: Accelerating Gaussian Process Inference with Sparse Inducing Kernel Approximations for Bayesian Deep Learning

HardwareDGX agent

arXiv:2605.26509v1 Announce Type: new Abstract: Gaussian processes (GPs) provide a principled Bayesian framework for uncertainty estimation, but their computational complexity severely limits scalabil

Sources: Taiwan suspects three people smuggled at least one Nvidia chip shipment to China after first exporting them to Japan; Taiwan detained them last week (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Taiwan suspects three people smuggled at least one Nvidia chip shipment to China after first exporting them to Japan; Taiwan detained them last week — Taiwan prosecutors suspect th

Tensormesh taps Nvidia, AMD and CoreWeave for funding to fix AI model memory problems

HardwareDGX agent

Tensormesh Inc. has hit upon a way to make artificial intelligence inference more efficient by eliminating the need for redundant computations, and its technology is so convincing that several of AI i

The AI GPU market has belonged to hyperscalers, but AMD’s MI350P is coming for the enterprise

HardwareDGX agent

Enterprise GPU access is expanding as Advanced Micro Devices Inc. lowers barriers with the air-cooled Instinct MI350P GPU and a ready-to-run software stack designed for standard enterprise servers. Th

Towards Feedback-to-Plan Decisions for Self-Evolving LLM Agents in CUDA Kernel Generation

HardwareDGX agent

arXiv:2605.26720v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong empirical gains as self-evolving agents for CUDA kernel generation, driven by feedback-conditioned planni

We're open-sourcing the Unigram tokenizer we rebuilt to reduce CPU utilization by 5-6x. Small rerankers and embedders run in single-digit mi…

HardwareDGX agent

We're open-sourcing the Unigram tokenizer we rebuilt to reduce CPU utilization by 5-6x. Small rerankers and embedders run in single-digit milliseconds on GPU, making CPU tokenization a meaningful shar

What’s New for Game Developers in NVIDIA RTX: DLSS 4.5 for UE5 and Multilingual AI Characters

HardwareDGX agent

NVIDIA DLSS 4.5 introduces Dynamic Multi Frame Generation and a second-generation transformer model for Super Resolution to enhance game image quality and frame rates for Unreal Engine 5. The new Tens

Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU

HardwareDGX agent

arXiv:2605.26118v1 Announce Type: cross Abstract: Porting deep learning algorithms to new hardware accelerators requires developers to repeatedly apply the same low-level optimizations -- quantization

26 May 2026

A universal vision transformer for fast calorimeter simulations

HardwareDGX agent

arXiv:2601.05289v2 Announce Type: replace-cross Abstract: The high-dimensional complex nature of detectors makes fast calorimeter simulations a prime application for modern generative machine learning

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 …

HardwareDGX agent

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 all the way to 24.04 the other day). Asked Codex on my Mac t

AgentIR: A Workload-Adaptive Cascade Retrieval Substrate for Long-Term Conversational Memory

HardwareDGX agent

arXiv:2605.25092v1 Announce Type: cross Abstract: Long-term conversational memory is a retrieval workload classical IR was not built for: the index grows during the query stream, query types shift int

AI readiness in telecommunications

HardwareDGX agent

This Databricks blog post likely examines the prerequisites and preparedness required for telecommunications companies to successfully implement and deploy AI solutions. It probable covers organizatio

As agentic AI surges, CPUs and air-cooled infrastructure move to the fore

HardwareDGX agent

Agents are turning up the heat for enterprises, turning air-cooled AI infrastructure into a boardroom priority. While GPUs have largely dominated the AI conversation, CPUs are increasingly coming into

Beyond the Aggregation Dilemma: Prior-Retaining Decoupled Learning for Multimodal Graphs

HardwareDGX agent

arXiv:2605.24684v1 Announce Type: cross Abstract: Multimodal Attributed Graph Learning (MAGL) integrates intrinsic node attributes with structural topology via graph aggregation. However, as pretraine

Build high-performance generative AI systems with Strands Agents, NVIDIA NIM, and Amazon Bedrock AgentCore

HardwareDGX agent

In this post you'll learn how to build a multi-agent campaign review system that demonstrates parallel reasoning, context persistence, and traceable execution paths using an integrated architecture th

Develop High-Performance GPU Kernels in C++ with NVIDIA CUDA Tile

HardwareDGX agent

This NVIDIA Developer article guides developers on creating optimized GPU kernels using C++ with NVIDIA's CUDA Tile technology, which enables efficient parallel computation on graphics processors. The

Dropbox founder Drew Houston is stepping down as CEO after 19 years to become executive chairman, replaced by Ashraf Alkarmi, who is SVP and GM of Dropbox Core (Jonathan Vanian/CNBC)

HardwareDGX agent

Jonathan Vanian / CNBC: Dropbox founder Drew Houston is stepping down as CEO after 19 years to become executive chairman, replaced by Ashraf Alkarmi, who is SVP and GM of Dropbox Core — Drew Houston f

H^{2}MT: Semantic Hierarchy-Aware Hierarchical Memory Transformer

HardwareDGX agent

arXiv:2605.24930v1 Announce Type: new Abstract: Transformer-based LLMs achieve strong results on many language tasks; however, long inputs remain challenging because context windows are finite, and pr

I have been very impressed by @SemiAnalysis_ . I think of myself as a wide ranging systems engineer, looking for value at every level from t…

HardwareDGX agent

I have been very impressed by @SemiAnalysis_ . I think of myself as a wide ranging systems engineer, looking for value at every level from the chip specs to the user interface, but SA exposes me to ad

Inference Time Context Sparsity: Illusion or Opportunity?

HardwareDGX agent

arXiv:2605.24168v1 Announce Type: new Abstract: Sparsity has long been a central theme in LLM efficiency, but its role in context processing remains unresolved. As LLM workloads shift toward longer co

Initial benchmarks show Nvidia's Vera CPU, which features 88 in-house-designed Olympus cores, packs a heavy-hitting punch, beating Intel's and AMD's x86_64 CPUs (Michael Larabel/Phoronix)

HardwareDGX agent

Michael Larabel / Phoronix: Initial benchmarks show Nvidia's Vera CPU, which features 88 in-house-designed Olympus cores, packs a heavy-hitting punch, beating Intel's and AMD's x86_64 CPUs — NVIDIA's

Inside the 800VDC Revolution – Part 1

HardwareDGX agent

This article examines the emerging 800V direct current (800VDC) electrical architecture in electric vehicles, which represents a significant advancement over traditional 400V systems by enabling faste

Inside the 800VDC Revolution – Part 1 Four-Phase 800VDC Transition, Power Rack Economics, SST, Equipment Content/MW Build, Supplier Implicat…

HardwareDGX agent

Inside the 800VDC Revolution – Part 1 Four-Phase 800VDC Transition, Power Rack Economics, SST, Equipment Content/MW Build, Supplier Implications https://newsletter.semianalysis.com/p/inside-the-800vdc

Learning, Solving and Optimizing PDEs with TensorGalerkin: an efficient high-performance Galerkin assembly algorithm

HardwareDGX agent

arXiv:2602.05052v3 Announce Type: replace Abstract: We present a unified algorithmic framework for the numerical solution, constrained optimization, and physics-informed learning of PDEs with a variat

MathOptAI.jl: Embed trained machine learning predictors into JuMP models

HardwareDGX agent

arXiv:2507.03159v2 Announce Type: replace Abstract: We present exttt{MathOptAI.jl}, an open-source Julia library for embedding trained machine learning predictors into a JuMP model. exttt{MathOptAI.jl

Nvidia officially retires its GeForce Control Panel app after 20 years, following the porting of all of its major features to the Nvidia app (Tom Warren/The Verge)

HardwareDGX agent

Tom Warren / The Verge: Nvidia officially retires its GeForce Control Panel app after 20 years, following the porting of all of its major features to the Nvidia app — Nvidia has now ported across all

OpenRouter raises $113M to bring order to enterprise AI inference routing

HardwareDGX agent

Artificial intelligence inference routing startup OpenRouter Inc. today announced it raised 113 million in new funding led by CapitalG, Alphabet Inc.’s independent growth fund, to push the envelope fo

Optimizing Multidimensional Scaling in Gini Metric Spaces

HardwareDGX agent

arXiv:2605.25124v1 Announce Type: new Abstract: The Gini Multidimensional Scaling (Gini MDS) framework extends the Euclidean multidimensional scaling. We introduce a Gini pseudo-distance based on valu

Parallel Differentiable Reachability for Learning and Planning with Certified Neural Dynamics and Controllers

HardwareDGX agent

arXiv:2605.25346v1 Announce Type: cross Abstract: Neural network (NN) dynamics models and control policies achieve strong performance in robotics, but providing sound guarantees under uncertainty rema

Paris 2.0: A Decentralized Diffusion Model for Video Generation

HardwareDGX agent

arXiv:2605.26064v1 Announce Type: cross Abstract: We present Paris 2.0, the first video generation model pre-trained through decentralized computation. Its training recipe builds upon Paris 1.0 (arXiv

Physical Analogue Kolmogorov-Arnold Networks based on Reconfigurable Nonlinear-Processing Units

HardwareDGX agent

arXiv:2602.07518v2 Announce Type: replace-cross Abstract: Kolmogorov-Arnold Networks (KANs) shift neural computation from linear layers to learnable nonlinear edge functions, but implementing these no

Rambus targets agentic AI workloads with faster client memory chipset

HardwareDGX agent

Rambus Inc. today announced a complete DDR5 9600 client memory module chipset designed to push PC memory speeds to 9,600 megatransfers per second, targeting the bandwidth and capacity demands of agent

recommended reading. i too am very done with people anthropomorphizing a bunch of matrices on a GPU cluster, especially if the same people d…

HardwareDGX agent

recommended reading. i too am very done with people anthropomorphizing a bunch of matrices on a GPU cluster, especially if the same people do not give two fucks about actual human beings. More musings

Run Key Genomics and Protein Folding Workloads Faster with NVIDIA RTX PRO 4500 Blackwell

HardwareDGX agent

The NVIDIA RTX PRO 4500 accelerates genomic sequencing and drug discovery by slashing processing times. This professional desktop GPU features 32GB of ultra-fast GDDR7 memory and Blackwell architectur

SA-Kura: An Energy-Efficient Systolic Array Accelerator for Locally-Coupled Kuramoto Drift in Diffusion Sampling

HardwareDGX agent

arXiv:2605.24016v1 Announce Type: cross Abstract: Diffusion inference remains costly for edge deployment, yet existing accelerators focus almost exclusively on score networks because standard drift is

We aren’t going to do this again so quickly, are we? Rising demand results in higher costs. Higher costs result in lower demand. It is almos…

HardwareDGX agent

We aren’t going to do this again so quickly, are we? Rising demand results in higher costs. Higher costs result in lower demand. It is almost like some sort of equilibrium is being achieved. But there

25 May 2026

Anthropic is really a mixed bag; they rightly support moderate regulation, but they invite wildly anthropomorphic language about LLMs, and s…

HardwareDGX agent

Anthropic is really a mixed bag; they rightly support moderate regulation, but they invite wildly anthropomorphic language about LLMs, and steal IP just like their peers. Anthropic are the bad guys. W

China’s Huawei unveils new sanctions-busting chip architecture that replaces Moore’s Law

HardwareDGX agent

Chinese electronics giant Huawei Technologies Co. Ltd. today unveiled a new chip design framework that it says will help it to close the gap in the semiconductor industry with global leaders like Taiw

Convex Low-resource Accent-Robust Language Detection in Speech Recognition

HardwareDGX agent

arXiv:2605.23235v1 Announce Type: new Abstract: Globalization and multiculturalism continue to produce increasingly diverse speech varieties. Yet current spoken dialogue systems frequently fail on und

Dell and Nvidia target the data chaos standing between AI pilots and real results

HardwareDGX agent

As organizations push to convert AI ambition into defensible business results, many are discovering that infrastructure alone is not enough. Instead, the real challenge lies in reconciling decades of

DFSAttn: Dynamic Fine-grained Sparse Attention for Efficient Video Generation

HardwareDGX agent

arXiv:2605.23445v1 Announce Type: new Abstract: Diffusion transformers have achieved remarkable success in high-quality video generation, yet their reliance on spatiotemporal 3D full attention incurs

Graph Learning via Logic-Based Weisfeiler-Leman Variants and Tabularization

HardwareDGX agent

arXiv:2508.10651v3 Announce Type: replace Abstract: We present a novel approach for graph classification based on tabularizing graph data via new variants of the Weisfeiler-Leman algorithm and then ap

If I wrote a short, cheap, electronic sequel to Taming Silicon Valley, perhaps with a title like “Getting to an AI that is good for all of h…

HardwareDGX agent

If I wrote a short, cheap, electronic sequel to Taming Silicon Valley, perhaps with a title like “Getting to an AI that is good for all of humanity” (perhaps donating proceeds) would you read it and g

MiniCPM5-1B is an impressive release in the 1B class! @OpenBMB https://huggingface.co/collections/openbmb/minicpm5 ✨ 1B - Apache 2.0 ✨ Hybri…

HardwareDGX agent

MiniCPM5-1B is an impressive release in the 1B class! @OpenBMB https://huggingface.co/collections/openbmb/minicpm5 ✨ 1B - Apache 2.0 ✨ Hybrid reasoning with Think / No-Think modes ✨ 128K context ✨ Run

PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion

HardwareDGX agent

arXiv:2605.23902v1 Announce Type: new Abstract: Most practical high-resolution text-to-image systems, including latent diffusion and autoregressive models, perform generation in a compact latent space

Sources: Meta, Google, and Amazon execs met Vatican officials on April 29, as part of a quiet lobbying push ahead of Pope Leo XIV's first AI encyclical (Océane Herrero/Politico)

HardwareDGX agent

Océane Herrero / Politico: Sources: Meta, Google, and Amazon execs met Vatican officials on April 29, as part of a quiet lobbying push ahead of Pope Leo XIV's first AI encyclical — As Leo XIV prepares

Taxes screw over both the companies and the states they’re in. Taxes didn’t build those big companies. They show up after the success. They …

HardwareDGX agent

Taxes screw over both the companies and the states they’re in. Taxes didn’t build those big companies. They show up after the success. They sure as hell didn’t create Silicon Valley. The sky-high taxe

Testing ZIT and Flux-1 with 'NVIDIA PiD — Pixel Diffusion Decoder'

HardwareDGX agent

NVIDIA's Pixel Diffusion Decoder (PiD) is an open-source decoder that replaces VAE decoders in image generation pipelines without retraining, producing sharper fine details and textures through a lear

← Previous
1…1516171819…29
Next →