AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
10 Jun 2026

Flash-GMM: A Memory-Efficient Kernel for Scalable Soft Clustering

HardwareDGX agent

arXiv:2606.10896v1 Announce Type: new Abstract: We present extbf{Flash-GMM}, a fused Triton kernel for efficient computation of Gaussian Mixture Models (GMMs) over large-scale data in a single GPU pas

How Torc hit 90% GPU utilization and other stories on scaling AI with Ray from Discord, Cubist, and Coinbase

HardwareDGX agent

This article recaps presentations from Ray Day NYC featuring case studies from Discord, Cubist, and Coinbase on scaling AI workloads using the Ray distributed computing framework, with a focus on Torc

In good faith and with no judgment (mistakes happen), I truly hope that Anthropic will hear the feedback and change course on this. Anthropi…

HardwareDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

In good faith and with no judgment (mistakes happen), I truly hope that Anthropic will hear the feedback and change course on this. Anthropic is a company that has been raising awareness about AI mani

Neura Robotics to raise up to $1.4B from Nvidia-backed consortium

HardwareDGX agent

Automation startup Neura Robotics GmbH is raising a funding round worth up to 1.4 billion from a group of prominent investors. The company detailed today that the consortium includes Amazon.com Inc.,

One-Step Residual Shifting Diffusion for Image Super-Resolution via Distillation

HardwareDGX agent

arXiv:2503.13358v5 Announce Type: replace Abstract: Diffusion models for super-resolution (SR) produce high-quality visual results but require expensive computational costs. Despite the development of

OpenRTLSet: A Fully Open-Source Dataset for Large Language Model-based Verilog Module Design

Model ReleasesDGX agent

arXiv:2606.10285v1 Announce Type: new Abstract: OpenRTLSet introduces the largest fully open-source dataset for hardware design, offering over 131,000 diverse Verilog code samples to the research comm

Run DiffusionGemma on NVIDIA for Developer-Ready, High-Throughput Text Generation

HardwareDGX agent

DiffusionGemma is an experimental open model built for exceptionally fast text generation that NVIDIA has optimized to run on GeForce RTX GPUs, RTX PRO, and DGX Spark systems. Rather than generating t

Sampling Triangulations and Calabi-Yau Threefolds with Autoregressive GNNs

HardwareDGX agent

arXiv:2605.27770v2 Announce Type: cross Abstract: We introduce `dualGNN', an autoregressive message-passing GNN for sampling fine, regular triangulations (FRTs) of convex polytopes. dualGNN operates o

Sources: Germany's Neura Robotics, which builds AI-powered humanoid robots, raised 1.4B from Tether, Qualcomm, Amazon, Nvidia, and others at a ~7B valuation (Financial Times)

HardwareDGX agent

Financial Times: Sources: Germany's Neura Robotics, which builds AI-powered humanoid robots, raised 1.4B from Tether, Qualcomm, Amazon, Nvidia, and others at a ~7B valuation — Crypto group Tether, Ama

Sources: OpenAI is in advanced talks to lease a proposed 10GW data center campus in Ohio as part of a deal that could include financial backing from Nvidia (Anissa Gardizy/The Information)

HardwareDGX agent

Anissa Gardizy / The Information: Sources: OpenAI is in advanced talks to lease a proposed 10GW data center campus in Ohio as part of a deal that could include financial backing from Nvidia — OpenAI i

SpenseGPT: Practical One-shot Pruning Enabling Sparse and Dense GEMMs for LLM Inference

HardwareDGX agent

arXiv:2606.10445v1 Announce Type: cross Abstract: Semi-structured 2:4 sparsity is widely supported by modern accelerators, providing up to a 2x theoretical speedup. However, its strict 50% sparsity co

torch-sla: Differentiable Sparse Linear Algebra with Adjoint Solvers and Sparse Tensor Parallelism for PyTorch

HardwareDGX agent

arXiv:2601.13994v3 Announce Type: replace-cross Abstract: Differentiable sparse linear algebra is foundational for scientific machine learning, yet PyTorch lacks a unified library for it: torch.sparse

Unifying Data, Memory, and Compute Efficiency in LLM training: A Survey

HardwareDGX agent

arXiv:2606.10706v1 Announce Type: cross Abstract: Resource constraints increasingly determine what can be trained, fine-tuned, and deployed in large language models (LLMs), yet efficiency is often stu

Usage share of OpenAI grew vs Anthropic yesterday despite Mythos 5 / Fable 5 launch Multiple power users at SemiAnalysis tried Mythos / Fabl…

HardwareDGX agent

Usage share of OpenAI grew vs Anthropic yesterday despite Mythos 5 / Fable 5 launch Multiple power users at SemiAnalysis tried Mythos / Fable Got refusals for nonsensical reasons Got pissed off at Ant

9 Jun 2026

Accelerating Birkhoff Projection for Manifold-Constrained Hyper-Connections

HardwareDGX agent

arXiv:2606.07574v1 Announce Type: cross Abstract: Manifold-constrained hyper-connections (mHCs) have recently been proposed as a principled extension of hyper-connections, where the residual mixing ma

Accelerating Federated Learning Research with AI Agents and NVIDIA FLARE Auto-FL

HardwareDGX agent

NVIDIA FLARE Auto-FL automates federated learning research by constraining agent actions through a control plane, enforcing fixed benchmark contracts, and using an experiment ledger to ensure reproduc

Aqua Boundary-Saliency Attention Module for Lightweight Underwater Salient Instance Segmentation Detection Transformer

HardwareDGX agent

arXiv:2606.08002v1 Announce Type: new Abstract: Underwater instance segmentation integrates pixel-level mask prediction and instance-level discrimination for marine resource exploration, ecological mo

BRAIN: Bayesian Reasoning via Active Inference for Agentic and Embodied Intelligence in Mobile Networks

HardwareDGX agent

arXiv:2602.14033v1 Announce Type: cross Abstract: Future sixth-generation (6G) mobile networks will demand artificial intelligence (AI) agents that are not only autonomous and efficient, but also capa

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly …

HardwareDGX agent

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly degrade its IQ so that the average engineer won't notice. We

FOCUS specification eyes AI token economics as AI billing complexity hits a new frontier

HardwareDGX agent

AI token economics is creating a data normalization crisis across the technology stack as enterprises scramble to apply consistent financial controls to GPU clusters, AI factories and token-based cons

for those keeping track at home it was 34 days between signing this deal and launching Mythos-class model GA to the world. https://x.com/lee…

HardwareDGX agent

for those keeping track at home it was 34 days between signing this deal and launching Mythos-class model GA to the world. https://x.com/leerob/status/2052059466821198061?s=20 building on @nvidia stac

FuseFSS: Efficient Secure LLM Inference with Function Secret Sharing

HardwareDGX agent

arXiv:2606.09551v1 Announce Type: cross Abstract: Two-server secure inference allows a client to query a hosted large language model (LLM) without revealing prompts or embeddings. Recent GPU systems b

@GaryMarcus Hey Gary, I'm reading one of your books right now, ‘Taming Silicon Valley’. I have to say, it's really opened my eyes to what's …

HardwareDGX agent

@GaryMarcus Hey Gary, I'm reading one of your books right now, ‘Taming Silicon Valley’. I have to say, it's really opened my eyes to what's going on out there in the AI industry! Looking forward to yo

GPU MODE has powered much of the public GPU kernel work online, with a permissive license from day one and generous credit from researchers,…

HardwareDGX agent

GPU MODE has powered much of the public GPU kernel work online, with a permissive license from day one and generous credit from researchers, NVIDIA, AMD, and others. Today we’re moving our datasets to

Haven't seen much about the Nvidia Cosmos 3 video model that dropped, what's up with that?

HardwareDGX agent

Nvidia Cosmos 3 is an open physical AI foundation model built on a mixture-of-transformers architecture that combines vision reasoning, world generation, and action prediction for reasoning, simulatio

Hybridizing Equilibrium Propagation with Ising Machines for Efficient Energy-Based Learning

HardwareDGX agent

arXiv:2606.09112v1 Announce Type: cross Abstract: The rapid evolution of artificial intelligence has led to substantial advances in deep neural networks. Nonetheless, conventional GPU-based training r

Kunlun: Establishing Scaling Laws for Massive-Scale Recommendation Systems through Unified Architecture Design

HardwareDGX agent

arXiv:2602.10016v3 Announce Type: replace-cross Abstract: Deriving predictable scaling laws that govern the relationship between model performance and computational investment is crucial for designing

Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT

HardwareDGX agent

arXiv:2601.20408v2 Announce Type: replace-cross Abstract: Enterprise LLM deployment faces a critical scalability challenge: organizations must optimize models systematically to scale AI initiatives wi

MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning

HardwareDGX agent

arXiv:2606.08039v1 Announce Type: new Abstract: Robotic simulators are a cornerstone of modern research in aerial robotics, serving both as a vehicle for the development of new control algorithms and

NVIDIA Confidential Computing to Help Expand Apple’s Private Cloud Compute

HardwareDGX agent

NVIDIA GPUs with Confidential Computing are now used for confidential inference in Apple’s Private Cloud Compute (PCC), as it expands beyond Apple’s data centers to Google Cloud. Unveiled during Apple

Report: GKE Inference Gateway delivers up to 92% faster AI responses

Model ReleasesDGX agent

As generative AI moves from experimental pilots to massive production environments, the efficiency of your infrastructure becomes the ultimate differentiator. One way to get the most out of it and min

Resource-aware Computation-Communication Overlap for multi-GPU ML Workloads

HardwareDGX agent

arXiv:2606.09200v1 Announce Type: cross Abstract: The rapid growth of large-scale machine learning (ML) has made distributed training across multiple GPUs a fundamental component of modern ML systems.

RTL-BenchLS: A Large-Scale Benchmark for RTL Reasoning and Generation with Large Language Models

Model ReleasesDGX agent

arXiv:2606.08976v1 Announce Type: new Abstract: LLM-based RTL generation and reasoning is a promising direction for hardware design automation. High-quality benchmarks are critical infrastructure for

Scale Robot Reinforcement Learning with NVIDIA Isaac Lab on Amazon SageMaker AI

HardwareDGX agent

In this post, we show how to train robot policies for the Unitree H1 humanoid with NVIDIA Isaac Lab on Amazon SageMaker AI across two compute options: Amazon SageMaker HyperPod and Amazon SageMaker Tr

SoccerNet 2026 Player-Centric Ball-Action Spotting:Retraining and Post-Processing Extensions to the FOOTPASS Baselines

HardwareDGX agent

arXiv:2606.09679v1 Announce Type: new Abstract: We describe our system for the SoccerNet 2026 Player-Centric Ball-Action Spotting Challenge, which requires predicting who performs which action and whe

STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank Control

HardwareDGX agent

arXiv:2606.08382v1 Announce Type: cross Abstract: Low-rank projection has emerged as a promising approach for compressing the KV cache by exploiting hidden-dimension redundancy. However, prior methods

TUDSR: Twice Upsampling-Diffusion for Higher Super-Resolution

HardwareDGX agent

arXiv:2606.09608v1 Announce Type: new Abstract: Diffusion-based generative models have achieved remarkable success in real-world image super-resolution (SR). With tiled diffusion techniques, these mod

Ultra Flash: Scaling Real-Time Streaming Video Generation to High Resolutions

HardwareDGX agent

arXiv:2606.09150v1 Announce Type: new Abstract: While recent autoregressive video diffusion models achieve remarkable streaming quality, they remain confined to low resolutions (e.g., 480P), leaving e

What EU regulations does to AI

HardwareDGX agent

EU regulations, particularly the AI Act, establish comprehensive compliance requirements for AI systems including risk-based classification, transparency obligations, and restrictions on high-risk app

8 Jun 2026

Ablation Study of Block Size, Weight Precision, and Scale Precision in NVFP4 Inference for Low-Power Edge-Efficient Neural Networks

ResearchDGX agent

arXiv:2606.06527v1 Announce Type: cross Abstract: Energy-efficient edge inference requires reducing arithmetic cost, memory traffic, and hardware overhead. This paper presents an ablation-focused stud

Accelerated Fourier SAT (AFSAT): Fully Realising a GPU-based Symmetric Pseudo-Boolean SAT Solver

HardwareDGX agent

arXiv:2606.06641v1 Announce Type: new Abstract: We present Accelerated Fourier SAT (AFSAT), a GPU-accelerated solver for pseudo-Boolean satisfiability based on continuous local search (CLS). AFSAT rea

b9556

Local AiDGX agent

llama.cpp b9556 release adds support for AMD RDNA3.5 graphics hardware (gfx1152 and gfx1153) in its HIP backend . The release includes compiled binaries for multiple platforms including macOS, Linux,

Broadband Hyperspectral 3D Imaging using Dispersed Structured Light

HardwareDGX agent

arXiv:2605.25757v2 Announce Type: replace Abstract: Hyperspectral 3D imaging enables the capture of dense spectral information and scene geometry but has traditionally been confined to narrow spectral

China's Unitree Will Dominate Global Robotics

HardwareDGX agent

Unitree, a Chinese robotics company, is positioned to become a dominant force in global robotics due to its advanced quadruped and humanoid robot technologies, competitive pricing, and rapid innovatio

China's Unitree Will Dominate Global Robotics The Fastest Iteration Cycle In Next-Gen Robotics Should See Unprecedented Acceleration https:/…

HardwareDGX agent

China's Unitree Will Dominate Global Robotics The Fastest Iteration Cycle In Next-Gen Robotics Should See Unprecedented Acceleration https://newsletter.semianalysis.com/p/chinas-unitree-will-dominate-

Good take My guess is - demand for intelligence is near infinite - but 80% of workloads will be running on 99% cheaper models within 12-18 m…

HardwareDGX agent

Good take My guess is - demand for intelligence is near infinite - but 80% of workloads will be running on 99% cheaper models within 12-18 months - 20% of workloads will still run on latest gen models

How the UK Is Turning Sovereign AI Ambition Into Action With NVIDIA Technologies

HardwareDGX agent

A year ago at London Tech Week, NVIDIA founder and CEO Jensen Huang and U.K. Prime Minister Keir Starmer made a declaration: the U.K. would be an AI maker, not an AI taker. At this year’s event, NVIDI

I'll never forget @GaryMarcus' comment at last autumn's @MediaTechDem conference in MTL - something to the effect of 'what if Nvidia are jus…

HardwareDGX agent

I'll never forget @GaryMarcus' comment at last autumn's @MediaTechDem conference in MTL - something to the effect of 'what if Nvidia are just the people who sold shovels in the gold rush'? Striking pa

LuMamba: Latent Unified Mamba for Electrode Topology-Invariant and Efficient EEG Modeling

HardwareDGX agent

arXiv:2603.19100v2 Announce Type: replace Abstract: Electroencephalography (EEG) enables non-invasive monitoring of brain activity across clinical and neurotechnology applications, yet building founda

macOS 27 requires Apple Silicon, as Apple draws down the Intel Mac era

HardwareDGX agent

macOS 27 will be compatible with Apple silicon Macs only, requiring an M-series chip or newer. macOS 26 Tahoe is the final major macOS version for Intel-based Macs. The incompatibility stems from the

@METR_Evals previously on @cognition_labs https://x.com/swyx/status/2062611218196771017?s=20

HardwareDGX agent

@METR_Evals previously on @cognition_labs https://x.com/swyx/status/2062611218196771017?s=20 Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog

MVSegNet: A Lightweight Boundary-Aware Network for Fetal Lateral Ventricle Segmentation and Atrial Width Estimation in Prenatal Ultrasound

HardwareDGX agent

arXiv:2606.06958v1 Announce Type: new Abstract: Fetal ventriculomegaly is assessed by measuring the atrial width of the lateral ventricle in prenatal ultrasound. Accurate segmentation is essential for

Nvidia and LG expand their partnership in AI, robotics, and mobility tech, including building an AI factory and co-designing future data center architectures (Choi Eun-Kyung/The Chosun Daily)

HardwareDGX agent

Choi Eun-Kyung / The Chosun Daily: Nvidia and LG expand their partnership in AI, robotics, and mobility tech, including building an AI factory and co-designing future data center architectures — Jense

NVIDIA and LG Group Build an AI Factory to Advance Physical AI, Mobility and AI Infrastructure

HardwareDGX agent

NVIDIA and LG Group are building an AI factory to accelerate LG Group’s next wave of AI-driven businesses, spanning robotics, autonomous driving, data center technologies and GPU cloud services. The A

Nvidia partners with South Korea’s SK hynix, Naver and Doosan to expand the country’s AI infrastructure

HardwareDGX agent

Artificial intelligence chipmaker Nvidia Corp. today announced a slate of new partnerships with some of South Korea’s biggest technology, including the memory chip supplier SK hynix Inc., internet gia

Predictive Style Matching: Natural and Robust Humanoid Locomotion

ResearchDGX agent

arXiv:2606.07083v1 Announce Type: new Abstract: Reinforcement learning has become the prevailing approach to humanoid locomotion control: policies transfer reliably from simulation to hardware and rec

RhinoVLA Technical Report

Model ReleasesDGX agent

arXiv:2606.07383v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for robotic manipulation, but real-time deployment on edge hardware remains challengin

Scalable Joint Resource Allocation for SLO-Constrained LLM Inference in Heterogeneous GPU Clouds

HardwareDGX agent

arXiv:2604.07472v2 Announce Type: replace Abstract: Serving large language model (LLM) inference in cloud environments requires jointly optimizing model selection, GPU provisioning, parallelism config

TorchKM: A GPU-Oriented Library for Kernel Learning and Model Selection

HardwareDGX agent

arXiv:2606.06742v1 Announce Type: new Abstract: TorchKM is an open-source library for kernel machines, including support vector machines, kernel logistic regression, and kernel quantile regression, wi

Train Models Faster with JAX and MaxText Using NVFP4 on NVIDIA Blackwell

HardwareDGX agent

This NVIDIA technical article addresses how low-bit mixed-precision pre-training accelerates large language model training, where optimizing numerical precision can reduce training time across distrib

← Previous
1…1920212223…75
Next →