AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
9 Aug 2026

Open Model: Google Weather Next 2

HardwareDGX agent

I am not a meteorologist, but I just read a very interesting article: https://arstechnica.com/science/2026/08/deepminds-hurricane-model-bought-forecasters-an-extra-day/ In a paper published on Thursda

7 Aug 2026

CUDA-L2: Surpassing cuBLAS Performance for Matrix Multiplication through Reinforcement Learning

HardwareDGX agent

arXiv:2512.02551v4 Announce Type: replace-cross Abstract: In this paper, we propose CUDA-L2, a system that combines large language models (LLMs) and reinforcement learning (RL) to automatically optimi

IcFuzz: Fuzzing Isaac Sim with Semantic Stage Guidance and Multi-level Mutation

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.06088v1 Announce Type: new Abstract: Robotics simulators serve as a foundational infrastructure for embodied AI, facilitating safe and scalable robotic system development. NVIDIA Isaac Sim

Learning to Rank Tensor Network Contraction Plans for GPU-Accelerated Quantum Circuit Simulation

HardwareDGX agent

arXiv:2608.05819v1 Announce Type: new Abstract: Classical simulation remains essential for developing and validating quantum algorithms, but its cost grows rapidly with circuit size. Tensor-network co

PaDoc: Layout-Grounded Parallel Decoding for Document Parsing

HardwareDGX agent

arXiv:2608.06146v1 Announce Type: new Abstract: End-to-end document parsers provide a unified interface, but serialize page layouts and regional contents into one autoregressive sequence. This formula

Parallel-in-Time Nonlinear Optimal Control via GPU-native Sequential Convex Programming

HardwareDGX agent

arXiv:2603.10711v3 Announce Type: replace Abstract: Real-time solution of nonlinear optimal control problems remains challenging on embedded robotic hardware, where conventional solvers often rely on

Sources: Nvidia agrees to invest 2B in Lancium, the power infrastructure developer of the Stargate campus in Texas, plus 1B more if it hits certain thresholds (The Information)

HardwareDGX agent

The Information: Sources: Nvidia agrees to invest 2B in Lancium, the power infrastructure developer of the Stargate campus in Texas, plus 1B more if it hits certain thresholds — Nvidia has agreed to i

Sources: the US Commerce Department's BIS is reviewing how Chinese AI companies access Nvidia chips overseas, including by legally renting foreign data centers (Mackenzie Hawkins/Bloomberg)

HardwareDGX agent

Mackenzie Hawkins / Bloomberg: Sources: the US Commerce Department's BIS is reviewing how Chinese AI companies access Nvidia chips overseas, including by legally renting foreign data centers — A key U

6 Aug 2026

AI-based single-shot structured-light depth reconstruction for real-time laparoscopic surgical guidance

HardwareDGX agent

arXiv:2608.05109v1 Announce Type: cross Abstract: Significance. Accurate intraoperative depth perception is important for autonomous and semi-autonomous robotic laparoscopic surgery. Conventional frin

AMD acquires Taalas to hardwire AI models into silicon

HardwareDGX agent

Advanced Micro Devices Inc. said today it has agreed to buy Taalas Inc., a Toronto startup that hardwires artificial intelligence models directly into silicon, in a deal that pushes the chipmaker deep

AMD acquires Toronto-based Taalas, which integrates model weights directly into silicon to boost inference performance, for an undisclosed sum (Tobias Mann/The Register)

HardwareDGX agent

Tobias Mann / The Register: AMD acquires Toronto-based Taalas, which integrates model weights directly into silicon to boost inference performance, for an undisclosed sum — Early tech demos show model

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

HardwareDGX agent

arXiv:2608.04041v1 Announce Type: new Abstract: Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classific

ContextMaster: Interactive Multi-Shot Video Creation via Fixed-Budget Sparse Context Routing

HardwareDGX agent

arXiv:2608.04956v1 Announce Type: new Abstract: Recent video models increasingly support generation, reference conditioning, and editing within a single model, yet typically expose them as separate op

Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs (Anna Tong/Forbes)

HardwareDGX agent

Anna Tong / Forbes: Docs: US data labeling companies, like Surge AI and Mercor, that sell training datasets to US AI labs and the government are also selling them to Chinese labs — The same Silicon Va

GASP: GPU-Accelerated Safe Planner for Real-Time Collision-Aware Motion Generation with Latent Trajectory Sampling

HardwareDGX agent

arXiv:2608.04612v1 Announce Type: new Abstract: We present GASP, a GPU-Accelerated Safe Planner for real-time, collision-aware joint-space motion generation in known environments. GASP combines a clam

Gaussian-LIC2: LiDAR-Inertial-Camera Gaussian Splatting SLAM

HardwareDGX agent

arXiv:2507.04004v3 Announce Type: replace Abstract: This paper presents the first photo-realistic LiDAR-Inertial-Camera Gaussian Splatting SLAM system that simultaneously addresses visual quality, geo

Hugging Face Storage Buckets are now on http://Vast.ai Connect your HF Storage Bucket as a Cloud Connection in your Vast settings, and every…

HardwareDGX agent

Hugging Face Storage Buckets are now on http://Vast.ai Connect your HF Storage Bucket as a Cloud Connection in your Vast settings, and every GPU instance you rent can pull datasets and checkpoints str

Into the Omniverse: How Open World Models Push the Frontier of Physical AI

HardwareDGX agent

In July, NVIDIA joined more than 200 companies and organizations in signing “Open Weights and American AI Leadership,” an open letter arguing that AI leadership will be measured not by any single fron

Ooredoo, Nvidia, Nokia, and Indosat launch Zankore, Indonesia's first dedicated AI compute and neocloud platform; Ooredoo, which owns a 49% stake, commits $800M (Melissa Hancock/Fortune)

HardwareDGX agent

Melissa Hancock / Fortune: Ooredoo, Nvidia, Nokia, and Indosat launch Zankore, Indonesia's first dedicated AI compute and neocloud platform; Ooredoo, which owns a 49% stake, commits $800M — Ooredoo Gr

RORA: Realistic Object Reconstruction with Articulation

HardwareDGX agent

arXiv:2608.04842v1 Announce Type: new Abstract: Replicating real-world environments into simulation by realistic visual representation like NeRF and 3D Gaussian Splatting (3DGS) has emerged as an effe

SparseDitto: Customizing GPU Kernels for Different Sparsity Patterns with LLM-Based Agentic System

HardwareDGX agent

arXiv:2608.05033v1 Announce Type: cross Abstract: Sparse matrix kernels are fundamental to scientific computing, graph analytics, and machine learning. Their GPU performance depends strongly on the in

StaticSegFormer: An Efficient High-Performance Semantic Segmentation Based on Static Structured Pruning

HardwareDGX agent

arXiv:2608.04811v1 Announce Type: new Abstract: Structured pruning enhances the efficiency of deep neural networks (DNNs) by eliminating groups of parameters during inference. Previous methods mostly

Sydney-based AI data center company Firmus raised 2B in funding from Coatue, Nvidia, and others at a 10.5B post-money valuation, up from $5.5B in April (Nichiket Sunil/Reuters)

HardwareDGX agent

Nichiket Sunil / Reuters: Sydney-based AI data center company Firmus raised 2B in funding from Coatue, Nvidia, and others at a 10.5B post-money valuation, up from 5.5B in April — Firmus said on Friday

5 Aug 2026

A Deployment-Friendly Foundational Framework for Efficient Computational Pathology

HardwareDGX agent

arXiv:2602.14010v2 Announce Type: replace-cross Abstract: Pathology foundation models (PFMs) generalize well across computational pathology tasks but remain costly for gigapixel whole-slide image anal

Accelerating Dynamic Graph Clustering on GPU Architectures with cuGraph

HardwareDGX agent

arXiv:2608.03695v1 Announce Type: cross Abstract: This work addresses community detection in temporal networks through GPU-accelerated extensions of spectral clustering and modularity-based algorithms

AcceptMoE: Commitment-Weighted Self-Sizing Verifier Expert Sets for Efficient MoE Speculative Decoding

HardwareDGX agent

arXiv:2608.02989v1 Announce Type: cross Abstract: Speculative decoding verifies a tree of draft tokens in one target-model forward pass. For a mixture-of-experts (MoE) target, however, parallel verifi

Compound and Parallel Modes of Tropical Convolutional Neural Networks

HardwareDGX agent

arXiv:2504.06881v2 Announce Type: replace-cross Abstract: Convolutional neural networks (CNNs) are foundational to many state-of-the-art computer vision systems, yet their reliance on multiplication-i

Learning and Clustering on Temporal Graphs: Principles, Primitives, and Pooling

HardwareDGX agent

arXiv:2608.03696v1 Announce Type: new Abstract: This work focuses on the problem of learning on temporal graphs, with particular emphasis on the task of clustering: obtaining coarse-grained representa

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

HardwareDGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs

HardwareDGX agent

arXiv:2608.03036v1 Announce Type: cross Abstract: Large Language Models (LLMs) are integrated into software systems and AI services, making efficient LLM serving a concern for software engineering. Se

RTX 3090 MiniMax H3 Speed Comparison: FP8 Scaled vs INT8 ConvRot (W8A8)

HardwareDGX agent

Setup: GPU: RTX 3090 24GB RAM: 32GB ComfyUI 0.30.0 PyTorch 2.13.0+cu130 CUDA 13.0 SageAttention enabled (sageattention-2.2.0+cu130torch2.10.0andhigher.post6-cp310-abi3-win_amd64) Spectrum node + Euler

4 Aug 2026

A Fortran General-Purpose Transpiler: Proof of Concept

HardwareDGX agent

arXiv:2608.00130v1 Announce Type: cross Abstract: Fortran has been the cornerstone of high-performance computing for decades and remains unmatched in many domains. Yet the language faces an expertise

AI cloud startup Volta Infra raised 300M led by a16z and Altimeter at a 2.4B valuation and says it landed a $10B contract with an unnamed leading AI developer (Dina Bass/Bloomberg)

HardwareDGX agent

Dina Bass / Bloomberg: AI cloud startup Volta Infra raised 300M led by a16z and Altimeter at a 2.4B valuation and says it landed a 10B contract with an unnamed leading AI developer — Volta Infra Holdi

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory…

HardwareDGX agent

AI compute is going to orbit. 🚀 @SpaceX’s Starmind AI1 satellite compute payload is powered by NVIDIA Vera Rubin NVL72, bringing AI factory compute closer to the stars. The next chapter of AI infrastr

As AI Increases Demands on Memory, Storage Steps Up

HardwareDGX agent

Surging AI demands are driving the need for massive datasets and context windows that burst past the confines of system memory. But rising needs aren’t met by simply adding more storage capacity. What

Bole: Efficient Tree Speculation for Hybrid-Attention Language Models

HardwareDGX agent

arXiv:2608.01651v1 Announce Type: cross Abstract: Hybrid-attention large language models combine full attention with recurrent linear attention to reduce long-context inference costs, yet their autore

Choosing the right orchestration layer for your AI use cases

HardwareDGX agent

Compute scarcity is not the only struggle AI teams face. Optimal utilization is also key. Friction also emerges when they outgrow informal coordination methods such as shared spreadsheets, manual SSH

early sparks of rsi? Ali Taha @waterloo_intern explains how http://Z.ai's GLM-5.2 (@Zai_org) profiled its SGLang serving path and rewrote bo…

HardwareDGX agent

early sparks of rsi? Ali Taha @waterloo_intern explains how http://Z.ai's GLM-5.2 (@Zai_org) profiled its SGLang serving path and rewrote bottlenecked GPU kernels, (still lacks reliable judgment) grea

Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super

HardwareDGX agent

NVIDIA Alpamayo 2 Super is a publicly available 34‑billion‑parameter vision–language–action model that merges a 32‑B Cosmos 3 Super Reasoner with a 2‑B action‑expert diffusion network. It produces uni

Has anyone been working on a solid setup for DSV4F on x2+ R9700s?

HardwareDGX agent

I'm hoping that one of you guys has been working on an inference engine or has somehow found improvements to running DSV4F on RDNA4 multi-GPU setups. I am currently building a custom inference engine

HetRoute Heterogeneous and Cost-aware Collaborative Routing Framework for Distributed Edge MoE Inference

HardwareDGX agent

arXiv:2608.00577v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models have become a dominant architecture for large-scale AI services, yet deploying them over geo-distributed heterogeneous

MiniWorld: Democratizing the Training of Video World Models from Scratch

HardwareDGX agent

arXiv:2608.01127v1 Announce Type: new Abstract: Video world models predict future observations conditioned on historical observations and control signals, enabling long-horizon generation through auto

NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use

HardwareDGX agent

For robotaxis and other autonomous vehicles (AVs), the hardest problems aren’t the everyday scenarios. They’re the rare, complex situations that are difficult to anticipate and train for. Handling the

NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US

HardwareDGX agent

NVIDIA is participating in the U.S. National Science Foundation’s (NSF) State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanc

Nvidia open sources cuFile API, accelerating GPU read/write capability for high-speed storage

HardwareDGX agent

As artificial intelligence applications become ever hungrier for faster access to data, Nvidia Corp. today announced it is open-sourcing the application programming interface for its powerful cuFile v

Open Secure AI Alliance proposes SAFE guidelines as membership tops 120

HardwareDGX agent

The Open Secure AI Alliance today proposed a set of guidelines for reporting cybersecurity incidents involving artificial intelligence agents, one week after the group was formed. The proposal is call

Partial FC: Training 10 Million Identities on a Single Machine

HardwareDGX agent

arXiv:2010.05222v3 Announce Type: replace Abstract: Training face recognition models with millions of identities is challenging because classifier storage, logit memory, and computation grow linearly

Sources: Anthropic agreed to a $10B deal for computing capacity in Norway from Nvidia-backed AI cloud startup Volta Infra, which says the deal is for six years (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Anthropic agreed to a 10B deal for computing capacity in Norway from Nvidia-backed AI cloud startup Volta Infra, which says the deal is for six years — Anthropic PBC has struck a 1

SparseKAN: Compressing Kolmogorov--Arnold Networks Across Basis Functions, Neurons, and Bits

HardwareDGX agent

arXiv:2608.00859v1 Announce Type: new Abstract: Kolmogorov--Arnold Networks (KANs) replace scalar edge weights with learnable univariate functions parameterized by multiple basis coefficients. This in

Stipple: Real-Time Incremental Gaussian Splatting with Visual-Inertial Tracking

HardwareDGX agent

arXiv:2608.00931v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3DGS) provides efficient rendering of photo-realistic scenes, but its heavy preprocessing and training steps make it a poor fit

Toward Geometry-Scalable Whole-Body Touch for Humanoids: A 3D-Printed Conformal EIT Skin

HardwareDGX agent

arXiv:2608.02080v1 Announce Type: new Abstract: Whole-body tactile sensing is a prerequisite for humanoids that operate in contact-rich human environments, but conventional taxel arrays scale poorly w

3 Aug 2026

DeltaServe: Host-Agnostic Co-Serving of Inference and Fine-Tuning for LLMs

HardwareDGX agent

arXiv:2607.28848v1 Announce Type: cross Abstract: LLM serving systems are provisioned for peak load to meet strict latency targets, leaving substantial GPU compute idle whenever traffic falls below pe

Event-Based Upper-Body Humanoid Teleoperation Under Challenging Illumination

HardwareDGX agent

arXiv:2607.29227v1 Announce Type: new Abstract: We present a real-time upper-body human-to-humanoid motion imitation framework driven by neuromorphic event-based vision. This work addresses practical

GPU-Accelerated ANNS: Quantized for Speed, Built for Change

HardwareDGX agent

arXiv:2601.07048v5 Announce Type: replace-cross Abstract: Approximate nearest neighbor search (ANNS) is a core problem in machine learning and information retrieval applications. GPUs offer a promisin

How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure

HardwareDGX agent

A combined architecture using KAI Scheduler and vCluster enables multiple teams to run fully isolated Kubernetes tenant clusters on a shared GPU node, each with its own control plane, RBAC, CRDs, and

Implicit Machine Learning Force Fields Accelerate Molecular Dynamics Simulations

HardwareDGX agent

arXiv:2607.29158v1 Announce Type: cross Abstract: We introduce implicit machine learning force fields (I-MLFFs), which replace explicit stacks of neural network layers with self-consistent fixed-point

London-based AI chip startup Olix raised 312M led by Fundomo, with participation from Arm, at a 3.3B valuation, up from 1B+ after raising 220M in February (Tim Bradshaw/Financial Times)

HardwareDGX agent

Tim Bradshaw / Financial Times: London-based AI chip startup Olix raised 312M led by Fundomo, with participation from Arm, at a 3.3B valuation, up from 1B+ after raising 220M in February — British ent

NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage

HardwareDGX agent

NVIDIA’s Vera BlueField‑4 STX Storage Processor combines an 88‑core Armv9.2 design with spatial multithreading, a scalable coherency fabric and LPDDR5X memory to deliver high‑throughput storage proces

OsteoCAD: A Human-in-the-Loop Cloud-Edge Framework for Bone Tumor Segmentation

HardwareDGX agent

arXiv:2607.29266v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Deep Learning (DL) have notably advanced medical image analysis, yet many health- care organizations struggle to adop

Point2Radio: A Foundation Model for Cross-Scene Radio Fields from Material-Aware Point Clouds

HardwareDGX agent

arXiv:2607.28994v1 Announce Type: cross Abstract: High-fidelity radio fields are typically simulated for every scene--transmitter configuration or fitted separately to each scene, failing to exploit p

← Previous
1234…29
Next →