AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
Hardware

The Download: a nuclear landmark, and China eyes Nvidia chips

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Four nuclear reactors hit a big milestone in the US —Casey Cro

hardwaremit-tech-review
9 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

The Future of Meta Superintelligence: A 1 Year Progress Update

DGX agent

Meta has made significant progress on its superintelligence research agenda, which aims to develop advanced AI systems with broad capabilities. The update covers Meta's infrastructure investments, res

hardwaresemianalysis
9 Jul 2026
Hardware

The model has an advisor tool that natively escalates to a stronger model when needed. This model is hosted in the U.S. by Perplexity on Nvi…

DGX agent

The model has an advisor tool that natively escalates to a stronger model when needed. This model is hosted in the U.S. by Perplexity on Nvidia B200 GPUs. We will improve the model in research preview

hardwareperplexity--x
9 Jul 2026
Hardware

Towards Accurate and Fast Clinical Body Composition: A Resource-Efficient Hierarchical Segmentation Framework for Multi-Source CT

DGX agent

arXiv:2607.07177v1 Announce Type: cross Abstract: Background: Automated 3D segmentation of muscles and adipose tissue from CT is vital for body composition analysis, but multi-source data heterogeneit

hardwarearxiv-cs-cv
9 Jul 2026
Hardware

Trees from Marginals: Autoregressive drafting with factorized priors

DGX agent

arXiv:2607.06763v1 Announce Type: cross Abstract: Speculative decoding greatly increases the interactivity of autoregressive language models by trading off computation for extra tokens generated in a

hardwarearxiv-cs-cl
9 Jul 2026
Hardware

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min au…

DGX agent

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min audio input >Speaker labels + timestamps in one generation >Mul

hardwareclem-delangue--x
9 Jul 2026
Hardware

1/ GPU capacity is getting more distributed. Data often isn’t. Teams can find GPUs across clouds, Kubernetes, Slurm, or on-prem clusters, bu…

DGX agent

1/ GPU capacity is getting more distributed. Data often isn’t. Teams can find GPUs across clouds, Kubernetes, Slurm, or on-prem clusters, but then still have to move models, datasets, and checkpoints

hardwareclem-delangue--x
8 Jul 2026
Hardware

A look at Chinese lidar maker Hesai, blacklisted by the US DOD in 2024, as it expands in the US; Hesai says it has ~33% of the global automotive lidar market (CNBC)

DGX agent

CNBC: A look at Chinese lidar maker Hesai, blacklisted by the US DOD in 2024, as it expands in the US; Hesai says it has ~33% of the global automotive lidar market — Robots on the factory floor. Self-

hardwaretechmeme
8 Jul 2026
Hardware

Announcing our $130M Series A to build the Open Superintelligence Stack Led by Radical Ventures, with NVIDIA, Intel Capital, Dell Capital, a…

DGX agent

Announcing our $130M Series A to build the Open Superintelligence Stack Led by Radical Ventures, with NVIDIA, Intel Capital, Dell Capital, and existing investors Train, deploy, and continuously improv

hardwareclem-delangue--x
8 Jul 2026
Hardware

Anthropic 3Q26 Profit Over $1B: The Anthropic IPO Financials Sneak Peak

DGX agent

I cannot verify this information as accurate. The article claims Anthropic achieved over $1B in profit in Q3 2026, but I have no reliable data confirming Anthropic's actual financial performance, IPO

hardwaresemianalysis
8 Jul 2026
Hardware

Argentum targets the capital stack as the missing layer in AI infrastructure buildout

DGX agent

The AI infrastructure boom has trained the industry’s attention on silicon and power, but a more fundamental constraint within the capital stack is quietly throttling the speed of global data center d

hardwaresiliconangle
8 Jul 2026
Hardware

Congrats to prime intellect! Love partnering with them on LangChain labs work

DGX agent

Congrats to prime intellect! Love partnering with them on LangChain labs work Announcing our $130M Series A to build the Open Superintelligence Stack Led by Radical Ventures, with NVIDIA, Intel Capita

hardwareharrison-chase--x
8 Jul 2026
Hardware

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. T…

DGX agent

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. This is what happens when world-class AI infrastructure meets

hardwareelon-musk--x
8 Jul 2026
Hardware

DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression

DGX agent

arXiv:2607.06523v1 Announce Type: new Abstract: Long-context language model inference is increasingly limited by the memory bandwidth and capacity required to store key-value caches, yet existing comp

hardwarearxiv-cs-ai
8 Jul 2026
Hardware

Design-CP: Context Parallelism for Design of Protein Nanoparticles

DGX agent

arXiv:2607.05439v1 Announce Type: new Abstract: Many all-atom generative protein models can in principle design large multimeric complexes by jointly modelling all chains, but their quadratic token- a

hardwarearxiv-cs-lg
8 Jul 2026
Hardware

HJCD-IK: GPU-Accelerated Inverse Kinematics through Batched Hybrid Jacobian Coordinate Descent

DGX agent

arXiv:2510.07514v2 Announce Type: replace Abstract: Inverse Kinematics (IK) is a core problem in robotics, in which joint configurations are found to achieve a (collision-free) desired end-effector po

hardwarearxiv-cs-ro
8 Jul 2026
Hardware

Inference chip startup SambaNova valued at 11B in 1B funding round

DGX agent

Chip startup SambaNova Inc. today announced that it has raised 1 billion in funding at a 11 billion valuation. General Atlantic led the Series F round with contributions from more than a dozen others.

hardwaresiliconangle
8 Jul 2026
Hardware

InfluMatch: Frontier-Quality KOL Search at 4B-Model Cost

DGX agent

arXiv:2607.05968v1 Announce Type: cross Abstract: Matching influencers (KOLs) to free-form, multi-part Thai marketing criteria is today served either by keyword search over structured profiles, which

hardwarearxiv-cs-ai
8 Jul 2026
Model Releases

Is Your NPU Ready for LLMs? Dissecting the Hidden Efficiency Bottlenecks in Mobile LLM Inference

DGX agent

arXiv:2607.05475v1 Announce Type: cross Abstract: Deploying Large Language Models (LLMs) on mobile devices enhances privacy and reduces latency, but is severely bottlenecked by hardware inefficiency.

model-releasesarxiv-cs-ai
8 Jul 2026
Research

Leveraging Neural Graph Compilers in Machine Learning Research for Edge-Cloud Systems

DGX agent

arXiv:2504.20198v2 Announce Type: replace-cross Abstract: This work presents a comprehensive evaluation of neural network graph compilers across heterogeneous hardware platforms, addressing the critic

researcharxiv-cs-lg
8 Jul 2026
Hardware

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory

DGX agent

arXiv:2607.05511v1 Announce Type: new Abstract: Agentic video understanding equips models with long-term memory to autonomously process and respond to continuous, long-horizon multimodal streams. Howe

hardwarearxiv-cs-cv
8 Jul 2026
Hardware

Nvidia lost ~$1T in market value in less than two months, dropping 16% from its May all-time-high, trading at 18x forward earnings, its lowest level since 2019 (Bloomberg)

DGX agent

Bloomberg: Nvidia lost ~1T in market value in less than two months, dropping 16% from its May all-time-high, trading at 18x forward earnings, its lowest level since 2019 — After losing roughly 1 trill

hardwaretechmeme
8 Jul 2026
Hardware

Open, convenient and predictable: Introducing Provisioned Throughput

DGX agent

Provisioned Throughput gives you reserved inference capacity for frontier open models like MiniMax M3 and GLM-5.2. Token-based pricing, a 99% uptime SLA, and up to 90% lower cost than proprietary APIs

hardwaretogether-ai-blog
8 Jul 2026
Hardware

Sources: China plans to let some of its biggest AI companies buy a small number of Nvidia's H200 chips to offset a domestic computing shortage in recent months (Qianer Liu/The Information)

DGX agent

Qianer Liu / The Information: Sources: China plans to let some of its biggest AI companies buy a small number of Nvidia's H200 chips to offset a domestic computing shortage in recent months — China pl

hardwaretechmeme
8 Jul 2026
Hardware

Sources: Reno-based AI chip startup Positron is in talks to raise ~750M in two phases, at valuations of 3.5B in the first tranche and ~$5B in the second (Bloomberg)

DGX agent

Bloomberg: Sources: Reno-based AI chip startup Positron is in talks to raise ~750M in two phases, at valuations of 3.5B in the first tranche and ~5B in the second — AI chip startup Positron is in talk

hardwaretechmeme
8 Jul 2026
Hardware

SUSE, NVIDIA, and Vultr Simplify the Path from AI Pilots to Production

DGX agent

SUSE, NVIDIA, and Vultr have partnered to create a streamlined solution for deploying AI applications from pilot phase to production environments. The collaboration leverages SUSE's AI Factory framewo

hardwarevultr
8 Jul 2026
Hardware

Tensordyne targets AI inference market with logarithmic math and Juniper-derived rack architecture

DGX agent

The race to serve AI inference faster and cheaper is exposing the hard limits of conventional chip architecture. As demand for real-time AI responses accelerates, the industry’s standard response — st

hardwaresiliconangle
8 Jul 2026
Hardware

Tuning-Free Latent Diffusion Models for Ultrahigh-Resolution Image Editing

DGX agent

arXiv:2607.06136v1 Announce Type: new Abstract: Recent diffusion-based generative models have shown impressive performance in image generation and editing. However, due to memory limitations and the h

hardwarearxiv-cs-cv
8 Jul 2026
Hardware

UBEP: Re-architecting Expert Parallelism Communication Library for Production Superpods

DGX agent

arXiv:2607.06202v1 Announce Type: cross Abstract: The deployment of Mixture-of-Experts (MoE) models on production high-bandwidth superpods, such as NVIDIA's NVL72/576 and Huawei's CloudMatrix384, intr

hardwarearxiv-cs-ai
8 Jul 2026
Hardware

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a han…

DGX agent

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a handful of labs. RL changes who can build frontier AI and just wo

hardwareswyx--x
8 Jul 2026
Hardware

We're releasing ZML/LLMD, our homegrown LLM server built on top of our homegrown high performance heterogeneous inference stack. It ships wi…

DGX agent

We're releasing ZML/LLMD, our homegrown LLM server built on top of our homegrown high performance heterogeneous inference stack. It ships with 5 architectures out of the box: NVIDIA, AMD, Metal, Intel

hardwareyann-lecun--x
8 Jul 2026
Hardware

A Deep Learning-based surrogate model for Severe Accidents in nuclear reactors using ASTEC

DGX agent

arXiv:2607.04450v1 Announce Type: cross Abstract: Integral codes like the Accident Source Term Evaluation Code (ASTEC) are powerful tools to study the physics of Severe Accidents (SAs) in nuclear reac

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

AI debt financing is on an insane trajectory, per @SemiAnalysis_:

DGX agent

AI debt financing is on an insane trajectory, per @SemiAnalysis_: Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters Over 7T AI debt by 2029, There can be no N

hardwaregary-marcus--x
7 Jul 2026
Hardware

AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters

DGX agent

Max single-threaded CPUs at scale are a new category of CPUs built for the agentic AI era. Across the creation and deployment of an agentic system, the CPU is on the critical path for reasoning, respo

hardwarenvidia-blog
7 Jul 2026
Hardware

AquaGen: Scaling generative models to molecular dynamics precision on thousands of atoms

DGX agent

arXiv:2607.03513v1 Announce Type: cross Abstract: We present AquaGen, the first all-atom, explicit solvent, periodic-boundary-condition-aware generative model that produces molecular configurations fr

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

CPR: Chained Perceptual Refinement for Coarse-to-Fine Medical Image Classification

DGX agent

arXiv:2607.02591v1 Announce Type: new Abstract: High resolution medical images contain fine grained, spatially sparse cues that are critical for diagnosis, yet preserving full resolution incurs substa

hardwarearxiv-cs-cv
7 Jul 2026
Hardware

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

DGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Fast Asymptotically Optimal Kinodynamic Planning via Vectorization

DGX agent

arXiv:2607.03987v1 Announce Type: new Abstract: Sampling-based motion planners have been shown to be effective for systems with complex kinodynamic constraints and high dimensionality. However, these

hardwarearxiv-cs-ro
7 Jul 2026
Model Releases

From Arithmetic to Logic: The Resilience of Logic and Lookup-Based Neural Networks Under Parameter Bit-Flips

DGX agent

arXiv:2603.22770v2 Announce Type: replace-cross Abstract: The deployment of deep neural networks (DNNs) in safety-critical edge environments necessitates robustness against hardware-induced bit-flip e

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

Graph Sparse Sampling: Breaking the Curse of the Horizon in Continuous MDP Planning

DGX agent

arXiv:2607.05359v1 Announce Type: new Abstract: Planning under uncertainty in continuous domains is essential for autonomous systems, yet computationally demanding. Tree-based search methods such as M

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

Lights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy Consumption

DGX agent

arXiv:2607.04553v1 Announce Type: cross Abstract: We present a bidirectional framework for estimating the energy consumption of text-to-video (T2V) and text-to-video-audio (T2VA) models from architect

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community

DGX agent

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development can still be gated b

hardwarenvidia-blog
7 Jul 2026
Hardware

NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads

DGX agent

NVIDIA Vera is a CPU designed to help AI factories scale agentic AI and reinforcement learning by shortening CPU execution time, increasing task throughput, and enabling smarter, longer-thinking agent

hardwarenvidia-developer
7 Jul 2026
Hardware

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

DGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

hardwarearxiv-cs-ai
7 Jul 2026
Hardware

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization

DGX agent

arXiv:2607.02834v1 Announce Type: new Abstract: Molecular optimization often starts from a pretrained generative model that captures a broad prior over valid molecular structures. At test time, howeve

hardwarearxiv-cs-lg
7 Jul 2026
Hardware

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the lates…

DGX agent

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the latest Isaac models and frameworks to the open robotics community.

hardwareclem-delangue--x
7 Jul 2026
Hardware

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

DGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

DGX agent

arXiv:2504.08791v3 Announce Type: replace-cross Abstract: On-device inference offers privacy, offline use, and instant response, but consumer hardware restricts large language models (LLMs) to low thr

model-releasesarxiv-cs-ai
7 Jul 2026
← Previous
1…1819202122…93
Next →