AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
9 Jul 2026

A Practical Guide to GPU-Initiated Communication for Molecular Dynamics at Scale

HardwareDGX agent

This guide covers GPU-initiated communication techniques using NVSHMEM, which allows GPUs to directly communicate with other GPU memory spaces, leveraging global address space for higher bandwidth and

AI agent startup Lyzr reportedly raising 100M at 500M valuation

HardwareDGX agent

Lyzr Inc., a startup that helps enterprises build artificial intelligence agents, is reportedly raising a funding round worth about 100 million. Bloomberg today cited sources as saying that the deal h

compute daddy @dylan522p has spoken

HardwareDGX agent

compute daddy @dylan522p has spoken The Future of Meta Superintelligence: A 1 Year Progress Update A top tier RL environment startup spawns out of thin air, the most aggressive compute ramp we've ever

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DDN targets GPU efficiency with AI data infrastructure as the make-or-break layer

HardwareDGX agent

The race to build AI factories is well underway, and the winning organizations have learned that AI data infrastructure determines whether or not GPU investments pay off, while others are still scramb

DYNA-PRUNER: Input-Adaptive Data-Model Co-Pruning for Efficient and Scalable Spatio-Temporal Media Prediction

HardwareDGX agent

arXiv:2606.15346v2 Announce Type: replace Abstract: Spatio-temporal prediction supports radar/satellite nowcasting and city-scale traffic monitoring, but modern models are often too expensive for real

Efficient Long-Horizon Learning for Learned Optimization

HardwareDGX agent

arXiv:2607.06772v1 Announce Type: new Abstract: Learned optimization aims to improve upon hand-designed optimizers (e.g., Adam and Muon) by meta-learning small neural network optimizers over a distrib

GeForce NOW Turns Up the Heat With New GeForce RTX 5080-Powered Toronto Server

HardwareDGX agent

This GFN Thursday brings more games, more power and more ways to play on GeForce NOW. The cloud gaming service is expanding with a new GeForce RTX 5080-powered server in Toronto, bringing dedicated hi

Infinite Worlds with Versatile Interactions

HardwareDGX agent

arXiv:2607.07534v1 Announce Type: new Abstract: We present LingBot-World 2.0 (also known as LingBot-World-Infinity), an advanced iteration of LingBot-World featuring four distinct upgrades. (1) Our mo

MiLSD: A Micro Line-Segment Detector for Resource-Constrained Devices

HardwareDGX agent

arXiv:2607.06600v1 Announce Type: cross Abstract: Line segment detection is a key building block in visual SLAM, 3D reconstruction, and industrial inspection. Recent deep learning methods have greatly

Prime Intellect raises 130M at 1B valuation for its AI training platform

HardwareDGX agent

Artificial intelligence training startup Prime Intellect Inc. has raised 130 million in funding from a group of prominent investors. The consortium included Nvidia Corp.’s NVentures, Intel Capital and

Shanghai-based GPU maker Iluvatar CoreX raised ~902M in a Hong Kong share sale; the company's stock has soared 257% since its January IPO, which raised 473M (Bloomberg)

HardwareDGX agent

Bloomberg: Shanghai-based GPU maker Iluvatar CoreX raised ~902M in a Hong Kong share sale; the company's stock has soared 257% since its January IPO, which raised 473M — Shanghai Iluvatar CoreX Semico

Synthetic Data Generation for Financial AI Research with NVIDIA NeMo

HardwareDGX agent

NVIDIA's NeMo framework addresses the challenge of limited, imbalanced financial NLP data by generating synthetic financial news headlines to fill gaps for trading research, risk modeling, and surveil

The Download: a nuclear landmark, and China eyes Nvidia chips

HardwareDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Four nuclear reactors hit a big milestone in the US —Casey Cro

The Future of Meta Superintelligence: A 1 Year Progress Update

HardwareDGX agent

Meta has made significant progress on its superintelligence research agenda, which aims to develop advanced AI systems with broad capabilities. The update covers Meta's infrastructure investments, res

The model has an advisor tool that natively escalates to a stronger model when needed. This model is hosted in the U.S. by Perplexity on Nvi…

HardwareDGX agent

The model has an advisor tool that natively escalates to a stronger model when needed. This model is hosted in the U.S. by Perplexity on Nvidia B200 GPUs. We will improve the model in research preview

Towards Accurate and Fast Clinical Body Composition: A Resource-Efficient Hierarchical Segmentation Framework for Multi-Source CT

HardwareDGX agent

arXiv:2607.07177v1 Announce Type: cross Abstract: Background: Automated 3D segmentation of muscles and adipose tissue from CT is vital for body composition analysis, but multi-source data heterogeneit

Trees from Marginals: Autoregressive drafting with factorized priors

HardwareDGX agent

arXiv:2607.06763v1 Announce Type: cross Abstract: Speculative decoding greatly increases the interactivity of autoregressive language models by trading off computation for extra tokens generated in a

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min au…

HardwareDGX agent

🤗 we just opensourced our tiny beast on @huggingface >0.9B ASR model >Apache license 2.0 >128k long-context transcription >Up to ~90-min audio input >Speaker labels + timestamps in one generation >Mul

8 Jul 2026

1/ GPU capacity is getting more distributed. Data often isn’t. Teams can find GPUs across clouds, Kubernetes, Slurm, or on-prem clusters, bu…

HardwareDGX agent

1/ GPU capacity is getting more distributed. Data often isn’t. Teams can find GPUs across clouds, Kubernetes, Slurm, or on-prem clusters, but then still have to move models, datasets, and checkpoints

A look at Chinese lidar maker Hesai, blacklisted by the US DOD in 2024, as it expands in the US; Hesai says it has ~33% of the global automotive lidar market (CNBC)

HardwareDGX agent

CNBC: A look at Chinese lidar maker Hesai, blacklisted by the US DOD in 2024, as it expands in the US; Hesai says it has ~33% of the global automotive lidar market — Robots on the factory floor. Self-

Announcing our $130M Series A to build the Open Superintelligence Stack Led by Radical Ventures, with NVIDIA, Intel Capital, Dell Capital, a…

HardwareDGX agent

Announcing our $130M Series A to build the Open Superintelligence Stack Led by Radical Ventures, with NVIDIA, Intel Capital, Dell Capital, and existing investors Train, deploy, and continuously improv

Anthropic 3Q26 Profit Over $1B: The Anthropic IPO Financials Sneak Peak

HardwareDGX agent

I cannot verify this information as accurate. The article claims Anthropic achieved over $1B in profit in Q3 2026, but I have no reliable data confirming Anthropic's actual financial performance, IPO

Argentum targets the capital stack as the missing layer in AI infrastructure buildout

HardwareDGX agent

The AI infrastructure boom has trained the industry’s attention on silicon and power, but a more fundamental constraint within the capital stack is quietly throttling the speed of global data center d

Congrats to prime intellect! Love partnering with them on LangChain labs work

HardwareDGX agent

Congrats to prime intellect! Love partnering with them on LangChain labs work Announcing our $130M Series A to build the Open Superintelligence Stack Led by Radical Ventures, with NVIDIA, Intel Capita

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. T…

HardwareDGX agent

Congrats to @SpaceXAI on Grok 4.5 — trained on NVIDIA GB300 NVL72 systems and purpose-built for coding, agentic tasks, and knowledge work. This is what happens when world-class AI infrastructure meets

DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression

HardwareDGX agent

arXiv:2607.06523v1 Announce Type: new Abstract: Long-context language model inference is increasingly limited by the memory bandwidth and capacity required to store key-value caches, yet existing comp

Design-CP: Context Parallelism for Design of Protein Nanoparticles

HardwareDGX agent

arXiv:2607.05439v1 Announce Type: new Abstract: Many all-atom generative protein models can in principle design large multimeric complexes by jointly modelling all chains, but their quadratic token- a

HJCD-IK: GPU-Accelerated Inverse Kinematics through Batched Hybrid Jacobian Coordinate Descent

HardwareDGX agent

arXiv:2510.07514v2 Announce Type: replace Abstract: Inverse Kinematics (IK) is a core problem in robotics, in which joint configurations are found to achieve a (collision-free) desired end-effector po

Inference chip startup SambaNova valued at 11B in 1B funding round

HardwareDGX agent

Chip startup SambaNova Inc. today announced that it has raised 1 billion in funding at a 11 billion valuation. General Atlantic led the Series F round with contributions from more than a dozen others.

InfluMatch: Frontier-Quality KOL Search at 4B-Model Cost

HardwareDGX agent

arXiv:2607.05968v1 Announce Type: cross Abstract: Matching influencers (KOLs) to free-form, multi-part Thai marketing criteria is today served either by keyword search over structured profiles, which

Is Your NPU Ready for LLMs? Dissecting the Hidden Efficiency Bottlenecks in Mobile LLM Inference

Model ReleasesDGX agent

arXiv:2607.05475v1 Announce Type: cross Abstract: Deploying Large Language Models (LLMs) on mobile devices enhances privacy and reduces latency, but is severely bottlenecked by hardware inefficiency.

Leveraging Neural Graph Compilers in Machine Learning Research for Edge-Cloud Systems

ResearchDGX agent

arXiv:2504.20198v2 Announce Type: replace-cross Abstract: This work presents a comprehensive evaluation of neural network graph compilers across heterogeneous hardware platforms, addressing the critic

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory

HardwareDGX agent

arXiv:2607.05511v1 Announce Type: new Abstract: Agentic video understanding equips models with long-term memory to autonomously process and respond to continuous, long-horizon multimodal streams. Howe

Nvidia lost ~$1T in market value in less than two months, dropping 16% from its May all-time-high, trading at 18x forward earnings, its lowest level since 2019 (Bloomberg)

HardwareDGX agent

Bloomberg: Nvidia lost ~1T in market value in less than two months, dropping 16% from its May all-time-high, trading at 18x forward earnings, its lowest level since 2019 — After losing roughly 1 trill

Open, convenient and predictable: Introducing Provisioned Throughput

HardwareDGX agent

Provisioned Throughput gives you reserved inference capacity for frontier open models like MiniMax M3 and GLM-5.2. Token-based pricing, a 99% uptime SLA, and up to 90% lower cost than proprietary APIs

Sources: China plans to let some of its biggest AI companies buy a small number of Nvidia's H200 chips to offset a domestic computing shortage in recent months (Qianer Liu/The Information)

HardwareDGX agent

Qianer Liu / The Information: Sources: China plans to let some of its biggest AI companies buy a small number of Nvidia's H200 chips to offset a domestic computing shortage in recent months — China pl

Sources: Reno-based AI chip startup Positron is in talks to raise ~750M in two phases, at valuations of 3.5B in the first tranche and ~$5B in the second (Bloomberg)

HardwareDGX agent

Bloomberg: Sources: Reno-based AI chip startup Positron is in talks to raise ~750M in two phases, at valuations of 3.5B in the first tranche and ~5B in the second — AI chip startup Positron is in talk

SUSE, NVIDIA, and Vultr Simplify the Path from AI Pilots to Production

HardwareDGX agent

SUSE, NVIDIA, and Vultr have partnered to create a streamlined solution for deploying AI applications from pilot phase to production environments. The collaboration leverages SUSE's AI Factory framewo

Tensordyne targets AI inference market with logarithmic math and Juniper-derived rack architecture

HardwareDGX agent

The race to serve AI inference faster and cheaper is exposing the hard limits of conventional chip architecture. As demand for real-time AI responses accelerates, the industry’s standard response — st

Tuning-Free Latent Diffusion Models for Ultrahigh-Resolution Image Editing

HardwareDGX agent

arXiv:2607.06136v1 Announce Type: new Abstract: Recent diffusion-based generative models have shown impressive performance in image generation and editing. However, due to memory limitations and the h

UBEP: Re-architecting Expert Parallelism Communication Library for Production Superpods

HardwareDGX agent

arXiv:2607.06202v1 Announce Type: cross Abstract: The deployment of Mixture-of-Experts (MoE) models on production high-bandwidth superpods, such as NVIDIA's NVL72/576 and Huawei's CloudMatrix384, intr

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a han…

HardwareDGX agent

We raised 130M @ 1B for our series A To build the open superintelligence stack for everyone Pre-training concentrated frontier AI in a handful of labs. RL changes who can build frontier AI and just wo

We're releasing ZML/LLMD, our homegrown LLM server built on top of our homegrown high performance heterogeneous inference stack. It ships wi…

HardwareDGX agent

We're releasing ZML/LLMD, our homegrown LLM server built on top of our homegrown high performance heterogeneous inference stack. It ships with 5 architectures out of the box: NVIDIA, AMD, Metal, Intel

7 Jul 2026

A Deep Learning-based surrogate model for Severe Accidents in nuclear reactors using ASTEC

HardwareDGX agent

arXiv:2607.04450v1 Announce Type: cross Abstract: Integral codes like the Accident Source Term Evaluation Code (ASTEC) are powerful tools to study the physics of Severe Accidents (SAs) in nuclear reac

AI debt financing is on an insane trajectory, per @SemiAnalysis_:

HardwareDGX agent

AI debt financing is on an insane trajectory, per @SemiAnalysis_: Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters Over 7T AI debt by 2029, There can be no N

AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters

HardwareDGX agent

Max single-threaded CPUs at scale are a new category of CPUs built for the agentic AI era. Across the creation and deployment of an agentic system, the CPU is on the critical path for reasoning, respo

AquaGen: Scaling generative models to molecular dynamics precision on thousands of atoms

HardwareDGX agent

arXiv:2607.03513v1 Announce Type: cross Abstract: We present AquaGen, the first all-atom, explicit solvent, periodic-boundary-condition-aware generative model that produces molecular configurations fr

CPR: Chained Perceptual Refinement for Coarse-to-Fine Medical Image Classification

HardwareDGX agent

arXiv:2607.02591v1 Announce Type: new Abstract: High resolution medical images contain fine grained, spatially sparse cues that are critical for diagnosis, yet preserving full resolution incurs substa

Data Driven Optimization of GPU efficiency for Distributed LLM-Adapter Serving

HardwareDGX agent

arXiv:2602.24044v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) adapters enable low-cost model specialization, but introduce complex caching and scheduling challenges in distribut

Fast Asymptotically Optimal Kinodynamic Planning via Vectorization

HardwareDGX agent

arXiv:2607.03987v1 Announce Type: new Abstract: Sampling-based motion planners have been shown to be effective for systems with complex kinodynamic constraints and high dimensionality. However, these

From Arithmetic to Logic: The Resilience of Logic and Lookup-Based Neural Networks Under Parameter Bit-Flips

Model ReleasesDGX agent

arXiv:2603.22770v2 Announce Type: replace-cross Abstract: The deployment of deep neural networks (DNNs) in safety-critical edge environments necessitates robustness against hardware-induced bit-flip e

Graph Sparse Sampling: Breaking the Curse of the Horizon in Continuous MDP Planning

HardwareDGX agent

arXiv:2607.05359v1 Announce Type: new Abstract: Planning under uncertainty in continuous domains is essential for autonomous systems, yet computationally demanding. Tree-based search methods such as M

Lights, Camera, Carbon: Architectural Scaling Laws for Video Generation Energy Consumption

HardwareDGX agent

arXiv:2607.04553v1 Announce Type: cross Abstract: We present a bidirectional framework for estimating the energy consumption of text-to-video (T2V) and text-to-video-audio (T2VA) models from architect

NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community

HardwareDGX agent

Open source AI has shown how quickly developers can innovate when models, data and tools are shared. Robotics has the same opportunity, but advancements in physical AI development can still be gated b

NVIDIA Vera CPU Boosts AI Factory Throughput to Accelerate Agentic Workloads

HardwareDGX agent

NVIDIA Vera is a CPU designed to help AI factories scale agentic AI and reinforcement learning by shortening CPU execution time, increasing task throughput, and enabling smarter, longer-thinking agent

OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioning, Placement, and Scheduling

HardwareDGX agent

arXiv:2509.23722v2 Announce Type: replace-cross Abstract: Pipeline parallelism is widely used to train large language models (LLMs). However, increasing heterogeneity in model architectures exacerbate

On the Design Space of Discrete Diffusion Online Adaptation for Molecular Optimization

HardwareDGX agent

arXiv:2607.02834v1 Announce Type: new Abstract: Molecular optimization often starts from a pretrained generative model that captures a broad prior over valid molecular structures. At test time, howeve

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the lates…

HardwareDGX agent

Open-source robotics just leveled up. 🦾 @HuggingFace's LeRobot 0.6 now includes NVIDIA Isaac GR00T 1.7 and Isaac Teleop, bringing the latest Isaac models and frameworks to the open robotics community.

PEEK: Predictive Queue-Informed KV Cache Management for LLM Serving

HardwareDGX agent

arXiv:2607.02525v1 Announce Type: cross Abstract: We present PEEK, a lightweight scheduling and eviction framework for both online (streaming) and offline (batch) LLM serving; this paper focuses on th

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

Model ReleasesDGX agent

arXiv:2504.08791v3 Announce Type: replace-cross Abstract: On-device inference offers privacy, offline use, and instant response, but consumer hardware restricts large language models (LLMs) to low thr

← Previous
1…1415161718…75
Next →