AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
Hardware

This is your laptop… on AI

DGX agent

We're now deep into developer conference season, and one of the themes so far is the relentless conviction from Big Tech companies that AI is going to change everything about how we do everything. Nvi

hardwarethe-verge-ai
5 Jun 2026
Hardware

Today we published a technical blog post about Ideogram 4.0 — our goal is to enable more innovation and creativity. It's a 9.3B Diffusion Tr…

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

Today we published a technical blog post about Ideogram 4.0 — our goal is to enable more innovation and creativity. It's a 9.3B Diffusion Transformer trained from scratch, paired with a frozen 8B VLM

hardwareclem-delangue--x
5 Jun 2026
Hardware

Uncertainty-Aware Adaptive Sensor Fusion for Autonomous Navigation

DGX agent

arXiv:2606.05437v1 Announce Type: cross Abstract: This work introduces a hybrid deep learning approach integrated with an Unscented Kalman Filter (UKF) to enhance pose estimation accuracy in Visual-In

hardwarearxiv-cs-cv
5 Jun 2026
Hardware

US-traded chipmakers plunged on Friday, after Broadcom missed expectations; Nvidia fell 6.19%, Micron fell 13.25%, AMD fell 10.86%, and Broadcom 7.92% (Noel Randewich/Reuters)

DGX agent

Noel Randewich / Reuters: US-traded chipmakers plunged on Friday, after Broadcom missed expectations; Nvidia fell 6.19%, Micron fell 13.25%, AMD fell 10.86%, and Broadcom 7.92% — U.S.-traded chipmaker

hardwaretechmeme
5 Jun 2026
Hardware

World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis

DGX agent

arXiv:2606.05979v1 Announce Type: new Abstract: We propose world-language-action (WLA) models as a new class of embodied foundation models. WLA takes textual instructions, images, and robot states as

hardwarearxiv-cs-ro
5 Jun 2026
Hardware

AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2606.04484v1 Announce Type: new Abstract: We present AgentJet, a distributed swarm training framework for large language model (LLM) agent reinforcement learning. Unlike centralized frameworks t

hardwarearxiv-cs-ai
4 Jun 2026
Local Ai

b9515

DGX agent

llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Build b9515 is an interme

local-aillama-cpp-releases
4 Jun 2026
Hardware

Cartridges at Scale: Training Modular KV Caches over Large Document Collections

DGX agent

arXiv:2606.04557v1 Announce Type: new Abstract: Large Language Models can reason over long contexts, yet prefilling millions of tokens is wasteful as much of the content remains static across queries.

hardwarearxiv-cs-cl
4 Jun 2026
Hardware

DSA: Dynamic Step Allocation for Fast Autoregressive Video Generation

DGX agent

arXiv:2606.04432v1 Announce Type: new Abstract: Video diffusion transformers have achieved state-of-the-art visual quality, but their high inference cost remains a major bottleneck for real-time appli

hardwarearxiv-cs-cv
4 Jun 2026
Hardware

Ex-OpenAI Tech Lead, Justin Lebar joins SemiAnalysis as an Visiting Fellow to Burn $10,000 in 3 hours to find dozens of AMDGPU LLVM, x86 LLV…

DGX agent

Ex-OpenAI Tech Lead, Justin Lebar joins SemiAnalysis as an Visiting Fellow to Burn $10,000 in 3 hours to find dozens of AMDGPU LLVM, x86 LLVM, NVPTX bugs 00:00 - Intro & Justin’s background 00:59 - Ho

hardwaredylan-patel--x
4 Jun 2026
Hardware

Fitting scattered data with optional monotonicity constraints on GPU: LipFit package

DGX agent

arXiv:2606.04670v1 Announce Type: cross Abstract: This paper presents a method of multivariate scattered data interpolation and approximation that produces optimal Lipschitz-continuous approximation,

hardwarearxiv-cs-lg
4 Jun 2026
Hardware

Forecast: Fun Ahead — 18 Games Join in June to Stream on GeForce NOW

DGX agent

June’s forecast with GeForce NOW: 100% chance of gaming. GeForce NOW is lining up new adventures for the month, from big-name blockbusters to quirky indies ready for the spotlight. Members can dive in

hardwarenvidia-blog
4 Jun 2026
Hardware

Generalist AI raises 400M at 2B valuation to build general intelligence for robotics

DGX agent

Artificial intelligence startup Generalist AI Inc., a startup building embodied robotics intelligence, said today it has raised 400 million in new funding, bringing the company’s valuation to 2 billio

hardwaresiliconangle
4 Jun 2026
Hardware

I love that people reposting what we said miss most of what the note says. Happens all the time.

DGX agent

Dylan Patel observes that when people repost or share his notes on social media, they frequently misrepresent or overlook the main points of his original message. He notes this is a recurring pattern

hardwaredylan-patel--x
4 Jun 2026
Hardware

LLM Compression with Jointly Optimizing Architectural and Quantization choices

DGX agent

arXiv:2606.04063v1 Announce Type: cross Abstract: Deploying large language models (LLMs) is challenging due to their significant memory and computational requirements. While some methods address this

hardwarearxiv-cs-ai
4 Jun 2026
Hardware

Nvidia snaps up Kumo AI, a predictive AI startup known for its extreme accuracy

DGX agent

Nvidia Corp. has bagged itself another artificial intelligence startup, acquiring four-year-old model maker Kumo AI Inc. The company designs AI models focused on making extremely accurate business pre

hardwaresiliconangle
4 Jun 2026
Hardware

OpenAI is in deep, deep trouble They are low on capital relative to the massive cash they are burning and there is only so much capital in t…

DGX agent

OpenAI is in deep, deep trouble They are low on capital relative to the massive cash they are burning and there is only so much capital in the world. No rational person would sell their Bitcoin or Nvi

hardwaregary-marcus--x
4 Jun 2026
Hardware

Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models

DGX agent

arXiv:2606.04287v1 Announce Type: cross Abstract: Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity,

hardwarearxiv-cs-ai
4 Jun 2026
Hardware

SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference

DGX agent

arXiv:2606.04511v1 Announce Type: new Abstract: Sparse attention reduces compute and memory bandwidth for long-context LLM inference. However, two key challenges remain: (1) KV cache capacity still gr

hardwarearxiv-cs-cl
4 Jun 2026
Model Releases

StepPRM-RTL: Stepwise Process-Reward Guided LLM Fine-Tuning for Enhanced RTL Synthesis

DGX agent

arXiv:2606.04246v1 Announce Type: new Abstract: Automatic generation of RTL code for digital hardware designs remains challenging due to long-horizon reasoning, multi-step dependencies, and strict cor

model-releasesarxiv-cs-ai
4 Jun 2026
Hardware

TITAN-FedAnil+: Trust-Based Adaptive Blockchain Federated Learning for Resource-Constrained Intelligent Enterprises

DGX agent

arXiv:2606.04388v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as an effective paradigm for collaborative intelligence while preserving data privacy. However, data heterogeneity

hardwarearxiv-cs-ai
4 Jun 2026
Hardware

Together AI provides the inference stack behind both: high-throughput serving on the latest NVIDIA Blackwell GPUs for agentic workloads, and…

DGX agent

Together AI provides the inference stack behind both: high-throughput serving on the latest NVIDIA Blackwell GPUs for agentic workloads, and TensorRT engines plus event-driven streaming I/O for low-la

hardwaretogether-ai--x
4 Jun 2026
Hardware

Announcing Foundry Managed Compute: Run open models in Microsoft Foundry

DGX agent

Microsoft Foundry Managed Compute is a new GPU platform-as-a-service for hosting open-source and custom AI models behind the same endpoint, SDKs, and bill as frontier models. The post Announcing Found

hardwaremicrosoft-foundry
3 Jun 2026
Hardware

CoreWeave’s Vera Rubin milestone sets stage for theCUBE’s agentic AI coverage

DGX agent

The agentic AI era is putting new pressure on the infrastructure stack, and CoreWeave Inc.’s latest milestone gives the conversation a sharper edge. This week, the company announced that it has comple

hardwaresiliconangle
3 Jun 2026
Hardware

Floating Point: The Origin Story

DGX agent

This article explores the historical development and origins of floating-point number representation in computing, likely covering how early computer scientists and engineers designed methods to repre

hardwarethe-chip-letter
3 Jun 2026
Hardware

GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning

DGX agent

arXiv:2606.02857v1 Announce Type: cross Abstract: Zeroth-order (ZO) optimization is a memory-efficient alternative to backpropagation for fine-tuning large language models, but its deployment is limit

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

MOSAIC: Efficient Mixture-of-Agent Scheduling via Adaptive Aggregation and Inference Concurrency

DGX agent

arXiv:2606.03014v1 Announce Type: new Abstract: Mixture-of-Agents (MoA) systems improve reasoning accuracy by routing each query to multiple expert LLMs and aggregating their outputs. Efficiently exec

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

Nvidia acquired Kumo, which sells predictive AI software to enterprises, a source says for 400M+; PitchBook: Kumo raised 37M at a $250M valuation in 2022 (The Information)

DGX agent

The Information: Nvidia acquired Kumo, which sells predictive AI software to enterprises, a source says for 400M+; PitchBook: Kumo raised 37M at a 250M valuation in 2022 — Nvidia has bought Kumo AI, a

hardwaretechmeme
3 Jun 2026
Hardware

NVIDIA Enables the Next Era Of Physical AI Research With Agent Skills For Autonomous Vehicles, Robotics And Vision AI

DGX agent

At CVPR, NVIDIA is unveiling new physical AI agent skills that help researchers and developers speed the development of autonomous vehicles, robots and vision AI systems. The core challenge in physica

hardwarenvidia-blog
3 Jun 2026
Hardware

NVIDIA Isaac Sim: Enabling Scalable, GPU-Accelerated Simulation for Robotics

DGX agent

arXiv:2606.03551v1 Announce Type: new Abstract: Simulation has become a core infrastructure for robotics research. Unlike previous simulators, NVIDIA Isaac Sim leverages GPU acceleration to enable lar

hardwarearxiv-cs-ro
3 Jun 2026
Hardware

NVIDIA Research Unlocks Advanced Grasping, Smarter Autonomous Driving and Agent Training at Scale

DGX agent

What makes a robot gripper useful isn’t that it can pick up one object — it’s that it can pick up the next one, and the one after that, with a tool it’s never held before. What makes an autonomous veh

hardwarenvidia-blog
3 Jun 2026
Hardware

Pruning Deep Neural Networks via the Marchenko--Pastur Distribution

DGX agent

arXiv:2606.02608v1 Announce Type: new Abstract: We study a Marchenko--Pastur (MP) random-matrix approach to pruning deep neural networks with very small post-pruning fine-tuning budgets. The main prac

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

Quobly raises $150M for its silicon spin qubit technology

DGX agent

French quantum computing startup Quobly SAS today announced that it has raised €130 million, or about 150 million, to commercialize its technology. The Series A round was led by publicly traded chipma

hardwaresiliconangle
3 Jun 2026
Hardware

Scaling Multi Agent Reinforcement Learning for Underwater Acoustic Tracking via Autonomous Vehicles

DGX agent

arXiv:2505.08222v3 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) offer a cost-effective solution for scientific missions such as underwater tracking. Reinforcement learning (RL) has

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

Speedrunning Tabular Foundation Model Pretraining

DGX agent

arXiv:2606.03681v1 Announce Type: new Abstract: Pretraining cost is a major bottleneck for research on tabular foundation models, slowing the iteration cycle for new architectures, priors, and optimiz

hardwarearxiv-cs-lg
3 Jun 2026
Hardware

To Boldly Go: The Case for Space Datacenters

DGX agent

This article explores the potential viability and benefits of locating data centers in space rather than on Earth, likely discussing advantages such as reduced cooling costs due to the vacuum environm

hardwaresemianalysis
3 Jun 2026
Hardware

Towards Compact Autonomous Driving Perception with Balanced Learning and Multi-sensor Fusion

DGX agent

arXiv:2606.02979v1 Announce Type: cross Abstract: We present a novel compact deep multi-task learning model to handle various autonomous driving perception tasks in one forward pass. The model perform

hardwarearxiv-cs-ai
3 Jun 2026
Hardware

APB-V: Accelerating Long-Video Understanding via Sequence-Parallelism-aware Approximate Attention

DGX agent

arXiv:2601.21444v2 Announce Type: replace-cross Abstract: The efficiency of long-video inference remains a critical bottleneck, mainly due to the dense computation in the prefill stage of Large Multim

hardwarearxiv-cs-ai
2 Jun 2026
Local Ai

b9471

DGX agent

B9471 is a build release of llama.cpp, an open-source C/C++ implementation for LLM inference that enables running large language models on consumer hardware with optimized performance. This intermedia

local-aillama-cpp-releases
2 Jun 2026
Hardware

BudgetDraft: Acceptance-Aware Multi-View Training for Sparse-KV Speculative Decoding

DGX agent

arXiv:2606.00144v1 Announce Type: cross Abstract: Speculative decoding speeds up autoregressive decoding by using a drafter to propose multiple tokens that a verifier validates in parallel. In resourc

hardwarearxiv-cs-ai
2 Jun 2026
Hardware

Cohort-Scale Neural Atlases of Ultrasound Video

DGX agent

arXiv:2606.00890v1 Announce Type: new Abstract: Ultrasound is the most widely used real-time imaging modality in clinical practice, yet per-frame video annotation remains a major bottleneck: expert la

hardwarearxiv-cs-cv
2 Jun 2026
Hardware

CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving

DGX agent

arXiv:2603.28768v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) has recently emerged as the mainstream architecture for efficiently scaling large language models while maintaining n

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Deploy Agentic-Ready AI at the Edge with Memory Efficiency in NVIDIA JetPack 7.2

DGX agent

NVIDIA JetPack 7.2 enables deployment of AI agents to edge devices with optimized memory and performance for real-world applications. The release directly supports one-command deployment of NVIDIA Nem

hardwarenvidia-developer
2 Jun 2026
Hardware

Design Space Exploration of DMA based Finer-Grain Compute Communication Overlap

DGX agent

arXiv:2512.10236v2 Announce Type: replace-cross Abstract: Modern ML workloads demand distributing training and inference across multiple GPUs. However, these parallelization techniques often suffer fr

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

DuetServe: Harmonizing Prefill and Decode for LLM Serving via Adaptive GPU Multiplexing

DGX agent

arXiv:2511.04791v2 Announce Type: replace Abstract: Modern LLM serving systems must sustain high throughput while meeting strict latency SLOs across two distinct inference phases: compute-intensive pr

hardwarearxiv-cs-lg
2 Jun 2026
Hardware

Eyettention II: A Dual-Sequence Architecture for Modeling Fixation Location, Within-Word Landing Position, and Fixation Duration in Reading

DGX agent

arXiv:2606.01964v1 Announce Type: new Abstract: The way our eyes move while reading provides valuable insights into both the reader's cognitive processes and the properties of the text. In particular,

hardwarearxiv-cs-cl
2 Jun 2026
Hardware

First technical Deepdive on M3 on the internet😎

DGX agent

First technical Deepdive on M3 on the internet😎 MiniMax-M3 combines 1M context, native multimodality, and MiniMax Sparse Attention. The next layer is serving it efficiently: KV-block-major sparse atte

hardwaretogether-ai--x
2 Jun 2026
Hardware

Industrial Software Leaders Build Secure, Autonomous AI Engineers With NVIDIA NemoClaw

DGX agent

Accelerated computing has revolutionized industrial engineering, compressing simulation times from weeks to hours. Today’s remaining challenges sit in the end-to-end workflow surrounding the simulatio

hardwarenvidia-blog
2 Jun 2026
← Previous
1…2627282930…94
Next →