AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
15 Apr 2026

ISSCC 2026: NVIDIA & Broadcom CPO, HBM4 & LPDDR6, TSMC Active LSI, Logic-Based SRAM, UCIe-S and More

HardwareDGX agent

ISSCC 2026 previews cutting-edge semiconductor research and industry presentations covering co-packaged optics (CPO) developments from NVIDIA and Broadcom, next-generation memory standards including H

@lmsysorg @sgl_project @vllm_project @CoreWeave @nebiusai @nscale @togethercompute @Togethercompute enables AI labs and enterprises to run f…

HardwareDGX agent

@lmsysorg @sgl_project @vllm_project @CoreWeave @nebiusai @nscale @togethercompute @Togethercompute enables AI labs and enterprises to run frontier models and agents on the NVIDIA Blackwell platform—g

New Adobe Premiere Color Grading Mode Accelerated on NVIDIA GPUs

HardwareDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The NAB Show 2026 trade show, running April 18-22 in Las Vegas, is set to showcase a wave of new features and optimizations for top video editing applications. Bringing together over 60,000 content pr

PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving

HardwareDGX agent

arXiv:2604.12171v1 Announce Type: cross Abstract: Pipeline parallelism (PP) is widely used to partition layers of large language models (LLMs) across GPUs, enabling scalable inference for large models

Q&A with Jensen Huang on Nvidia's supply chain moat, competition from ASICs like Google's TPU, investing in AI labs and neoclouds, selling to China, and more (Dwarkesh Patel/Dwarkesh Podcast)

HardwareDGX agent

Dwarkesh Patel / Dwarkesh Podcast: Q&A with Jensen Huang on Nvidia's supply chain moat, competition from ASICs like Google's TPU, investing in AI labs and neoclouds, selling to China, and more — “If o

Roop Unleashed 4.3.1 not fully utilizing RTX 5070 Ti / 5080X (Low GPU/RAM usage)

HardwareDGX agent

This Reddit thread from r/StableDiffusion discusses a performance issue where Roop Unleashed version 4.3.1 — a face-swapping tool commonly used alongside Stable Diffusion — fails to fully utilize the

Scalable Trajectory Generation for Whole-Body Mobile Manipulation

HardwareDGX agent

arXiv:2604.12565v1 Announce Type: cross Abstract: Robots deployed in unstructured environments must coordinate whole-body motion -- simultaneously moving a mobile base and arm -- to interact with the

Security Researchers - Get Rewarded for Bug Reports!

HardwareDGX agent

Discover how Vultr values and rewards the ethical disclosure of security vulnerabilities with its new bug bounty program. Report issues, get compensated, and help improve Vultr's web ecosystem through

Set up your domains with Vultr DNS!

HardwareDGX agent

Discover Vultr DNS, a free, geographically distributed anycast DNS service that improves browsing speed and reliability by connecting visitors to the closest DNS cluster. Start configuring your domain

Simplify Global Traffic Management with Vultr Global Load Balancers

HardwareDGX agent

Vultr Global Load Balancers is a managed service that distributes incoming traffic across multiple regions and backend servers worldwide, helping businesses improve application availability, reduce la

Social media intern made me a terminally online LEGO 😭

HardwareDGX agent

Dylan Patel, a semiconductor and AI industry analyst known for his work at SemiAnalysis, shared a post on X (formerly Twitter) humorously reacting to a LEGO-style avatar or depiction of himself create

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token…

HardwareDGX agent

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token Town' - the crack team of AI Engineers and Model Behavior En

VFA: Relieving Vector Operations in Flash Attention with Global Maximum Pre-computation

HardwareDGX agent

arXiv:2604.12798v1 Announce Type: cross Abstract: FlashAttention-style online softmax enables exact attention computation with linear memory by streaming score tiles through on-chip memory and maintai

Vultr Easily Meets EU Digital Readiness Operations Act (DORA) Requirements in Independent Assessment

HardwareDGX agent

Vultr, a cloud infrastructure provider, has successfully met the requirements of the EU's Digital Operational Resilience Act (DORA) according to an independent assessment, demonstrating its compliance

Vultr Managed Kafka® Connect Enhancements: Faster and Simpler Data Streaming

HardwareDGX agent

Enhance real-time data streaming with Vultr Managed Apache Kafka®. With expanded Kafka Connect support, data sources and destinations can be quickly integrated with just a few clicks. Get started toda

14 Apr 2026

AIRA_2: Overcoming Bottlenecks in AI Research Agents

HardwareDGX agent

arXiv:2603.26499v2 Announce Type: replace Abstract: Existing research has identified three structural performance bottlenecks in AI research agents: (1) synchronous single-GPU execution constrains sam

Building Custom Atomistic Simulation Workflows for Chemistry and Materials Science with NVIDIA ALCHEMI Toolkit

HardwareDGX agent

NVIDIA's ALCHEMI Toolkit is a PyTorch-native framework that lets researchers build custom GPU-accelerated atomistic simulation workflows for chemistry and materials science applications. The initial r

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

HardwareDGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

CASK: Core-Aware Selective KV Compression for Reasoning Traces

HardwareDGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

HardwareDGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts

HardwareDGX agent

arXiv:2604.10496v1 Announce Type: new Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) ar

Credo to acquire Israeli silicon photonics startup DustPhotonics for up to $1.3 billion

HardwareDGX agent

Fabless chip design company Credo Technology Group Holding Ltd. today announced that it has reached a deal to acquire Israeli silicon photonics startup DustPhotonics Ltd. for up to 1.3 billion in cash

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

HardwareDGX agent

arXiv:2410.21316v2 Announce Type: replace-cross Abstract: Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of pa

From Decision Trees to Boolean Logic: A Fast and Unified SHAP Algorithm

HardwareDGX agent

arXiv:2511.09376v2 Announce Type: replace Abstract: SHapley Additive exPlanations (SHAP) is a key tool for interpreting decision tree ensembles by assigning contribution values to features. It is wide

From operational to analytical: The unified Spanner Graph and BigQuery Graph solution

HardwareDGX agent

Understanding the relationships within your data is crucial for uncovering hidden insights and building intelligent applications. However, managing operational (OLTP) and analytical (OLAP) graph workl

Generative Design for Direct-to-Chip Liquid Cooling for Data Centers

HardwareDGX agent

arXiv:2604.10941v1 Announce Type: cross Abstract: Rapid growth in artificial intelligence (AI) workloads is driving up data center power densities, increasing the need for advanced thermal management.

GPU-Accelerated Continuous-Time Successive Convexification for Contact-Implicit Legged Locomotion

HardwareDGX agent

arXiv:2604.09993v1 Announce Type: new Abstract: Contact-implicit trajectory optimization (CITO) enables the automatic discovery of contact sequences, but most methods rely on fine time discretization

GPU Acceleration of Sparse Fully Homomorphic Encrypted DNNs

HardwareDGX agent

arXiv:2604.11659v1 Announce Type: cross Abstract: Fully homomorphic encryption (FHE) has recently attracted significant attention as both a cryptographic primitive and a systems challenge. Given the l

Hardware Utilization and Inference Performance of Edge Object Detection Under Fault Injection

HardwareDGX agent

arXiv:2604.09631v1 Announce Type: cross Abstract: As deep learning models are deployed on resource constrained edge platforms in autonomous driving systems, reli able knowledge of hardware behavior un

IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7

HardwareDGX agent

This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The

IceCache: Memory-efficient KV-cache Management for Long-Sequence LLMs

HardwareDGX agent

arXiv:2604.10539v1 Announce Type: cross Abstract: Key-Value (KV) cache plays a crucial role in accelerating inference in large language models (LLMs) by storing intermediate attention states and avoid

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GP…

HardwareDGX agent

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GPU, PyTorch & OS - Multiple kernel versions coexist in one pr

Lifetime-Aware Design for Item-Level Intelligence at the Extreme Edge

HardwareDGX agent

arXiv:2509.08193v2 Announce Type: replace-cross Abstract: We present FlexiFlow, a lifetime-aware design framework for item-level intelligence (ILI) where computation is integrated directly into dispos

MoBo, CPU, RAM suggestion | I'm terrible at guessing hardware | no gpu

HardwareDGX agent

This Reddit thread from r/ollama features a user seeking community recommendations for motherboard, CPU, and RAM components to build a system optimized for running Ollama locally without a dedicated G

NVIDIA Ising Introduces AI-Powered Workflows to Build Fault-Tolerant Quantum Systems

HardwareDGX agent

NVIDIA Ising is the world's first family of open-source quantum AI models, designed to help researchers and enterprises build quantum processors capable of running useful applications. The family span

NVIDIA NVbandwidth: Your Essential Tool for Measuring GPU Interconnect and Memory Performance

HardwareDGX agent

NVIDIA NVbandwidth is a performance analysis tool designed to measure memory bandwidth and latency between various components in GPU-equipped systems, providing detailed measurements of data transfer

Nvidia unveils Ising AI models for quantum error correction and calibration

HardwareDGX agent

Technology and computing giant Nvidia Corp. today announced the release of Ising, the world’s first open artificial intelligence model family aimed at quantum computing calibration and error correctio

PA-SFM: Tracker-free differentiable acoustic radiation for freehand 3D photoacoustic imaging

HardwareDGX agent

arXiv:2604.09643v1 Announce Type: new Abstract: Three-dimensional (3D) handheld photoacoustic tomography typically relies on bulky and expensive external positioning sensors to correct motion artifact

Parisians: we're running an open source AI art hackathon with LTX + NVIDIA this Saturday

HardwareDGX agent

A Reddit post on r/StableDiffusion announces an in-person open source AI art hackathon held in Paris, co-organized with LTX (Lightricks' open source AI video model) and NVIDIA. The event likely invite

Read the full research post with NVIDIA: http://cursor.com/blog/multi-agent-kernels

HardwareDGX agent

Cursor has published a research post in collaboration with NVIDIA exploring multi-agent kernel systems, likely detailing how multiple AI agents can work together to handle complex coding or computatio

Real-Time Voicemail Detection in Telephony Audio Using Temporal Speech Activity Features

HardwareDGX agent

arXiv:2604.09675v1 Announce Type: cross Abstract: Outbound AI calling systems must distinguish voicemail greetings from live human answers in real time to avoid wasted agent interactions and dropped c

Record-Remix-Replay: Hierarchical GPU Kernel Optimization using Evolutionary Search

HardwareDGX agent

arXiv:2604.11109v1 Announce Type: cross Abstract: As high-performance computing and AI workloads become increasingly dependent on GPUs, maintaining high performance across rapidly evolving hardware ge

Relational Probing: LM-to-Graph Adaptation for Financial Prediction

HardwareDGX agent

arXiv:2604.10212v1 Announce Type: new Abstract: Language models can be used to identify relationships between financial entities in text. However, while structured output mechanisms exist, prompting-b

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

HardwareDGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

HardwareDGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

StreamServe: Adaptive Speculative Flows for Low-Latency Disaggregated LLM Serving

HardwareDGX agent

arXiv:2604.09562v1 Announce Type: cross Abstract: Efficient LLM serving must balance throughput and latency across diverse, bursty workloads. We introduce StreamServe, a disaggregated prefill decode s

Sustainable Transformer Neural Network Acceleration with Stochastic Photonic Computing

HardwareDGX agent

arXiv:2604.09759v1 Announce Type: cross Abstract: Transformers achieve state-of-the-art performance in natural language processing, vision, and scientific computing, but demand high computation and me

Tessera: Unlocking Heterogeneous GPUs through Kernel-Granularity Disaggregation

HardwareDGX agent

arXiv:2604.10180v1 Announce Type: cross Abstract: Disaggregation maps parts of an AI workload to different types of GPUs, offering a path to utilize modern heterogeneous GPU clusters. However, existin

The Transformation Edge: Why the modern CFO is becoming the enterprise’s AI co-architect

HardwareDGX agent

There are moments in tech when the stack shifts. And then there are moments when everything shifts. This is the latter. After three decades in Silicon Valley — and now building a bridge between Palo A

Tree Training: Accelerating Agentic LLMs Training via Shared Prefix Reuse

HardwareDGX agent

arXiv:2511.00413v4 Announce Type: replace Abstract: Agentic large language model (LLM) training often involves multi-turn interaction trajectories that branch into multiple execution paths due to conc

VTC: DNN Compilation with Virtual Tensors for Data Movement Elimination

HardwareDGX agent

arXiv:2604.09558v1 Announce Type: cross Abstract: With the widening gap between compute and memory operation latencies, data movement optimizations have become increasingly important for DNN compilati

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to ap…

HardwareDGX agent

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to apply it to optimizing CUDA kernels. In 3 weeks, it delivered

what's the best place to buy GPU server?

HardwareDGX agent

This Reddit thread from r/ollama discusses community recommendations for purchasing or renting GPU servers to run Ollama and local LLMs. It likely covers options ranging from dedicated GPU server prov

13 Apr 2026

$200/month is enough to buy an H100 GPU for 6 hours every workday

HardwareDGX agent

Soumith Chintala shared a post highlighting that $200 per month is sufficient to rent access to an NVIDIA H100 GPU for approximately 6 hours every workday, making high-end AI compute more accessible t

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference

HardwareDGX agent

arXiv:2604.08584v1 Announce Type: cross Abstract: Long-context LLMs increasingly rely on extended, reusable prefill prompts for agents and domain Q&A, pushing attention and KV-cache to become the domi

Fast Model-guided Instance-wise Adaptation Framework for Real-world Pansharpening with Fidelity Constraints

HardwareDGX agent

arXiv:2604.08903v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs

HardwareDGX agent

arXiv:2512.20033v2 Announce Type: replace Abstract: We present FlashLips, a two-stage, mask-free lip-sync system that decouples lips control from rendering and achieves real-time performance, with our

Free AI Voice Cloning with Qwen3 TTS — Google Colab Notebook (works on free tier, no GPU needed)

HardwareDGX agent

This Reddit post shares a Google Colab notebook that enables free AI voice cloning using Alibaba's Qwen3-TTS, which offers voice cloning, voice design, and ultra-high-quality human-like speech generat

High-dimensional inference for the gamma-ray sky with differentiable programming

HardwareDGX agent

arXiv:2604.08648v1 Announce Type: cross Abstract: We motivate the use of differentiable probabilistic programming techniques in order to account for the large model-space inherent to astrophysical gam

Is an nvidia DGK Spark or similar worth it?

HardwareDGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

← Previous
1…26272829
Next →