AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,730 results
Hardware

VFA: Relieving Vector Operations in Flash Attention with Global Maximum Pre-computation

DGX agent

arXiv:2604.12798v1 Announce Type: cross Abstract: FlashAttention-style online softmax enables exact attention computation with linear memory by streaming score tiles through on-chip memory and maintai

hardwarearxiv-cs-ai
15 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Hardware

Vultr Easily Meets EU Digital Readiness Operations Act (DORA) Requirements in Independent Assessment

DGX agent

Vultr, a cloud infrastructure provider, has successfully met the requirements of the EU's Digital Operational Resilience Act (DORA) according to an independent assessment, demonstrating its compliance

hardwarevultr
15 Apr 2026
Hardware

Vultr Managed Kafka® Connect Enhancements: Faster and Simpler Data Streaming

DGX agent

Enhance real-time data streaming with Vultr Managed Apache Kafka®. With expanded Kafka Connect support, data sources and destinations can be quickly integrated with just a few clicks. Get started toda

hardwarevultr
15 Apr 2026
Hardware

AIRA_2: Overcoming Bottlenecks in AI Research Agents

DGX agent

arXiv:2603.26499v2 Announce Type: replace Abstract: Existing research has identified three structural performance bottlenecks in AI research agents: (1) synchronous single-GPU execution constrains sam

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Building Custom Atomistic Simulation Workflows for Chemistry and Materials Science with NVIDIA ALCHEMI Toolkit

DGX agent

NVIDIA's ALCHEMI Toolkit is a PyTorch-native framework that lets researchers build custom GPU-accelerated atomistic simulation workflows for chemistry and materials science applications. The initial r

hardwarenvidia-developer
14 Apr 2026
Hardware

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

DGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

CASK: Core-Aware Selective KV Compression for Reasoning Traces

DGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

DGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts

DGX agent

arXiv:2604.10496v1 Announce Type: new Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) ar

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

Credo to acquire Israeli silicon photonics startup DustPhotonics for up to $1.3 billion

DGX agent

Fabless chip design company Credo Technology Group Holding Ltd. today announced that it has reached a deal to acquire Israeli silicon photonics startup DustPhotonics Ltd. for up to 1.3 billion in cash

hardwaresiliconangle
14 Apr 2026
Hardware

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

DGX agent

arXiv:2410.21316v2 Announce Type: replace-cross Abstract: Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of pa

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

From Decision Trees to Boolean Logic: A Fast and Unified SHAP Algorithm

DGX agent

arXiv:2511.09376v2 Announce Type: replace Abstract: SHapley Additive exPlanations (SHAP) is a key tool for interpreting decision tree ensembles by assigning contribution values to features. It is wide

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

From operational to analytical: The unified Spanner Graph and BigQuery Graph solution

DGX agent

Understanding the relationships within your data is crucial for uncovering hidden insights and building intelligent applications. However, managing operational (OLTP) and analytical (OLAP) graph workl

hardwaregoogle-cloud-ai
14 Apr 2026
Hardware

Generative Design for Direct-to-Chip Liquid Cooling for Data Centers

DGX agent

arXiv:2604.10941v1 Announce Type: cross Abstract: Rapid growth in artificial intelligence (AI) workloads is driving up data center power densities, increasing the need for advanced thermal management.

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

GPU-Accelerated Continuous-Time Successive Convexification for Contact-Implicit Legged Locomotion

DGX agent

arXiv:2604.09993v1 Announce Type: new Abstract: Contact-implicit trajectory optimization (CITO) enables the automatic discovery of contact sequences, but most methods rely on fine time discretization

hardwarearxiv-cs-ro
14 Apr 2026
Hardware

GPU Acceleration of Sparse Fully Homomorphic Encrypted DNNs

DGX agent

arXiv:2604.11659v1 Announce Type: cross Abstract: Fully homomorphic encryption (FHE) has recently attracted significant attention as both a cryptographic primitive and a systems challenge. Given the l

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

Hardware Utilization and Inference Performance of Edge Object Detection Under Fault Injection

DGX agent

arXiv:2604.09631v1 Announce Type: cross Abstract: As deep learning models are deployed on resource constrained edge platforms in autonomous driving systems, reli able knowledge of hardware behavior un

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7

DGX agent

This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The

hardwarer-ollama
14 Apr 2026
Hardware

IceCache: Memory-efficient KV-cache Management for Long-Sequence LLMs

DGX agent

arXiv:2604.10539v1 Announce Type: cross Abstract: Key-Value (KV) cache plays a crucial role in accelerating inference in large language models (LLMs) by storing intermediate attention states and avoid

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GP…

DGX agent

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GPU, PyTorch & OS - Multiple kernel versions coexist in one pr

hardwareclem-delangue--x
14 Apr 2026
Hardware

Lifetime-Aware Design for Item-Level Intelligence at the Extreme Edge

DGX agent

arXiv:2509.08193v2 Announce Type: replace-cross Abstract: We present FlexiFlow, a lifetime-aware design framework for item-level intelligence (ILI) where computation is integrated directly into dispos

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

MoBo, CPU, RAM suggestion | I'm terrible at guessing hardware | no gpu

DGX agent

This Reddit thread from r/ollama features a user seeking community recommendations for motherboard, CPU, and RAM components to build a system optimized for running Ollama locally without a dedicated G

hardwarer-ollama
14 Apr 2026
Hardware

NVIDIA Ising Introduces AI-Powered Workflows to Build Fault-Tolerant Quantum Systems

DGX agent

NVIDIA Ising is the world's first family of open-source quantum AI models, designed to help researchers and enterprises build quantum processors capable of running useful applications. The family span

hardwarenvidia-developer
14 Apr 2026
Hardware

NVIDIA NVbandwidth: Your Essential Tool for Measuring GPU Interconnect and Memory Performance

DGX agent

NVIDIA NVbandwidth is a performance analysis tool designed to measure memory bandwidth and latency between various components in GPU-equipped systems, providing detailed measurements of data transfer

hardwarenvidia-developer
14 Apr 2026
Hardware

Nvidia unveils Ising AI models for quantum error correction and calibration

DGX agent

Technology and computing giant Nvidia Corp. today announced the release of Ising, the world’s first open artificial intelligence model family aimed at quantum computing calibration and error correctio

hardwaresiliconangle
14 Apr 2026
Hardware

PA-SFM: Tracker-free differentiable acoustic radiation for freehand 3D photoacoustic imaging

DGX agent

arXiv:2604.09643v1 Announce Type: new Abstract: Three-dimensional (3D) handheld photoacoustic tomography typically relies on bulky and expensive external positioning sensors to correct motion artifact

hardwarearxiv-cs-cv
14 Apr 2026
Hardware

Parisians: we're running an open source AI art hackathon with LTX + NVIDIA this Saturday

DGX agent

A Reddit post on r/StableDiffusion announces an in-person open source AI art hackathon held in Paris, co-organized with LTX (Lightricks' open source AI video model) and NVIDIA. The event likely invite

hardwarer-stablediffusion
14 Apr 2026
Hardware

Read the full research post with NVIDIA: http://cursor.com/blog/multi-agent-kernels

DGX agent

Cursor has published a research post in collaboration with NVIDIA exploring multi-agent kernel systems, likely detailing how multiple AI agents can work together to handle complex coding or computatio

hardwarecursor--x
14 Apr 2026
Hardware

Real-Time Voicemail Detection in Telephony Audio Using Temporal Speech Activity Features

DGX agent

arXiv:2604.09675v1 Announce Type: cross Abstract: Outbound AI calling systems must distinguish voicemail greetings from live human answers in real time to avoid wasted agent interactions and dropped c

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Record-Remix-Replay: Hierarchical GPU Kernel Optimization using Evolutionary Search

DGX agent

arXiv:2604.11109v1 Announce Type: cross Abstract: As high-performance computing and AI workloads become increasingly dependent on GPUs, maintaining high performance across rapidly evolving hardware ge

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Relational Probing: LM-to-Graph Adaptation for Financial Prediction

DGX agent

arXiv:2604.10212v1 Announce Type: new Abstract: Language models can be used to identify relationships between financial entities in text. However, while structured output mechanisms exist, prompting-b

hardwarearxiv-cs-cl
14 Apr 2026
Hardware

SatReg: Regression-based Neural Architecture Search for Lightweight Satellite Image Segmentation

DGX agent

arXiv:2604.10306v1 Announce Type: new Abstract: As Earth-observation workloads move toward onboard and edge processing, remote-sensing segmentation models must operate under tight latency and energy c

hardwarearxiv-cs-cv
14 Apr 2026
Hardware

SCNO: Spiking Compositional Neural Operator -- Towards a Neuromorphic Foundation Model for Nuclear PDE Solving

DGX agent

arXiv:2604.11625v1 Announce Type: cross Abstract: Neural operators have emerged as powerful surrogates for partial differential equation (PDE) solvers, yet they are typically trained as monolithic mod

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

StreamServe: Adaptive Speculative Flows for Low-Latency Disaggregated LLM Serving

DGX agent

arXiv:2604.09562v1 Announce Type: cross Abstract: Efficient LLM serving must balance throughput and latency across diverse, bursty workloads. We introduce StreamServe, a disaggregated prefill decode s

hardwarearxiv-cs-ai
14 Apr 2026
Hardware

Sustainable Transformer Neural Network Acceleration with Stochastic Photonic Computing

DGX agent

arXiv:2604.09759v1 Announce Type: cross Abstract: Transformers achieve state-of-the-art performance in natural language processing, vision, and scientific computing, but demand high computation and me

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

Tessera: Unlocking Heterogeneous GPUs through Kernel-Granularity Disaggregation

DGX agent

arXiv:2604.10180v1 Announce Type: cross Abstract: Disaggregation maps parts of an AI workload to different types of GPUs, offering a path to utilize modern heterogeneous GPU clusters. However, existin

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

The Transformation Edge: Why the modern CFO is becoming the enterprise’s AI co-architect

DGX agent

There are moments in tech when the stack shifts. And then there are moments when everything shifts. This is the latter. After three decades in Silicon Valley — and now building a bridge between Palo A

hardwaresiliconangle
14 Apr 2026
Hardware

Tree Training: Accelerating Agentic LLMs Training via Shared Prefix Reuse

DGX agent

arXiv:2511.00413v4 Announce Type: replace Abstract: Agentic large language model (LLM) training often involves multi-turn interaction trajectories that branch into multiple execution paths due to conc

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

VTC: DNN Compilation with Virtual Tensors for Data Movement Elimination

DGX agent

arXiv:2604.09558v1 Announce Type: cross Abstract: With the widening gap between compute and memory operation latencies, data movement optimizations have become increasingly important for DNN compilati

hardwarearxiv-cs-lg
14 Apr 2026
Hardware

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to ap…

DGX agent

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to apply it to optimizing CUDA kernels. In 3 weeks, it delivered

hardwarecursor--x
14 Apr 2026
Hardware

what's the best place to buy GPU server?

DGX agent

This Reddit thread from r/ollama discusses community recommendations for purchasing or renting GPU servers to run Ollama and local LLMs. It likely covers options ranging from dedicated GPU server prov

hardwarer-ollama
14 Apr 2026
Hardware

$200/month is enough to buy an H100 GPU for 6 hours every workday

DGX agent

Soumith Chintala shared a post highlighting that $200 per month is sufficient to rent access to an NVIDIA H100 GPU for approximately 6 hours every workday, making high-end AI compute more accessible t

hardwaresoumith-chintala--x
13 Apr 2026
Hardware

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference

DGX agent

arXiv:2604.08584v1 Announce Type: cross Abstract: Long-context LLMs increasingly rely on extended, reusable prefill prompts for agents and domain Q&A, pushing attention and KV-cache to become the domi

hardwarearxiv-cs-ai
13 Apr 2026
Hardware

Fast Model-guided Instance-wise Adaptation Framework for Real-world Pansharpening with Fidelity Constraints

DGX agent

arXiv:2604.08903v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

hardwarearxiv-cs-cv
13 Apr 2026
Hardware

FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs

DGX agent

arXiv:2512.20033v2 Announce Type: replace Abstract: We present FlashLips, a two-stage, mask-free lip-sync system that decouples lips control from rendering and achieves real-time performance, with our

hardwarearxiv-cs-cv
13 Apr 2026
Hardware

Free AI Voice Cloning with Qwen3 TTS — Google Colab Notebook (works on free tier, no GPU needed)

DGX agent

This Reddit post shares a Google Colab notebook that enables free AI voice cloning using Alibaba's Qwen3-TTS, which offers voice cloning, voice design, and ultra-high-quality human-like speech generat

hardwarer-stablediffusion
13 Apr 2026
Hardware

High-dimensional inference for the gamma-ray sky with differentiable programming

DGX agent

arXiv:2604.08648v1 Announce Type: cross Abstract: We motivate the use of differentiable probabilistic programming techniques in order to account for the large model-space inherent to astrophysical gam

hardwarearxiv-cs-lg
13 Apr 2026
Hardware

Is an nvidia DGK Spark or similar worth it?

DGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

hardwarer-ollama
13 Apr 2026
← Previous
1…3334353637
Next →