AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,477 results
16 Apr 2026

SemiFA: An Agentic Multi-Modal Framework for Autonomous Semiconductor Failure Analysis Report Generation

HardwareDGX agent

arXiv:2604.13236v1 Announce Type: new Abstract: Semiconductor failure analysis (FA) requires engineers to examine inspection images, correlate equipment telemetry, consult historical defect records, a

Sources: OpenAI agrees to pay Cerebras $20B+ to use its server chips, double the amount previously associated with the deal, and may receive equity in Cerebras (The Information)

HardwareDGX agent

The Information: Sources: OpenAI agrees to pay Cerebras $20B+ to use its server chips, double the amount previously associated with the deal, and may receive equity in Cerebras — OpenAI recently struc

to be clear, NVIDIA is NOT a car

HardwareDGX agent

Dylan Patel of SemiAnalysis posted a clarification on X distinguishing NVIDIA from automotive companies, likely in response to comparisons or confusion surrounding NVIDIA's growing presence in the aut

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
15 Apr 2026

Bandwidth Overage Alerts

HardwareDGX agent

Stay informed with new bandwidth usage alerts on Vultr.com as your instances near their transfer limits. For details, visit your Vultr Console's settings section. Give us your feedback on Twitter @Vul

Best part are retrieval and search details: >Agent queries differ from human queries & what good results means changes too >Parallel queries…

HardwareDGX agent

Best part are retrieval and search details: >Agent queries differ from human queries & what good results means changes too >Parallel queries and ranking are both tools to the same outcome >Top-K preci

Data Security Compliance Made Simple with Vultr

HardwareDGX agent

Discover how Vultr prioritizes data security and compliance, ensuring your data is protected across 32 cloud data center regions worldwide. Explore comprehensive compliance reports and robust certific

DreamStereo: Towards Real-Time Stereo Inpainting for HD Videos

HardwareDGX agent

arXiv:2604.12270v1 Announce Type: new Abstract: Stereo video inpainting, which aims to fill the occluded regions of warped videos with visually coherent content while maintaining temporal consistency,

finally: @simonlast + @sarahmsachs on Latent Space! Notion has rebuilt Notion AI 5 times. This is the first time Simon has told the entire s…

HardwareDGX agent

finally: @simonlast + @sarahmsachs on Latent Space! Notion has rebuilt Notion AI 5 times. This is the first time Simon has told the entire story. I've been trying to do this interview for ~3 years. We

Fine-Tuning LLMs for Report Summarization: Analysis on Supervised and Unsupervised Data

HardwareDGX agent

arXiv:2503.10676v2 Announce Type: replace-cross Abstract: We study the efficacy of fine-tuning Large Language Models (LLMs) for the specific task of report (government archives, news, intelligence rep

Get Organized with Cloud Server Tagging

HardwareDGX agent

Introducing instance tagging for better organization and easier server management on Vultr. Try tagging now and stay tuned for upcoming features like scheduled backups! Share your feedback on Twitter

GPU stays sometimes at 100% usage even when done replying. Is it normal?

HardwareDGX agent

This r/ollama post addresses a commonly reported behavior where Ollama's GPU usage remains at or near 100% even after a model has finished generating a response. Certain models appear to 'hang' after

I chatted with @ysmulki about MatX, chip design and where silicon designed for LLMs is headed (8:17) Tightly coupling SRAM and HBM on one ch…

HardwareDGX agent

I chatted with @ysmulki about MatX, chip design and where silicon designed for LLMs is headed (8:17) Tightly coupling SRAM and HBM on one chip (14:03) More MoE FLOPS, smaller KV cache load (16:08) Num

Important Update

HardwareDGX agent

Vultr informs customers about a security incident involving unauthorized access to a third-party vendor's marketing data affecting limited customer information. No Vultr systems, customer subscription

ISSCC 2026: NVIDIA & Broadcom CPO, HBM4 & LPDDR6, TSMC Active LSI, Logic-Based SRAM, UCIe-S and More

HardwareDGX agent

ISSCC 2026 previews cutting-edge semiconductor research and industry presentations covering co-packaged optics (CPO) developments from NVIDIA and Broadcom, next-generation memory standards including H

@lmsysorg @sgl_project @vllm_project @CoreWeave @nebiusai @nscale @togethercompute @Togethercompute enables AI labs and enterprises to run f…

HardwareDGX agent

@lmsysorg @sgl_project @vllm_project @CoreWeave @nebiusai @nscale @togethercompute @Togethercompute enables AI labs and enterprises to run frontier models and agents on the NVIDIA Blackwell platform—g

New Adobe Premiere Color Grading Mode Accelerated on NVIDIA GPUs

HardwareDGX agent

The NAB Show 2026 trade show, running April 18-22 in Las Vegas, is set to showcase a wave of new features and optimizations for top video editing applications. Bringing together over 60,000 content pr

Ollama broken on Apple M5 + macOS 26 — Metal shader crash on every model (500 error)

Local AiDGX agent

Ollama consistently fails to run any model on Apple M5 hardware running macOS 26, with the Metal backend failing to initialize and the runner process terminating with a 500 Internal Server Error. The

PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving

HardwareDGX agent

arXiv:2604.12171v1 Announce Type: cross Abstract: Pipeline parallelism (PP) is widely used to partition layers of large language models (LLMs) across GPUs, enabling scalable inference for large models

Q&A with Jensen Huang on Nvidia's supply chain moat, competition from ASICs like Google's TPU, investing in AI labs and neoclouds, selling to China, and more (Dwarkesh Patel/Dwarkesh Podcast)

HardwareDGX agent

Dwarkesh Patel / Dwarkesh Podcast: Q&A with Jensen Huang on Nvidia's supply chain moat, competition from ASICs like Google's TPU, investing in AI labs and neoclouds, selling to China, and more — “If o

Scalable Trajectory Generation for Whole-Body Mobile Manipulation

HardwareDGX agent

arXiv:2604.12565v1 Announce Type: cross Abstract: Robots deployed in unstructured environments must coordinate whole-body motion -- simultaneously moving a mobile base and arm -- to interact with the

Security Researchers - Get Rewarded for Bug Reports!

HardwareDGX agent

Discover how Vultr values and rewards the ethical disclosure of security vulnerabilities with its new bug bounty program. Report issues, get compensated, and help improve Vultr's web ecosystem through

Set up your domains with Vultr DNS!

HardwareDGX agent

Discover Vultr DNS, a free, geographically distributed anycast DNS service that improves browsing speed and reliability by connecting visitors to the closest DNS cluster. Start configuring your domain

Simplify Global Traffic Management with Vultr Global Load Balancers

HardwareDGX agent

Vultr Global Load Balancers is a managed service that distributes incoming traffic across multiple regions and backend servers worldwide, helping businesses improve application availability, reduce la

Social media intern made me a terminally online LEGO 😭

HardwareDGX agent

Dylan Patel, a semiconductor and AI industry analyst known for his work at SemiAnalysis, shared a post on X (formerly Twitter) humorously reacting to a LEGO-style avatar or depiction of himself create

Vultr Easily Meets EU Digital Readiness Operations Act (DORA) Requirements in Independent Assessment

HardwareDGX agent

Vultr, a cloud infrastructure provider, has successfully met the requirements of the EU's Digital Operational Resilience Act (DORA) according to an independent assessment, demonstrating its compliance

Vultr Managed Kafka® Connect Enhancements: Faster and Simpler Data Streaming

HardwareDGX agent

Enhance real-time data streaming with Vultr Managed Apache Kafka®. With expanded Kafka Connect support, data sources and destinations can be quickly integrated with just a few clicks. Get started toda

14 Apr 2026

A Compact and Efficient 1.251 Million Parameter Machine Learning CNN Model PD36-C for Plant Disease Detection: A Case Study

Model ReleasesDGX agent

arXiv:2604.11332v1 Announce Type: cross Abstract: Deep learning has markedly advanced image based plant disease diagnosis as improved hardware and dataset quality have enabled increasingly accurate ne

AIRA_2: Overcoming Bottlenecks in AI Research Agents

HardwareDGX agent

arXiv:2603.26499v2 Announce Type: replace Abstract: Existing research has identified three structural performance bottlenecks in AI research agents: (1) synchronous single-GPU execution constrains sam

Building Custom Atomistic Simulation Workflows for Chemistry and Materials Science with NVIDIA ALCHEMI Toolkit

HardwareDGX agent

NVIDIA's ALCHEMI Toolkit is a PyTorch-native framework that lets researchers build custom GPU-accelerated atomistic simulation workflows for chemistry and materials science applications. The initial r

Can LLMs Reason About Attention? Towards Zero-Shot Analysis of Multimodal Classroom Behavior

HardwareDGX agent

arXiv:2604.03401v2 Announce Type: replace-cross Abstract: Understanding student engagement usually requires time-consuming manual observation or invasive recording that raises privacy concerns. We pre

CASK: Core-Aware Selective KV Compression for Reasoning Traces

HardwareDGX agent

arXiv:2604.10900v1 Announce Type: new Abstract: In large language models performing long-form reasoning, the KV cache grows rapidly with decode length, creating bottlenecks in memory and inference sta

CodeQuant: Unified Clustering and Quantization for Enhanced Outlier Smoothing in Low-Precision Mixture-of-Experts

HardwareDGX agent

arXiv:2604.10496v1 Announce Type: new Abstract: Outliers have emerged as a fundamental bottleneck in preserving accuracy for low-precision large models, particularly within Mixture-of-Experts (MoE) ar

Credo to acquire Israeli silicon photonics startup DustPhotonics for up to $1.3 billion

HardwareDGX agent

Fabless chip design company Credo Technology Group Holding Ltd. today announced that it has reached a deal to acquire Israeli silicon photonics startup DustPhotonics Ltd. for up to 1.3 billion in cash

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

HardwareDGX agent

arXiv:2410.21316v2 Announce Type: replace-cross Abstract: Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of pa

From operational to analytical: The unified Spanner Graph and BigQuery Graph solution

HardwareDGX agent

Understanding the relationships within your data is crucial for uncovering hidden insights and building intelligent applications. However, managing operational (OLTP) and analytical (OLAP) graph workl

Generative Design for Direct-to-Chip Liquid Cooling for Data Centers

HardwareDGX agent

arXiv:2604.10941v1 Announce Type: cross Abstract: Rapid growth in artificial intelligence (AI) workloads is driving up data center power densities, increasing the need for advanced thermal management.

GPU-Accelerated Continuous-Time Successive Convexification for Contact-Implicit Legged Locomotion

HardwareDGX agent

arXiv:2604.09993v1 Announce Type: new Abstract: Contact-implicit trajectory optimization (CITO) enables the automatic discovery of contact sequences, but most methods rely on fine time discretization

IA local con NVIDIA RTX PRO™ 4000 Blackwell 16GB GDDR7

HardwareDGX agent

This Reddit post from the r/ollama community discusses running local AI/LLM workloads using the NVIDIA RTX PRO 4000 Blackwell GPU via Ollama, a framework for running large language models locally. The

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GP…

HardwareDGX agent

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GPU, PyTorch & OS - Multiple kernel versions coexist in one pr

NVIDIA NVbandwidth: Your Essential Tool for Measuring GPU Interconnect and Memory Performance

HardwareDGX agent

NVIDIA NVbandwidth is a performance analysis tool designed to measure memory bandwidth and latency between various components in GPU-equipped systems, providing detailed measurements of data transfer

Nvidia unveils Ising AI models for quantum error correction and calibration

HardwareDGX agent

Technology and computing giant Nvidia Corp. today announced the release of Ising, the world’s first open artificial intelligence model family aimed at quantum computing calibration and error correctio

PA-SFM: Tracker-free differentiable acoustic radiation for freehand 3D photoacoustic imaging

HardwareDGX agent

arXiv:2604.09643v1 Announce Type: new Abstract: Three-dimensional (3D) handheld photoacoustic tomography typically relies on bulky and expensive external positioning sensors to correct motion artifact

Parisians: we're running an open source AI art hackathon with LTX + NVIDIA this Saturday

HardwareDGX agent

A Reddit post on r/StableDiffusion announces an in-person open source AI art hackathon held in Paris, co-organized with LTX (Lightricks' open source AI video model) and NVIDIA. The event likely invite

Read the full research post with NVIDIA: http://cursor.com/blog/multi-agent-kernels

HardwareDGX agent

Cursor has published a research post in collaboration with NVIDIA exploring multi-agent kernel systems, likely detailing how multiple AI agents can work together to handle complex coding or computatio

Real-Time Voicemail Detection in Telephony Audio Using Temporal Speech Activity Features

HardwareDGX agent

arXiv:2604.09675v1 Announce Type: cross Abstract: Outbound AI calling systems must distinguish voicemail greetings from live human answers in real time to avoid wasted agent interactions and dropped c

Relational Probing: LM-to-Graph Adaptation for Financial Prediction

HardwareDGX agent

arXiv:2604.10212v1 Announce Type: new Abstract: Language models can be used to identify relationships between financial entities in text. However, while structured output mechanisms exist, prompting-b

StreamServe: Adaptive Speculative Flows for Low-Latency Disaggregated LLM Serving

HardwareDGX agent

arXiv:2604.09562v1 Announce Type: cross Abstract: Efficient LLM serving must balance throughput and latency across diverse, bursty workloads. We introduce StreamServe, a disaggregated prefill decode s

Sustainable Transformer Neural Network Acceleration with Stochastic Photonic Computing

HardwareDGX agent

arXiv:2604.09759v1 Announce Type: cross Abstract: Transformers achieve state-of-the-art performance in natural language processing, vision, and scientific computing, but demand high computation and me

The Transformation Edge: Why the modern CFO is becoming the enterprise’s AI co-architect

HardwareDGX agent

There are moments in tech when the stack shifts. And then there are moments when everything shifts. This is the latter. After three decades in Silicon Valley — and now building a bridge between Palo A

Tree Training: Accelerating Agentic LLMs Training via Shared Prefix Reuse

HardwareDGX agent

arXiv:2511.00413v4 Announce Type: replace Abstract: Agentic large language model (LLM) training often involves multi-turn interaction trajectories that branch into multiple execution paths due to conc

VTC: DNN Compilation with Virtual Tensors for Data Movement Elimination

HardwareDGX agent

arXiv:2604.09558v1 Announce Type: cross Abstract: With the widening gap between compute and memory operation latencies, data movement optimizations have become increasingly important for DNN compilati

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to ap…

HardwareDGX agent

We've been developing a multi-agent system that builds and maintains complex software autonomously. Recently, we partnered with NVIDIA to apply it to optimizing CUDA kernels. In 3 weeks, it delivered

13 Apr 2026

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference

HardwareDGX agent

arXiv:2604.08584v1 Announce Type: cross Abstract: Long-context LLMs increasingly rely on extended, reusable prefill prompts for agents and domain Q&A, pushing attention and KV-cache to become the domi

Fast Model-guided Instance-wise Adaptation Framework for Real-world Pansharpening with Fidelity Constraints

HardwareDGX agent

arXiv:2604.08903v1 Announce Type: new Abstract: Pansharpening aims to generate high-resolution multispectral (HRMS) images by fusing low-resolution multispectral (LRMS) and high-resolution panchromati

FlashLips: 100-FPS Mask-Free Latent Lip-Sync using Reconstruction Instead of Diffusion or GANs

HardwareDGX agent

arXiv:2512.20033v2 Announce Type: replace Abstract: We present FlashLips, a two-stage, mask-free lip-sync system that decouples lips control from rendering and achieves real-time performance, with our

Free AI Voice Cloning with Qwen3 TTS — Google Colab Notebook (works on free tier, no GPU needed)

HardwareDGX agent

This Reddit post shares a Google Colab notebook that enables free AI voice cloning using Alibaba's Qwen3-TTS, which offers voice cloning, voice design, and ultra-high-quality human-like speech generat

High-dimensional inference for the gamma-ray sky with differentiable programming

HardwareDGX agent

arXiv:2604.08648v1 Announce Type: cross Abstract: We motivate the use of differentiable probabilistic programming techniques in order to account for the large model-space inherent to astrophysical gam

Is an nvidia DGK Spark or similar worth it?

HardwareDGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

Nvidia hasn’t budged in six months. Coreweave is down 21% Oracle is down 50% Microsoft is down 25% Like it or not, the bubble has already st…

HardwareDGX agent

Gary Marcus, an AI skeptic and cognitive scientist, posted on X highlighting diverging stock performance in the AI infrastructure sector, noting that while Nvidia has remained relatively stable, CoreW

Ornn Compute Price Index: renting one Nvidia Blackwell GPU for an hour now costs 4.08, up 48% from 2.75 two months ago, driven by agentic AI demand (Wall Street Journal)

HardwareDGX agent

Wall Street Journal: Ornn Compute Price Index: renting one Nvidia Blackwell GPU for an hour now costs 4.08, up 48% from 2.75 two months ago, driven by agentic AI demand — AI companies are rationing of

← Previous
1…3334353637…75
Next →