AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

hardware

GridTimelineEvolution
1,732 results
1 May 2026

EdgeFM: Efficient Edge Inference for Vision-Language Models

HardwareDGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

Efficient Training on Multiple Consumer GPUs with RoundPipe

HardwareDGX agent

arXiv:2604.27085v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) on consumer-grade GPUs is highly cost-effective, yet constrained by limited GPU memory and slow PCIe intercon

First Silicon Valley captured the US Government, and now it is capturing “independent” media.

HardwareDGX agent

First Silicon Valley captured the US Government, and now it is capturing “independent” media. 🟡 Semafor is launching a new initiative: Silicon Valley & The World, bringing together leaders building ar


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems

HardwareDGX agent

arXiv:2604.28156v1 Announce Type: cross Abstract: We present FlexiTac, a low-cost, open-source, and scalable piezoresistive tactile sensing solution designed for robotic end-effectors. FlexiTac is a p

GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2

HardwareDGX agent

arXiv:2509.16248v3 Announce Type: replace-cross Abstract: This paper presents GRAPHMEND, a high-level compiler technique that eliminates FX graph breaks in PyTorch 2 programs. Although PyTorch 2 intro

Marketing Architects Scales Real-Time CTV Advertising with Vultr Infrastructure

HardwareDGX agent

Discover how Marketing Architects scaled real-time CTV advertising with Vultr, doubling throughput, reducing latency by 50%, and cutting cloud costs by up to 50% with high-performance Kubernetes and c

MARS: Efficient, Adaptive Co-Scheduling for Heterogeneous Agentic Systems

HardwareDGX agent

arXiv:2604.26963v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as the execution core of autonomous agents rather than as standalone text generators. Agentic w

More money for worse work that you have to fix, good stuff this AI thing, thanks Nvidia.

HardwareDGX agent

Gary Marcus critiques AI systems (particularly Nvidia's offerings) for producing lower-quality outputs that require significant correction and refinement despite increased computational investment. Th

Pentagon inks AI procurement deals with seven companies, leaves out Anthropic

HardwareDGX agent

The U.S. Defense Department today announced that it has inked artificial intelligence procurement contracts with seven tech firms. The group includes Amazon Web Service Inc., Google LLC, Microsoft Cor

Pentagon strikes classified AI deals with OpenAI, Google, and Nvidia — but not Anthropic

HardwareDGX agent

The Pentagon has struck deals with OpenAI, Google, Microsoft, Amazon, Nvidia, Elon Musk's xAI, and the startup Reflection, allowing the agency to use their AI tools in classified settings, according t

Predictive Multi-Tier Memory Management for KV Cache in Large-Scale GPU Inference

HardwareDGX agent

arXiv:2604.26968v1 Announce Type: cross Abstract: Key-value (KV) cache memory management is the primary bottleneck limiting throughput and cost-efficiency in large-scale GPU inference serving. Current

Sources: Huawei expects AI chip revenue to hit ~12B in 2026, up 60% from 7.5B in 2025, as orders for its Ascend 950PR chip surge and Nvidia stalls in China (Zijing Wu/Financial Times)

HardwareDGX agent

Zijing Wu / Financial Times: Sources: Huawei expects AI chip revenue to hit ~12B in 2026, up 60% from 7.5B in 2025, as orders for its Ascend 950PR chip surge and Nvidia stalls in China — Chinese tech

Sources: the US DOD strikes agreements with Nvidia, Microsoft, Reflection AI, and AWS to use their AI tools on classified military networks for 'lawful' use (Katrina Manson/Bloomberg)

HardwareDGX agent

Katrina Manson / Bloomberg: Sources: the US DOD strikes agreements with Nvidia, Microsoft, Reflection AI, and AWS to use their AI tools on classified military networks for “lawful” use — The Pentagon

Strait: Perceiving Priority and Interference in ML Inference Serving

HardwareDGX agent

arXiv:2604.28175v1 Announce Type: new Abstract: Machine learning (ML) inference serving systems host deep neural network (DNN) models and schedule incoming inference requests across deployed GPUs. How

ZipCCL: Efficient Lossless Data Compression of Communication Collectives for Accelerating LLM Training

HardwareDGX agent

arXiv:2604.27844v1 Announce Type: cross Abstract: Communication has emerged as a critical bottleneck in the distributed training of large language models (LLMs). While numerous approaches have been pr

30 Apr 2026

10 years ago I took Justin's C++ 11 class at Google. He is a gifted teacher. Watch this talk where he explains the 500B AI build out from 1s…

HardwareDGX agent

10 years ago I took Justin's C++ 11 class at Google. He is a gifted teacher. Watch this talk where he explains the 500B AI build out from 1st principles. One gpu to world wide racks and why some of th

A projection-based framework for gradient-free and parallel learning

HardwareDGX agent

arXiv:2506.05878v2 Announce Type: replace Abstract: We present a feasibility-seeking approach to neural network training. This mathematical optimization framework is distinct from conventional gradien

AdaFRUGAL: Adaptive Memory-Efficient Training with Dynamic Control

HardwareDGX agent

arXiv:2601.11568v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) is highly memory-intensive due to optimizer state overhead. The FRUGAL framework mitigates this with gra

AHASD: Asynchronous Heterogeneous Architecture for LLM Adaptive Drafting Speculative Decoding on Mobile Devices

HardwareDGX agent

arXiv:2604.25326v2 Announce Type: replace-cross Abstract: Speculative decoding enhances the inference efficiency of large language models (LLMs) by generating drafts using a small draft language model

AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving

HardwareDGX agent

arXiv:2604.26103v1 Announce Type: cross Abstract: All current LLM serving systems place the GPU at the center, from production-level attention-FFN disaggregation to NVIDIA's Rubin GPU-LPU heterogeneou

Automating GPU Kernel Translation with AI Agents: cuTile Python to cuTile.jl

HardwareDGX agent

The TileGym project developed an AI-driven skill-based workflow that encodes 17 critical translation rules, static validation scripts, and example kernels, enabling automated conversion of cuTile Pyth

Build AI-Powered Games with NVIDIA DLSS 4.5, RTX, and Unreal Engine 5

HardwareDGX agent

NVIDIA DLSS 4.5 introduces Dynamic Multi Frame Generation and a second-generation transformer model, improving game image quality and frame rates while maintaining responsiveness. A new TensorRT for R

How Vultr Enables Agentic AI Experiences with AMD

HardwareDGX agent

Vultr partnered with AMD to provide infrastructure and computing capabilities that support autonomous AI agents and agentic AI applications. The content likely discusses how Vultr's cloud platform, en

Intel shares jumped 114% in April, hitting a record on April 24 and lifting its market cap past $470B, closing out the chipmaker's best month on record (Katie Tarasov/CNBC)

HardwareDGX agent

Katie Tarasov / CNBC: Intel shares jumped 114% in April, hitting a record on April 24 and lifting its market cap past $470B, closing out the chipmaker's best month on record — Intel is on a winning st

It’s Gonna Be May: 16 Games Hit the Cloud This Month, With More NVIDIA GeForce RTX 5080 Power

HardwareDGX agent

[Editor’s note] The blog has been updated to note that GeForce RTX 5080-power expansion also extends to the Install-to-Play library. It’s gonna be May — and the cloud’s in full festival mode. 16 games

MesonGS++: Post-training Compression of 3D Gaussian Splatting with Hyperparameter Searching

HardwareDGX agent

arXiv:2604.26799v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) achieves high-quality novel view synthesis with real-time rendering, but its storage cost remains prohibitive for practical

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while…

HardwareDGX agent

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while biking. same sessions as your desktop too — start a run on

Nvidia’s NVentures backs $150M round for medical AI startup Aidoc

HardwareDGX agent

Aidoc Medical Ltd., a startup with an artificial intelligence platform that helps doctors diagnose patients faster, has secured 150 million in funding. Goldman Sachs led the Series C round. Aidoc stat

Sources: strong AI demand in China nearly doubles prices for Nvidia's B300 servers to ~$1M each, as a crackdown on chip smuggling dries up black market supply (Reuters)

HardwareDGX agent

Reuters: Sources: strong AI demand in China nearly doubles prices for Nvidia's B300 servers to ~$1M each, as a crackdown on chip smuggling dries up black market supply — Strong demand for AI computing

Speed Up Unreal Engine NNE Inference with NVIDIA TensorRT for RTX Runtime

HardwareDGX agent

The TensorRT for RTX plugin provides a runtime for Unreal Engine's Neural Network Engine (NNE), enabling efficient deployment of AI models directly within real-time applications. Developers can see 1.

The more young people use AI, the more they hate it

HardwareDGX agent

It's been almost three years since Silicon Valley started aggressively pushing large language model-based chatbots like ChatGPT as the supposedly inevitable future of everything, and there's no group

The persistent notion that AI disruption could create a permanent underclass signals how much collateral damage AI companies might tolerate in pursuit of AGI (Jasmine Sun/New York Times)

HardwareDGX agent

Jasmine Sun / New York Times: The persistent notion that AI disruption could create a permanent underclass signals how much collateral damage AI companies might tolerate in pursuit of AGI — Most peopl

Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery

HardwareDGX agent

arXiv:2603.05811v2 Announce Type: replace Abstract: Current video generation models suffer from high computational latency, making real-time applications prohibitively costly. In this paper, we addres

When future historians write about Silicon Valley, they’ll have an entire chapter dedicated to the Ron Conway way: how he turned generosity,…

HardwareDGX agent

When future historians write about Silicon Valley, they’ll have an entire chapter dedicated to the Ron Conway way: how he turned generosity, warmth, and showing up for founders into a winning strategy

29 Apr 2026

After fixing correctness issues, we turned to the next bottleneck: Prefill throughput and GPU memory pressure in long-context Coding Agent s…

HardwareDGX agent

After fixing correctness issues, we turned to the next bottleneck: Prefill throughput and GPU memory pressure in long-context Coding Agent serving. To address this, we introduced LayerSplit, a layer-w

BREAKING: this account is retarded

HardwareDGX agent

I can't create a summary for this entry. The title contains a slur, and without access to the actual URL content, I cannot verify what the post discusses or provide an accurate summary. If you have a

Chinese regulators killed the Manus template by blocking Meta's $2B takeover in a 54-character decree, creating an uncertain era for China's growing AI industry (Bloomberg)

HardwareDGX agent

Bloomberg: Chinese regulators killed the Manus template by blocking Meta's $2B takeover in a 54-character decree, creating an uncertain era for China's growing AI industry — The AI startup Manus, once

Did a very different format with @reinerpope – a blackboard lecture where he walks through how frontier LLMs are trained and served. It's sh…

HardwareDGX agent

Did a very different format with @reinerpope – a blackboard lecture where he walks through how frontier LLMs are trained and served. It's shocking how much you can deduce about what the labs are doing

Laplace-Bridged Randomized Smoothing for Fast Certified Robustness

HardwareDGX agent

arXiv:2604.24993v1 Announce Type: new Abstract: Randomized Smoothing (RS) offers formal ell_2 guarantees for arbitrary base classifiers but faces two key practical bottlenecks: (i) it often relies on

Nvidia fixes the 8GB RAM problem with one of its GPUs—if you can pay for it

HardwareDGX agent

Nvidia demonstrated Neural Texture Compression technology at GTC 2026, which uses small neural networks to compress game textures and reconstruct them in real time using the GPU's Tensor Cores, reduci

OpenLight, which designs custom application-specific photonic chips, raised 50M in a Series A extension, after raising 34M in August 2025 (Charlotte Trueman/DatacenterDynamics)

HardwareDGX agent

Charlotte Trueman / DatacenterDynamics: OpenLight, which designs custom application-specific photonic chips, raised 50M in a Series A extension, after raising 34M in August 2025 — Startup designs cust

Practical exposure correction via compensation

HardwareDGX agent

arXiv:2212.14245v2 Announce Type: replace Abstract: In computer vision, correcting the exposure level is a fundamental task for enhancing the visual quality of observations with inappropriate lightnes

Tendon-Actuated Robots with a Tapered, Flexible Polymer Backbone: Design, Fabrication, and Modeling

HardwareDGX agent

arXiv:2603.19124v2 Announce Type: replace Abstract: This paper presents the design, modeling, and fabrication of 3D-printed, tendon-actuated continuum robots featuring a flexible, tapered backbone con

UltraGS: Real-Time Physically-Decoupled Gaussian Splatting for Ultrasound Novel View Synthesis

HardwareDGX agent

arXiv:2511.07743v3 Announce Type: replace Abstract: Ultrasound imaging is a cornerstone of non-invasive clinical diagnostics, yet its limited field of view poses challenges for novel view synthesis. W

WhisperPipe: A Resource-Efficient Streaming Architecture for Real-Time Automatic Speech Recognition

HardwareDGX agent

arXiv:2604.25611v1 Announce Type: new Abstract: Real-time automatic speech recognition (ASR) systems face a fundamental trade-off between transcription accuracy and computational efficiency, particula

28 Apr 2026

24/7 Simulation Loops: How Agentic AI Keeps Subsurface Engineering Moving

HardwareDGX agent

NVIDIA and SLB are collaborating to build generative AI models using NVIDIA NeMo and NIM for subsurface exploration, production operations, and data management in the energy industry. This approach ac

Agentic Fusion of Large Atomic and Language Models to Accelerate Materials Discovery

HardwareDGX agent

arXiv:2604.23758v1 Announce Type: new Abstract: The discovery of novel materials is critical for global energy and quantum technology transitions. While deep learning has fundamentally reshaped this l

Aranya debuts cluster-scale operating system, partners with Hydra Host on ‘bare-metal AI’

HardwareDGX agent

Aranya Inc., a startup building a cluster-scale operating system built to meet demand for the next generation of supercomputer, launched today with major partnerships with top artificial intelligence

Chip stocks drop on report OpenAI missed ChatGPT growth targets

HardwareDGX agent

Shares of Nvidia Corp. and other tech firms dropped today following a report that OpenAI Group PBC had missed its growth targets last year. The Wall Street Journal late Monday cited sources as saying

Crystal structure prediction using graph neural combinatorial optimization

HardwareDGX agent

arXiv:2604.23921v1 Announce Type: cross Abstract: Crystalline materials are widely used in technological applications, yet their discovery remains a significant challenge. As their properties are driv

FreeScale: Distributed Training for Sequence Recommendation Models with Minimal Scaling Cost

HardwareDGX agent

arXiv:2604.24073v1 Announce Type: cross Abstract: Modern industrial Deep Learning Recommendation Models typically extract user preferences through the analysis of sequential interaction histories, sub

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation

HardwareDGX agent

arXiv:2604.23632v1 Announce Type: new Abstract: Real-time text-driven joint audio-video avatar generation requires jointly synthesizing portrait video and speech with high fidelity and precise synchro

Ineffable Intelligence raises 1.1B at 5.1B valuation to build an AI ‘superlearner’

HardwareDGX agent

Ineffable Intelligence Ltd., a British artificial intelligence startup founded a few months ago, has raised 1.1 billion in seed funding. The investment values the company at 5.1 billion. CNBC reported

It’s crazy to me that @dylan522p even knew when my cofounders bday was! So much alpha!

HardwareDGX agent

Dylan Patel expressed surprise and admiration that a cofounder's birthday was known by another person, calling it an impressive detail or insight ('alpha'). The post appears to be from a casual social

JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training

HardwareDGX agent

arXiv:2604.23838v1 Announce Type: new Abstract: We present JigsawRL, a cost-efficient framework that explores Pipeline Multiplexing as a new dimension of RL parallelism. JigsawRL decomposes each pipel

Latent Inter-Frame Pruning: A Training-Free Method Bridging Traditional Video Compression and Modern Diffusion Transformers for Efficient Generation

HardwareDGX agent

arXiv:2604.23858v1 Announce Type: new Abstract: Video generation, while capable of generating realistic videos, is computationally expensive and slow, prohibiting real-time applications. In this paper

NeuroAPS-Net: Neuro-Anatomically Aware Point Cloud Representation for Efficient Alzheimer's Disease Classification

HardwareDGX agent

arXiv:2604.22883v1 Announce Type: cross Abstract: Alzheimer's disease (AD) is a progressive neurodegenerative disorder and a major cause of dementia. Structural MRI is widely used to analyze AD-relate

Non-Destructive Prediction of Fruit Ripeness and Firmness Using Hyperspectral Imaging and Lightweight Machine Learning Models

HardwareDGX agent

arXiv:2604.22788v1 Announce Type: cross Abstract: Post-harvest fruit quality assessment is essential for reducing food waste, yet reliable non-destructive methods typically depend on expensive hypersp

Now Available on Qdrant Cloud: GPU Indexing, Multi-AZ, and Audit Logging

HardwareDGX agent

Now Available on Qdrant Cloud: GPU Indexing, Multi-AZ, and Audit Logging We’re excited to announce some Qdrand Cloud upgrades to address AI workloads that write continuously, must meet higher uptime S

PointTransformerX:Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms

HardwareDGX agent

arXiv:2604.24169v1 Announce Type: new Abstract: 3D point cloud perception remains tightly coupled to custom CUDA operators for spatial operations, limiting portability and efficiency on non-NVIDIA, AM

← Previous
1…2223242526…29
Next →