AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
7 May 2026

Full-chip CMP modelling based on Fully Convolutional Network leveraging White Light Interferometry

ResearchDGX agent

arXiv:2605.05062v1 Announce Type: new Abstract: As time-to-market is crucial in the Integrated Circuit (IC) industry, speeding up layout manufacturability verifi-cation is essential. Chemical-Mechanic

Model Quantization: Post-Training Quantization Using NVIDIA Model Optimizer

Local AiDGX agent

Post-training quantization is a technique that reduces model size and improves inference performance by converting weights and activations to lower precision formats after training is complete. NVIDIA

Quantization and Fast Inference (MEAP) - How much performance are you actually getting from quantization in production? [D]

ApplicationsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

This discussion examines the practical performance gains of quantization techniques when deployed in production environments, particularly addressing whether theoretical speedups translate to real-wor

Robots don't need a human face. Instead of building a humanoid face, we use a real human face as a controller for Reachy Mini's existing non…

IndustryDGX agent

Researchers developed a control method for Reachy Mini robots that uses a real human face as an input controller rather than building a humanoid facial interface. This approach leverages existing robo

Topology-Constrained Quantized nnUNet for Efficient and Anatomically Accurate 3D Tooth Segmentation

ApplicationsDGX agent

arXiv:2605.04201v1 Announce Type: new Abstract: We propose a topology-constrained quantized nnUNet framework for efficient and anatomically accurate 3D tooth segmentation, addressing the challenges of

We started running monthly research sessions at the @agi_inc office. First one was last Friday on PNM compilation. If you do on-device resea…

Local AiDGX agent

We started running monthly research sessions at the @agi_inc office. First one was last Friday on PNM compilation. If you do on-device research and want to come present, apply here: https://form.typef

6 May 2026

b9041

Local AiDGX agent

llama.cpp b9041 is a release build of llama.cpp, an LLM inference library written in C/C++ . The build identifier indicates this is an intermediate development release from the ggml-org/llama.cpp proj

Best value upgrade path from 12GB VRAM RTX4080, 16GB system RAM gaming laptop for local LLM inference?

Local AiDGX agent

This post discusses upgrade recommendations for improving local LLM inference performance on a gaming laptop with an RTX 4080 (12GB VRAM) and 16GB system RAM. Community members likely shared options f

BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation

SafetyDGX agent

arXiv:2605.03452v1 Announce Type: new Abstract: High-quality data collection is a fundamental cornerstone for training humanoid whole-body visuomotor policies. Current data acquisition paradigms predo

EMOVIS: Emotion-Optimized Image Processing

ResearchDGX agent

arXiv:2605.03131v1 Announce Type: cross Abstract: In cinematography, visual attributes such as color grading, contrast, and brightness are manipulated to reinforce the emotional narrative of a scene.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving

AgentsDGX agent

arXiv:2605.00831v1 Announce Type: cross Abstract: The rise of million-token, agent-based applications has placed unprecedented demands on large language model (LLM) inference services. The long-runnin

Hello again, everyone! Our latest Qwopus3.6-35B-A3B-v1 is now live, and it is once again breathtaking! Full HF space benchmark showcase and …

Model ReleasesDGX agent

Hello again, everyone! Our latest Qwopus3.6-35B-A3B-v1 is now live, and it is once again breathtaking! Full HF space benchmark showcase and write-up is in the comments, so you can make conclusions for

Local Truncation Error-Guided Neural ODEs for Large Scale Traffic Forecasting

Local AiDGX agent

arXiv:2605.03386v1 Announce Type: new Abstract: Spatiotemporal forecasting in physical systems, such as large-scale traffic networks, requires modeling a dual dynamic: continuous macroscopic rhythms a

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses

SafetyDGX agent

arXiv:2605.02900v1 Announce Type: cross Abstract: Embodied Artificial Intelligence (Embodied AI) integrates perception, cognition, planning, and interaction into agents that operate in open-world, saf

SHIELD: A Diverse Clinical Note Dataset and Distilled Small Language Models for Enterprise-Scale De-identification

Local AiDGX agent

arXiv:2605.03301v1 Announce Type: new Abstract: De-identification of clinical text remains essential for secondary use of electronic health records (EHRs), yet public benchmarks such as i2b2 2006/2014

SMoE: An Algorithm-System Co-Design for Pushing MoE to the Edge via Expert Substitution

SafetyDGX agent

arXiv:2508.18983v3 Announce Type: replace Abstract: The Mixture of Experts (MoE) architecture has emerged as a key technique for scaling Large Language Models by activating only a subset of experts pe

SpaceX Super Heavy Booster 19 rolls by as it heads to the launch site for testing

IndustryDGX agent

SpaceX's Super Heavy Booster 19 was transported to the launch site for testing purposes, as documented in a post by Elon Musk on X. This represents progress in SpaceX's development of the Starship lau

StateSMix: Online Lossless Compression via Mamba State Space Models and Sparse N-gram Context Mixing

Model ReleasesDGX agent

arXiv:2605.02904v1 Announce Type: new Abstract: We present StateSMix, a fully self-contained lossless compressor that couples an online-trained Mamba-style State Space Model (SSM) with sparse n-gram c

5 May 2026

Action Agent: Agentic Video Generation Meets Flow-Constrained Diffusion

Model ReleasesDGX agent

arXiv:2605.01477v1 Announce Type: new Abstract: We present Action Agent, a two-stage framework that unifies agentic navigation video generation with flow-constrained diffusion control for multi-embodi

Adaptation of AI-accelerated CFD Simulations to the IPU platform

TutorialsDGX agent

arXiv:2605.00462v1 Announce Type: cross Abstract: Intelligence Processing Units (IPU) have proven useful for many AI applications. In this paper, we evaluate them within the emerging field of AI for s

b9028

Local AiDGX agent

B9028 is a release build from the llama.cpp project, which provides LLM inference in C/C++. The llama.cpp project uses sequential build numbers (like b9028) to track intermediate releases of its LLM i

b9030

Local AiDGX agent

B9030 is a release build of llama.cpp, a C/C++ implementation for LLM inference . This build release would contain bug fixes, performance improvements, and new features added to the llama.cpp project

b9031

Local AiDGX agent

Release b9031 of llama.cpp optimizes backend loading by only loading backends when required, as specified in pull request #22290 . The release includes binary builds for multiple platforms including m

b9033

Local AiDGX agent

Release b9033 is a build of llama.cpp with a sync of ggml , published on May 5, 2026. The release provides compiled binaries across multiple platforms including macOS, Linux, Android, Windows, and ope

BadmintonGRF: A Multimodal Dataset and Benchmark for Markerless Ground Reaction Force Estimation in Badminton

Model ReleasesDGX agent

arXiv:2605.01876v1 Announce Type: new Abstract: Multimodal resources for non-periodic court sports with laboratory-grade sensing remain scarce: few publicly pair instrumented ground reaction force (GR

Barren Plateaus as Destructive Interference: A Diagnostic Framework and Implications for Structured Ansatzes

ResearchDGX agent

arXiv:2605.01319v1 Announce Type: cross Abstract: Barren plateaus (BPs) are usually described by the exponential suppression of gradient variance, but the mechanism by which gradient signal disappears

Bi-Level Reinforcement Learning Control for an Underactuated Blimp via Center-of-Mass Reconfiguration

SafetyDGX agent

arXiv:2605.01289v1 Announce Type: new Abstract: This paper investigates goal-directed tracking control of underactuated blimps with center-of-mass (CoM) reconfiguration. Unlike conventional overactuat

Cloud Engineer’s AI Toolkit: Sign up Now for a Developer Workshop Near You!

Model ReleasesDGX agent

The world of AI is rapidly shifting from experimental Large Language Models to an era of Agentic AI. In the agentic era, autonomous software agents act on behalf of employees and consumers—driving a f

CoFrGeNet: Continued Fraction Architectures for Language Generation

ResearchDGX agent

arXiv:2601.21766v3 Announce Type: replace Abstract: Transformers are arguably the preferred architecture for language generation. In this paper, inspired by continued fractions, we introduce a new fun

CoRAL: Contact-Rich Adaptive LLM-based Control for Robotic Manipulation

ApplicationsDGX agent

arXiv:2605.02600v1 Announce Type: new Abstract: While Large Language Models (LLMs) and Vision-Language Models (VLMs) demonstrate remarkable capabilities in high-level reasoning and semantic understand

CycleRL: Sim-to-Real Deep Reinforcement Learning for Robust Autonomous Bicycle Control

SafetyDGX agent

arXiv:2603.15013v2 Announce Type: replace Abstract: Autonomous bicycles offer a promising agile solution for urban mobility and last-mile logistics. However, conventional control strategies often stru

DBLP: Phase-Aware Bounded-Loss Transport for Burst-Resilient Distributed ML Training

Model ReleasesDGX agent

arXiv:2605.01989v1 Announce Type: new Abstract: Distributed machine learning (ML) training has become a necessity with the prevalence of billion to trillion-parameter-scale models. While prior work ha

From Characterization To Construction: Generative Quantum Circuit Synthesis from Gate Set Tomography Data

ResearchDGX agent

arXiv:2605.01367v1 Announce Type: cross Abstract: High-fidelity circuit execution on noisy intermediate-scale quantum devices is bottlenecked by compilation pipelines that disregard complex, correlate

HandelBot: Real-World Piano Playing via Fast Adaptation of Dexterous Robot Policies

SafetyDGX agent

arXiv:2603.12243v3 Announce Type: replace Abstract: Mastering dexterous manipulation with multi-fingered hands has been a grand challenge in robotics for decades. Despite its potential, the difficulty

Hybrid Quantum Reinforcement Learning with QAOA for Improved Vehicle Routing Optimization

SafetyDGX agent

arXiv:2605.01574v1 Announce Type: new Abstract: Vehicle Routing Problem (VRP) is one of the most complex NP-hard combinatorial optimization problem in transportation and logistics that requires a dyna

Learning Equivariant Neural-Augmented Object Dynamics From Few Interactions

TutorialsDGX agent

arXiv:2605.02699v1 Announce Type: cross Abstract: Learning data-efficient object dynamics models for robotic manipulation remains challenging, especially for deformable objects. A popular approach is

Machine Learning Enhanced Laser Spectroscopy for Multi-Species Gas Detection in Complex and Harsh Environments

SafetyDGX agent

arXiv:2605.01306v1 Announce Type: cross Abstract: Laser absorption spectroscopy (LAS) is a well-established technique for non-intrusive measurement of gas species in combustion and atmospheric environ

MolmoAct2: Action Reasoning Models for Real-world Deployment

Model ReleasesDGX agent

arXiv:2605.02881v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models aim to provide a single generalist controller for robots, but today's systems fall short on the criteria that matter

MultiSense-Pneumo: A Multimodal Learning Framework for Pneumonia Screening in Resource-Constrained Settings

ApplicationsDGX agent

arXiv:2605.02207v1 Announce Type: new Abstract: Pneumonia remains a leading global cause of morbidity and mortality, particularly in low resource settings where access to imaging, laboratory testing,

OphMAE: Bridging Volumetric and Planar Imaging with a Foundation Model for Adaptive Ophthalmological Diagnosis

Model ReleasesDGX agent

arXiv:2605.02714v1 Announce Type: new Abstract: The advent of foundation models has heralded a new era in medical artificial intelligence (AI), enabling the extraction of generalizable representations

Real-Time Text Transmission via LLM-Based Entropy Coding over Fixed-Rate Channels

Model ReleasesDGX agent

arXiv:2605.01991v1 Announce Type: cross Abstract: Learning, prediction, and compression are intimately connected: a model that accurately predicts the next symbol in a sequence can be coupled with a s

ShiftLIF: Efficient Multi-Level Spiking Neurons with Power-of-Two Quantization

ResearchDGX agent

arXiv:2605.01866v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are promising for edge sensing due to their event-driven computation and temporal filtering capability. However, standa

SRGAN-CKAN: Expressive Super-Resolution with Nonlinear Functional Operators under Minimal Resources

ResearchDGX agent

arXiv:2605.01459v1 Announce Type: new Abstract: Single-Image Super-Resolution (SISR) aims to reconstruct a High-Resolution (HR) image from a Low-Resolution (LR) observation, a fundamentally ill-posed

TAIL-Safe: Task-Agnostic Safety Monitoring for Imitation Learning Policies

SafetyDGX agent

arXiv:2605.01195v1 Announce Type: new Abstract: Recent imitation learning (IL) algorithms such as flow-matching and diffusion policies demonstrate remarkable performance in learning complex manipulati

TF1-EN-3M: Three Million Synthetic Moral Fables for Training Small, Open Language Models

Model ReleasesDGX agent

arXiv:2504.20605v2 Announce Type: replace Abstract: Moral stories are a time-tested vehicle for transmitting values, yet modern NLP lacks a large, structured corpus that couples coherent narratives wi

Toward a Scientific Discovery Engine for Weather and Climate Data: A Visual Analytics Workbench for Embedding-Based Exploration

SafetyDGX agent

arXiv:2605.00972v1 Announce Type: cross Abstract: Earth system science is producing increasingly large, high-dimensional datasets from physics based Earth system models to AI-based weather and climate

4 May 2026

A decade of quantum on the cloud

ResearchDGX agent

IBM Research reflects on ten years of quantum computing development and deployment through cloud-based platforms, documenting the evolution from early experimental systems to more practical quantum pr

AMD is adding HDMI 2.1 support for Linux. That's good news for the Steam Machine.

IndustryDGX agent

AMD has announced HDMI 2.1 support for Linux and SteamOS, enabling higher bandwidth for 4K 120Hz gaming and variable refresh rate (VRR) . The initial update includes HDMI Fixed Rate Link but excludes

Autoformalizing Memory Specifications with Agents

ApplicationsDGX agent

arXiv:2605.00058v1 Announce Type: cross Abstract: The primary goal of Design Verification (DV) is to ensure that a proposed chip design implementation (either in code, or physical form) exactly matche

b9016

Local AiDGX agent

Release b9016 is a build of llama.cpp, an LLM inference implementation in C/C++ . This build designation represents a specific intermediate version in the project's development cycle, similar to other

b9019

Local AiDGX agent

b9019 is a release of llama.cpp, a C/C++ implementation for LLM inference. As a release tag from the ggml-org/llama.cpp repository, it represents a specific commit or version of the project that inclu

b9020

Local AiDGX agent

b9020 is a build release of llama.cpp (a C/C++ implementation of LLM inference) released on May 4, 2026, with precompiled binaries available for multiple platforms including macOS, Android, OpenEuler,

b9022

Local AiDGX agent

B9022 is a release of llama.cpp created on May 4, 2026, with commit d8794ee signed with GitHub's verified signature. The llama.cpp project is the main playground for developing new features for the gg

b9025

Local AiDGX agent

b9025 is a release from llama.cpp, a C/C++ project for LLM inference . The release uses a build numbering system for intermediate versions of the library. As a llama.cpp release, it likely includes up

BlenderRAG: High-Fidelity 3D Object Generation via Retrieval-Augmented Code Synthesis

SafetyDGX agent

arXiv:2605.00632v1 Announce Type: new Abstract: Automatic generation of executable Blender code from natural language remains challenging, with state-of-the-art LLMs producing frequent syntactic error

Caracal: Causal Architecture via Spectral Mixing

Model ReleasesDGX agent

arXiv:2605.00292v1 Announce Type: new Abstract: The scalability of Large Language Models to long sequences is hindered by the quadratic cost of attention and the limitations of positional encodings. T

Cloud Is Closer Than It Appears: Revisiting the Tradeoffs of Distributed Real-Time Inference

Local AiDGX agent

arXiv:2605.00005v1 Announce Type: new Abstract: The increasing deployment of deep neural networks (DNNs) in cyber-physical systems (CPS) enhances perception fidelity, but imposes substantial computati

Conformalized Quantum DeepONet Ensembles for Scalable Operator Learning with Distribution-Free Uncertainty

SafetyDGX agent

arXiv:2605.00330v1 Announce Type: new Abstract: Operator learning enables fast surrogate modeling of high-dimensional dynamical systems, but existing approaches face two fundamental limitations: quadr

Foundational research powering efficient inference at scale

ToolsDGX agent

This article from Together AI discusses foundational research techniques and methodologies used to enable efficient large-scale language model inference. It likely covers optimization strategies, hard

From Images2Mesh: A 3D Surface Reconstruction Pipeline for Non-Cooperative Space Objects

Model ReleasesDGX agent

arXiv:2605.00147v1 Announce Type: new Abstract: On-orbit inspection imagery is crucial as it enables characterization of non-cooperative resident space objects, providing the geometry and structural c

← Previous
1…6566676869…75
Next →