AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,450 results
15 May 2026

Winning Lottery Tickets in Neural Networks via a Quantum-Inspired Classical Algorithm

ResearchDGX agent

arXiv:2605.13979v1 Announce Type: cross Abstract: Quantum machine learning (QML) aims to accelerate machine learning tasks by exploiting quantum computation. Previous work studied a QML algorithm for

14 May 2026

b9140

Local AiDGX agent

The search did not return specific information about the b9140 release. Based on the context, b9140 is a specific build number from the llama.cpp project releases. llama.cpp is an LLM inference implem

b9144

Local AiDGX agent

llama.cpp is an open source software library that performs inference on various large language models such as Llama. Build b9144 is an intermediate release version of the llama.cpp project from the gg

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cisco announces record revenue and 4,000 layoffs in the same day

IndustryDGX agent

Cisco announced nearly 4,000 job cuts as part of a restructuring to shift investment toward artificial intelligence while simultaneously raising its annual revenue forecast after strong hyperscaler de

Comparing tokens per second of common models

Local AiDGX agent

This Reddit post likely compares inference performance metrics across popular language models running on Ollama, measuring tokens per second as a key performance indicator. The post would help users a

DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense

SafetyDGX agent

arXiv:2605.13115v1 Announce Type: cross Abstract: Diffusion models depend on pseudo-random number generators (PRNGs) for latent noise sampling. We present DiffusionHijack, a supply-chain backdoor atta

Fuel for the AI factory: Data orchestration becomes Dell’s guiding star in the new age of machine intelligence

IndustryDGX agent

Dell Technologies Inc.’s launch of the AI Factory two years ago this month was positioned by the company as “advancements to manage and protect the data that fuels AI innovation.” This one phrase enca

Geometric Preconditioning and Curriculum Optimization for Trainable Variational Quantum Regression

Model ReleasesDGX agent

arXiv:2601.11942v3 Announce Type: replace Abstract: Variational quantum circuits are increasingly studied as continuous-function approximators, but quantum regression remains difficult to train when g

Manipulation Planning for Construction Activities with Repetitive Tasks

ResearchDGX agent

arXiv:2605.13754v1 Announce Type: new Abstract: In this paper, we study the problem of manipulation skill acquisition for performing construction activities consisting of repetitive tasks (e.g., build

MaskPro: Linear-Space Probabilistic Learning for Strict (N:M)-Sparsity on LLMs

SafetyDGX agent

arXiv:2506.12876v2 Announce Type: replace Abstract: The rapid scaling of large language models~(LLMs) has made inference efficiency a primary bottleneck in the practical deployment. To address this, s

N-vium: Mixture-of-Exits Transformer for Accelerated Exact Generation

Model ReleasesDGX agent

arXiv:2605.13190v1 Announce Type: cross Abstract: Improving the inference efficiency of autoregressive transformers typically means reducing FLOPs per token, usually through approximations that degrad

OpenAaaS: An Open Agent-as-a-Service Framework for Distributed Materials-Informatics Research

AgentsDGX agent

arXiv:2605.13618v1 Announce Type: cross Abstract: The Materials Genome Initiative catalyzed the proliferation of centralized platforms--SaaS, PaaS, and IaaS--that aggregate computational and experimen

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention

Model ReleasesDGX agent

arXiv:2605.13473v1 Announce Type: new Abstract: Linear attention and state-space models offer constant-memory alternatives to softmax attention, but often struggle with in-context associative recall.

SymbolSight: Minimizing Inter-Symbol Interference for Reading with Prosthetic Vision

ResearchDGX agent

arXiv:2601.17326v2 Announce Type: replace Abstract: Retinal prostheses restore limited visual perception, but low spatial resolution and temporal persistence make reading difficult. In sequential lett

TouchAnything: A Dataset and Framework for Bimanual Tactile Estimation from Egocentric Video

Model ReleasesDGX agent

arXiv:2605.13083v1 Announce Type: new Abstract: Egocentric human video data, which captures rich human-environment interactions and can be collected at scale, has become a key driver of embodied intel

When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering

SafetyDGX agent

arXiv:2602.22474v2 Announce Type: replace-cross Abstract: Policy steering is an emerging way to adapt robot behaviors at deployment-time: a learned verifier analyzes low-level action samples proposed

13 May 2026

Adversarial Effects on Expressibility and Trainability in Distributed Variational Quantum Algorithms

ResearchDGX agent

arXiv:2605.03629v1 Announce Type: cross Abstract: Distributed quantum algorithms offer a promising pathway to scale variational quantum algorithms beyond the constraints of noisy intermediate-scale qu

b9134

Local AiDGX agent

B9134 is a recent build release of llama.cpp, an open-source C/C++ library for efficient large language model inference. The release was tagged on May 13, 2026, representing an ongoing development ite

Bin Latent Transformer (BiLT): A shift-invariant autoencoder for calibration-free spectral unmixing of turbid media

Model ReleasesDGX agent

arXiv:2605.11829v1 Announce Type: cross Abstract: The accurate recovery of constituent-level optical properties from integrating sphere measurements is a central analytical challenge in pharmaceutical

Coordinated Diffusion: Generating Multi-Agent Behavior Without Multi-Agent Demonstrations

SafetyDGX agent

arXiv:2605.11485v1 Announce Type: new Abstract: Imitation learning powered by generative models has proven effective for modeling complex single-agent behaviors. However, teaching multi-agent systems,

http://blogs.nvidia.com/blog/rtx-ai-garage-hermes-agent-dgx-spark

Local AiDGX agent

NVIDIA's RTX AI Garage partnership with Nous Research demonstrates deploying Hermes agent models on DGX systems integrated with Apache Spark for accelerated AI workloads. The initiative showcases how

Improving the Performance and Learning Stability of Parallelizable RNNs Designed for Ultra-Low Power Applications

ResearchDGX agent

arXiv:2605.11855v1 Announce Type: new Abstract: Sequence learning is dominated by Transformers and parallelizable recurrent neural networks (RNNs) such as state-space models, yet learning long-term de

LTX 2.3 INT8 Benchmarks (2x Faster on Ampere)

Local AiDGX agent

LTX-2.3 INT8 benchmarks discuss INT8 quantization optimized for Ampere GPUs (RTX 30XX series), which offers a balance between speed, VRAM usage, and quality. These INT8 models are designed to speed up

Measuring Accuracy and Energy-to-Solution of Quantum Fine-Tuning of Foundational AI Models

ApplicationsDGX agent

arXiv:2605.02798v1 Announce Type: cross Abstract: We present an experimental study of energy-to-solution (ETS) of hybrid quantum-classical applications, enabled by direct instrumentation of power cons

Q&A with Amazon SVP of Devices and Alexa Panos Panay on Alexa+, the company's focus on devices for the home, its Leo satellite broadband business, and more (Rafe Rosner-Uddin/Financial Times)

IndustryDGX agent

Rafe Rosner-Uddin / Financial Times: Q&A with Amazon SVP of Devices and Alexa Panos Panay on Alexa+, the company's focus on devices for the home, its Leo satellite broadband business, and more — Execu

Real-Time Whole-Body Teleoperation of a Humanoid Robot Using IMU-Based Motion Capture with Sim2Sim and Sim2Real Validation

ResearchDGX agent

arXiv:2605.12347v1 Announce Type: new Abstract: Stable, low-latency whole-body teleoperation of humanoid robots is an open research challenge, complicated by kinematic mismatches between human and rob

Rethink the Role of Neural Decoders in Quantum Error Correction

SafetyDGX agent

arXiv:2605.12046v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for enabling quantum advantages, with decoding as a central algorithmic primitive. Owing to its importance

Rivian adds a new onboard AI assistant to its latest software update

IndustryDGX agent

Rivian released its new Rivian Assistant AI feature as part of software update 2026.15, enabling owners to interact more naturally with their vehicles through voice commands. Unlike phone-mirroring vo

ShardTensor: Domain Parallelism for Scientific Machine Learning

ResearchDGX agent

arXiv:2605.11111v1 Announce Type: cross Abstract: Scientific Machine Learning (SciML) faces unique challenges for extreme-resolution data, with mitigations that often fail to scale or degrade the accu

Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM

ResearchDGX agent

arXiv:2505.05772v2 Announce Type: replace Abstract: Transformer-based models are the foundation of modern machine learning, but their execution, particularly during autoregressive decoding in large la

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

SafetyDGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

12 May 2026

A Paired Point-of-Care Ultrasound Dataset for Image Quality Enhancement and Benchmarking via a cGAN Baseline

ResearchDGX agent

arXiv:2605.08282v1 Announce Type: cross Abstract: Purpose: We aim to enhance the image quality of point-of-care ultrasound (POCUS) devices using deep learning and a novel paired dataset of POCUS and h

A Physical Theory of Backpropagation: Exact Gradients from the Least-Action Principle

Local AiDGX agent

arXiv:2602.02281v2 Announce Type: replace-cross Abstract: Backpropagation is typically presented as a symbolic procedure: a backward pass topologically distinct from inference, with non-local error si

A Quantum Inspired Variational Kernel and Explainable AI Framework for Cross Region Solar and Wind Energy Forecasting

Model ReleasesDGX agent

arXiv:2605.09032v1 Announce Type: cross Abstract: Reliable short horizon forecasting of solar and wind generation is a structural prerequisite of any modern power system yet most published forecasters

Adaptive DNN Partitioning and Offloading in Heterogeneous Edge-Cloud Continuum

ResearchDGX agent

arXiv:2605.09623v1 Announce Type: cross Abstract: In recent years, the use of artificial intelligence on resource-constrained IoT devices has grown significantly. However, existing approaches to DNN p

b9123

Local AiDGX agent

b9123 is a release of llama.cpp created on May 12, 2026 . llama.cpp is an open source software library that performs inference on various large language models such as Llama , providing LLM inference

BEACON: A Multimodal Dataset for Learning Behavioral Fingerprints from Gameplay Data

Model ReleasesDGX agent

arXiv:2605.10867v1 Announce Type: cross Abstract: Continuous authentication in high-stakes digital environments requires datasets with fine-grained behavioral signals under realistic cognitive and mot

Contextual Plackett-Luce: An Efficient Neural Model for Probabilistic Sequence Selection under Ambiguity

ResearchDGX agent

arXiv:2605.09112v1 Announce Type: cross Abstract: Selecting a coherent sequence or subset of elements is a fundamental problem in structured prediction, arising in tasks such as detection, trajectory

Counterfactual Stress Testing for Image Classification Models

ApplicationsDGX agent

arXiv:2605.10894v1 Announce Type: new Abstract: Deep learning models in medical imaging often fail when deployed in new clinical environments due to distribution shifts in demographics, scanner hardwa

cuRegOT: A GPU-Accelerated Solver for Entropic-Regularized Optimal Transport

Model ReleasesDGX agent

arXiv:2605.08793v1 Announce Type: cross Abstract: Optimal transport (OT) has emerged as a fundamental tool in modern machine learning, yet its computational cost remains a significant bottleneck for l

DECO: Sparse Mixture-of-Experts with Dense-Comparable Performance on End-Side Devices

Model ReleasesDGX agent

arXiv:2605.10933v1 Announce Type: cross Abstract: While Mixture-of-Experts (MoE) scales model capacity without proportionally increasing computation, its massive total parameter footprint creates sign

Dolphin-CN-Dialect: Where Chinese Dialects Matter

ApplicationsDGX agent

arXiv:2605.08961v1 Announce Type: new Abstract: We present Dolphin-CN-Dialect, a streaming-capable ASR model with a focus on Chinese and dialect-rich scenarios. Compared to the previous version, Dolph

Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts

ApplicationsDGX agent

arXiv:2509.21892v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) models typically fix the number of activated experts k at both training and inference. However, real-world deployment

enclawed: A Configurable, Sector-Neutral Hardening Framework for Single-User AI Assistant Gateways

ApplicationsDGX agent

arXiv:2604.16838v2 Announce Type: replace-cross Abstract: We present enclawed, a hard-fork hardening framework built on the OpenClaw AI assistant gateway. enclawed targets deployments that need attest

End-to-End Keyword Spotting on FPGA Using Graph Neural Networks with a Neuromorphic Auditory Sensor

Local AiDGX agent

arXiv:2605.09570v1 Announce Type: new Abstract: With the rapid growth of mobile robotics and embedded intelligence, there is an increasing demand for efficient on-device data processing on edge platfo

Energy-based models for diagnostic reconstruction and analysis in a laboratory plasma device

ResearchDGX agent

arXiv:2605.08645v1 Announce Type: cross Abstract: Energy-based models (EBMs) provide a powerful and flexible way of learning a joint probability distribution over data by constructing an energy surfac

EROAS: 3D Efficient Reactive Obstacle Avoidance System for Autonomous Underwater Vehicles using 2.5D Forward-Looking Sonar

SafetyDGX agent

arXiv:2411.05516v3 Announce Type: replace Abstract: Autonomous Underwater Vehicles (AUVs) have advanced significantly in obstacle detection and path planning through sonar, cameras, and learning-based

Finer is Better (with the Right Scaling)

ResearchDGX agent

arXiv:2605.08565v1 Announce Type: new Abstract: Microscaling is a critical technique for preserving the quality of Large Language Models (LLMs) quantized to ultra-low precision formats. Intuitively, f

FLARE: One-Shot PE-Level Fault Localization in Systolic Arrays via Algebraic Test Vectors

Local AiDGX agent

arXiv:2605.08594v1 Announce Type: cross Abstract: Systolic arrays are the dominant compute fabric for neural network inference. Prior work has addressed column-level fault detection efficiently with u

Four Over Six: More Accurate NVFP4 Quantization with Adaptive Block Scaling

Model ReleasesDGX agent

arXiv:2512.02010v5 Announce Type: replace Abstract: As large language models have grown larger, interest has grown in low-precision numerical formats such as NVFP4 as a way to improve speed and reduce

Generalization Bounds of Emergent Communications for Agentic AI Networking

AgentsDGX agent

arXiv:2605.08613v1 Announce Type: new Abstract: The evolution of 6G networking toward agentic AI networking (AgentNet) systems requires a shift from traditional data pipelines to task-aware, agentic A

How to achieve truly serverless GPUs

TutorialsDGX agent

This article explains how to implement and deploy GPU workloads in a truly serverless architecture, likely covering Modal's approach to abstracting away infrastructure management while providing GPU c

IMPACT: An Implicit Active-Set Augmented Lagrangian for Fast Contact-Implicit Trajectory Optimization

ResearchDGX agent

arXiv:2605.09127v1 Announce Type: new Abstract: Contact-implicit trajectory optimization (CITO) has attracted growing attention as a unified framework for planning and control in contact-rich robotic

Kaczmarz Linear Attention

ResearchDGX agent

arXiv:2605.08587v1 Announce Type: cross Abstract: Long-context language modeling remains central to modern sequence modeling, but the quadratic cost of Transformer attention makes scaling computationa

Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers

SafetyDGX agent

arXiv:2410.14022v2 Announce Type: replace-cross Abstract: Human dexterity arises from combining high-level task reasoning with finger-level dexterity control and physical compliance at the muscle and

LiteMedCoT-VL: Parameter-Efficient Adaptation for Medical Visual Question Answering

Model ReleasesDGX agent

arXiv:2605.09384v1 Announce Type: cross Abstract: The reasoning gap between large and compact vision-language models (VLMs) limits the deployment of medical AI on portable clinical devices. Compact VL

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

Model ReleasesDGX agent

arXiv:2605.10777v1 Announce Type: new Abstract: The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling the

Loom: Hybrid Retrieval-Scoring Outfit Recommendation with Semantic Material Compatibility and Occasion-Aware Embedding Priors

ResearchDGX agent

arXiv:2605.09830v1 Announce Type: cross Abstract: We present Loom, an outfit recommendation system that combines neural embedding retrieval with structured domain scoring to generate complete, coheren

Markerless Head Tracking for Accurate and Accessible Neuronavigation

TutorialsDGX agent

arXiv:2602.07052v2 Announce Type: replace Abstract: Neuronavigation is widely used in biomedical research and interventions to guide the precise placement of instruments around the head to support pro

MicroViTv2: Beyond the FLOPS for Edge Energy-Friendly Vision Transformers

ResearchDGX agent

arXiv:2605.10148v1 Announce Type: new Abstract: The Vision Transformer (ViT) achieves remarkable accuracy across visual tasks but remains computationally expensive for edge deployment. This paper pres

← Previous
1…6364656667…75
Next →