AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,457 results
5 Jun 2026

HANDOFF: Humanoid Agentic Task-Space Whole-Body Control via Distilled Complementary Teachers

SafetyDGX agent

arXiv:2606.06493v1 Announce Type: new Abstract: For a humanoid robot to be deployed in the real world, the choice of command space (i.e., the interface between task planning and whole-body control) is

LadderMan: Learning Humanoid Perceptive Ladder Climbing

SafetyDGX agent

arXiv:2606.05873v1 Announce Type: cross Abstract: Humanoid robots hold great promise for operating in human-centered environments, yet ladder climbing remains one of the most challenging tasks due to

MoDex: A Diffusion Policy for Sequential Multi-Object Dexterous Grasping

SafetyDGX agent

arXiv:2606.05407v1 Announce Type: new Abstract: This work addresses sequentially grasping multiple objects with a single dexterous hand without releasing those already held. Most dexterous grasping me

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Seeking Counsel: Ongoing Targeted Campaign Against US Law Firms

SafetyDGX agent

Written by: Chad Reams, Tufail Ahmed, Keith Knapp, Ashley Frazer, Tyler McLellan Introduction From January through May 2026, Mandiant identified a financially motivated data theft extortion campaign e

TAGA: Terrain-aware Active Gaze Learning for Generalizable Agile Humanoid Locomotion

Local AiDGX agent

arXiv:2606.05880v1 Announce Type: new Abstract: Agile humanoid locomotion across diverse challenging terrain demands both wide perceptual coverage and precise local geometry understanding. Motivated b

v0.30.6-rc0

Local AiDGX agent

v0.30.6-rc0 is a release candidate that fixes kernel template instantiation so library symbols are exported correctly , following improvements from the v0.30 series. The v0.30 base release improved co

4 Jun 2026

Accelerate Edge AI Development with Foundry Local

Local AiDGX agent

Why edge AI development is still hard AI is no longer confined to cloud experiments. Developers are increasingly expected to deliver AI inside apps, devices, and edge systems where responsiveness, pri

And another open-weight release. Nemotron 3 Ultra has an ultra impressive capability:efficiency ratio! Design-wise, it carries forward the M…

Model ReleasesDGX agent

And another open-weight release. Nemotron 3 Ultra has an ultra impressive capability:efficiency ratio! Design-wise, it carries forward the Mamba-2-attention hybrid stack and LatentMoE introduced in th

AutoNumerics-Zero: Automated Discovery of State-of-the-Art Mathematical Functions

ResearchDGX agent

arXiv:2312.08472v2 Announce Type: replace-cross Abstract: Transcendental functions, such as the exponential, are central to scientific computing, yet they cannot be natively calculated by digital hard

b9500

Local AiDGX agent

B9500 is a release of llama.cpp that includes a Metal backend optimization reducing reset heartbeat timing from 500ms to 5ms . The release provides compiled binaries across multiple platforms includin

b9503

Local AiDGX agent

Release b9503 of llama.cpp fixes multimodal (mtmd) support by handling Gemma 4 audio projector embedding size , specifically removing the projection_dim from clip_n_mmproj_embd. This build includes pr

b9509

Local AiDGX agent

Release b9509 is a build version of llama.cpp, which provides LLM inference in C/C++ . As a specific build identifier in the llama.cpp project's release history, it likely contains bug fixes, performa

b9512

Local AiDGX agent

b9512 is a release build of llama.cpp, which provides LLM inference in C/C++ . The release uses a version numbering system with 'b' prefixes for intermediate builds, with newer releases building upon

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

Model ReleasesDGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

DiffAero: A GPU-Accelerated Differentiable Simulation Framework for Efficient Quadrotor Policy Learning

SafetyDGX agent

arXiv:2509.10247v1 Announce Type: cross Abstract: This letter introduces DiffAero, a lightweight, GPU-accelerated, and fully differentiable simulation framework designed for efficient quadrotor contro

dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats

Model ReleasesDGX agent

arXiv:2606.04115v1 Announce Type: cross Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engine…

Model ReleasesDGX agent

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU perfor

HighTide: An Agent-Curated Open-Source VLSI Benchmark Suite

Model ReleasesDGX agent

arXiv:2606.04126v1 Announce Type: cross Abstract: We introduce HighTide, an evolving AI-assisted benchmark suite. Specifically, the contributions are: (i) a diverse open-source suite spanning multiple

L^3: Large Lookup Layers

ResearchDGX agent

arXiv:2601.21461v3 Announce Type: replace-cross Abstract: Modern sparse language models typically achieve sparsity through Mixture-of-Experts (MoE) layers, which dynamically route tokens to dense MLP

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection

Model ReleasesDGX agent

arXiv:2606.04050v1 Announce Type: cross Abstract: Existing quantization methods are fundamentally limited by rigid, integer-based bit-widths (e.g., 2, 3-bit), resulting in a ``deployment gap' where La

NLLog: Lightweight, Explainable SOC Anomaly Detection via Log-to-Language Rewriting

ResearchDGX agent

arXiv:2606.04957v1 Announce Type: cross Abstract: System-generated logs underpin security monitoring, yet their rigid template-based format hinders both automated analysis and human comprehension. We

PerceptTwin: Semantic Scene Reconstruction for Iterative LLM Planning and Verification

SafetyDGX agent

arXiv:2606.04226v1 Announce Type: cross Abstract: Simulation environments are useful for both robot policy learning and planning verification and validation. Traditionally, the process of creating a s

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

Local AiDGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

SSSD: Simply-Scalable Speculative Decoding

ApplicationsDGX agent

arXiv:2411.05894v3 Announce Type: replace-cross Abstract: Speculative Decoding has emerged as a popular technique for accelerating inference in Large Language Models. However, most existing approaches

Teaching Robots to Say 'I Don't Know' : SENTINEL for Uncertainty-Aware SLAM

ResearchDGX agent

arXiv:2606.04853v1 Announce Type: new Abstract: Low-cost 2D LiDARs lack the intensity channel that higher-end sensors use to diagnose measurement failures, yet they are widely used on educational and

The Differentiable Auditory Loop (DAL): An ML Framework for Hyper-Personalized Hearing Aids

ApplicationsDGX agent

arXiv:2606.04103v1 Announce Type: cross Abstract: Conventional hearing aids rely on fixed, frequency-dependent amplification and compression to manage reduced sensitivity, which often fails to provide

The future of quantum takes center stage at NY Tech Week

ResearchDGX agent

IBM Research highlighted quantum computing developments and applications at NY Tech Week, showcasing the technology's potential impact on industry and research. The event featured discussions on quant

TransTac: Visuo-Tactile Modality Transition via Ultraviolet-Encoded Transparent Elastomers

SafetyDGX agent

arXiv:2606.04477v1 Announce Type: new Abstract: Vision-based tactile sensors (VBTS) recover high-resolution contact geometry but typically rely on opaque elastomer layers that prevent visual transpare

Uncertainty-Aware End-to-End Co-Design of Neural Network Processors: From Training and Mapping to Fabrication

ApplicationsDGX agent

arXiv:2606.04850v1 Announce Type: cross Abstract: Designing a neural network processor is an end-to-end co-design problem: network architecture and training budget determine the inference workload; ha

v0.30.5

Local AiDGX agent

v0.30.5 is a release candidate version of Ollama, a local large language model runner. The v0.30 series features improved compatibility and performance using llama.cpp and augments the MLX engine on A

Vectorized Online POMDP Planning

AgentsDGX agent

arXiv:2510.27191v5 Announce Type: replace-cross Abstract: Planning under partial observability is an essential capability of autonomous robots. The Partially Observable Markov Decision Process (POMDP)

3 Jun 2026

A Measurement-Driven Digital Twin Architecture for Plant-Level Biomass Estimation and Growth Forecasting in Hydroponic Systems

ResearchDGX agent

arXiv:2606.02796v1 Announce Type: new Abstract: Alternatives to soil-based horticulture, such as hydroponics, have been developed to respond to food distribution concerns for dense urban centers. A ne

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

Local AiDGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

ArrowFlow: Hierarchical Machine Learning in the Space of Permutations

Model ReleasesDGX agent

arXiv:2604.04087v2 Announce Type: replace Abstract: We introduce ArrowFlow, a machine learning architecture that operates entirely in the space of permutations. Its computational units are ranking fil

Assessing Region-Level EEG Contributions to Cognitive Workload Prediction

SafetyDGX agent

arXiv:2606.02598v1 Announce Type: new Abstract: Accurate and generalizable estimation of cognitive workload from electroencephalography (EEG) is critical for human-centered and safety-critical systems

ATLAS: A Large-Scale Evaluation Benchmark for Adversarial LiDAR Perception

Model ReleasesDGX agent

arXiv:2606.02924v1 Announce Type: new Abstract: Autonomous driving perception is typically evaluated on clean benchmark data, yet real-world deployment requires robustness to rare, structured, and pot

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

Model ReleasesDGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

b9491

Local AiDGX agent

b9491 is a release of llama.cpp , a C/C++ implementation of large language model inference that enables running LLMs locally with minimal dependencies. This release likely contains bug fixes, feature

b9493

Local AiDGX agent

B9493 is a release of llama.cpp, an LLM inference framework in C/C++ . This release includes updates to model support, such as centralized hidden activation mappings and additions for granite embeddin

b9494

Local AiDGX agent

The search results do not contain specific details about the b9494 release. Based on the context from the llama.cpp GitHub releases page, b9494 is an intermediate build release of llama.cpp, a C/C++ p

b9495

Local AiDGX agent

The search didn't return specific information about release b9495. Based on the available information about llama.cpp releases, here's a summary: llama.cpp enables LLM inference with minimal setup and

b9496

Local AiDGX agent

The search results do not contain specific information about release b9496. Based on the available context, b9496 is a build/release version of llama.cpp, a C/C++ implementation for efficient LLM infe

Bionic Human-Motion Style Transfer for Physically Executable Whole-Body Control of Humanoid Robots

SafetyDGX agent

arXiv:2606.03536v1 Announce Type: new Abstract: Expressive whole-body motion is important for humanoid robots operating in human environments, where robots are expected to move stably while presenting

CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA

Model ReleasesDGX agent

arXiv:2512.00360v2 Announce Type: replace Abstract: We study timestamped question answering over educational lecture videos under a single-GPU latency/memory budget. Given a natural-language query, th

Dynamic Short Convolutions Improve Transformers

SafetyDGX agent

arXiv:2606.03825v1 Announce Type: cross Abstract: Transformers have become the dominant architecture for large language models, largely due to the scalability and flexibility of attention, feed-forwar

Evaluating LLMs' Effectiveness on Real-World Consumer Device Repair Questions

Model ReleasesDGX agent

arXiv:2606.03331v1 Announce Type: cross Abstract: Consumer device repair is an important but underexplored testbed for large language models (LLMs). Repair tasks require reasoning over incomplete prob

Fixed-Time Dynamic Landing of Quadrotors using Adaptive Unscented Kalman Filtering and Nonlinear Model Predictive Control

ResearchDGX agent

arXiv:2606.02658v1 Announce Type: new Abstract: This paper introduces an estimation and control framework for dynamic landing of multi-rotor uncrewed aerial vehicles on moving platforms. The proposed

FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs

Model ReleasesDGX agent

arXiv:2506.01969v3 Announce Type: replace-cross Abstract: Efficient inference of Multi-Head Latent Attention (MLA) is challenged by deploying the DeepSeek-R1 671B model on a single Multi-GPU server. T

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

Model ReleasesDGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

Latent Activation Editing: Inference-Time Refinement of Learned Policies for Safer Multirobot Navigation

SafetyDGX agent

arXiv:2509.20623v2 Announce Type: replace Abstract: Reinforcement learning has enabled significant progress in complex domains such as coordinating and navigating multiple quadrotors. However, even we

Link Prediction or Perdition: the Seeds of Instability in Knowledge Graph Embeddings

ResearchDGX agent

arXiv:2606.03365v1 Announce Type: new Abstract: Embedding models (KGEMs) constitute the main link prediction approach to complete knowledge graphs. Standard evaluation protocols emphasize rank-based m

LiveBand: Live Accompaniment Generation in the Audio Domain

Model ReleasesDGX agent

arXiv:2606.03803v1 Announce Type: cross Abstract: We present LiveBand, a real-time system that generates high-fidelity music accompaniments to live audio input, respecting strict causal constraints. O

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

SafetyDGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

NetKV: Network-Aware Decode Instance Selection for Disaggregated LLM Inference

Local AiDGX agent

arXiv:2606.03910v1 Announce Type: cross Abstract: Disaggregated LLM inference forces the KV cache to traverse the datacenter network before decoding begins, so transfer time enters directly into the T

OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration

ResearchDGX agent

arXiv:2507.23035v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities across a wide range of applications, but demand substantial memory and comput

Optimizing Explicit Unit-Distance Lower-Bound Certificates

Model ReleasesDGX agent

arXiv:2606.03419v1 Announce Type: cross Abstract: The 2026 disproof of Erdos's unit-distance conjecture and Sawin's subsequent explicit quantitative refinement show that the maximum number u(n) of uni

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

SafetyDGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

PSViT: A Methodology for Structurally Pruning Spiking Vision Transformers

ResearchDGX agent

arXiv:2606.03257v1 Announce Type: cross Abstract: Spiking Vision Transformer (SViT) models are promising low-power ViT models for solving vision-based tasks with state-of-the-art performance. However,

QUIVER: Quantum-Informed Views for Enhanced Representations in Large ML Models

Model ReleasesDGX agent

arXiv:2606.02785v1 Announce Type: new Abstract: Large machine learning models benefit substantially from multimodal inputs that provide a complementary view of the same example. We introduce QUIVER (Q

Qwen3.6-35B-A3B on 2× GTX 1080 Ti with Ollama: ~20 tok/s + 3 gotchas (driver 570+, cuda_v12 for Pascal, quant fit on 22GB)

Local AiDGX agent

This post documents running the Qwen3.6-35B-A3B language model on dual GTX 1080 Ti GPUs using Ollama, achieving approximately 20 tokens per second. The author highlights three critical configuration r

← Previous
1…5758596061…75
Next →