AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
4,474 results
Model Releases

And another open-weight release. Nemotron 3 Ultra has an ultra impressive capability:efficiency ratio! Design-wise, it carries forward the M…

DGX agent

And another open-weight release. Nemotron 3 Ultra has an ultra impressive capability:efficiency ratio! Design-wise, it carries forward the Mamba-2-attention hybrid stack and LatentMoE introduced in th

model-releasessebastian-raschka--x
4 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

AutoNumerics-Zero: Automated Discovery of State-of-the-Art Mathematical Functions

DGX agent

arXiv:2312.08472v2 Announce Type: replace-cross Abstract: Transcendental functions, such as the exponential, are central to scientific computing, yet they cannot be natively calculated by digital hard

researcharxiv-cs-lg
4 Jun 2026
Local Ai

b9500

DGX agent

B9500 is a release of llama.cpp that includes a Metal backend optimization reducing reset heartbeat timing from 500ms to 5ms . The release provides compiled binaries across multiple platforms includin

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9503

DGX agent

Release b9503 of llama.cpp fixes multimodal (mtmd) support by handling Gemma 4 audio projector embedding size , specifically removing the projection_dim from clip_n_mmproj_embd. This build includes pr

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9509

DGX agent

Release b9509 is a build version of llama.cpp, which provides LLM inference in C/C++ . As a specific build identifier in the llama.cpp project's release history, it likely contains bug fixes, performa

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9512

DGX agent

b9512 is a release build of llama.cpp, which provides LLM inference in C/C++ . The release uses a version numbering system with 'b' prefixes for intermediate builds, with newer releases building upon

local-aillama-cpp-releases
4 Jun 2026
Model Releases

CodegenBench: Can LLMs Write Efficient Code Across Architectures?

DGX agent

arXiv:2606.04023v1 Announce Type: cross Abstract: While large language models (LLMs) have been extensively evaluated on code generation tasks for general-purpose programming and GPU-accelerated enviro

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

DiffAero: A GPU-Accelerated Differentiable Simulation Framework for Efficient Quadrotor Policy Learning

DGX agent

arXiv:2509.10247v1 Announce Type: cross Abstract: This letter introduces DiffAero, a lightweight, GPU-accelerated, and fully differentiable simulation framework designed for efficient quadrotor contro

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats

DGX agent

arXiv:2606.04115v1 Announce Type: cross Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engine…

DGX agent

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU perfor

model-releasesgeorgi-gerganov--x
4 Jun 2026
Model Releases

HighTide: An Agent-Curated Open-Source VLSI Benchmark Suite

DGX agent

arXiv:2606.04126v1 Announce Type: cross Abstract: We introduce HighTide, an evolving AI-assisted benchmark suite. Specifically, the contributions are: (i) a diverse open-source suite spanning multiple

model-releasesarxiv-cs-ai
4 Jun 2026
Research

L^3: Large Lookup Layers

DGX agent

arXiv:2601.21461v3 Announce Type: replace-cross Abstract: Modern sparse language models typically achieve sparsity through Mixture-of-Experts (MoE) layers, which dynamically route tokens to dense MLP

researcharxiv-cs-ai
4 Jun 2026
Model Releases

LiftQuant: Continuous Bit-Width LLM via Dimensional Lifting and Projection

DGX agent

arXiv:2606.04050v1 Announce Type: cross Abstract: Existing quantization methods are fundamentally limited by rigid, integer-based bit-widths (e.g., 2, 3-bit), resulting in a ``deployment gap' where La

model-releasesarxiv-cs-ai
4 Jun 2026
Research

NLLog: Lightweight, Explainable SOC Anomaly Detection via Log-to-Language Rewriting

DGX agent

arXiv:2606.04957v1 Announce Type: cross Abstract: System-generated logs underpin security monitoring, yet their rigid template-based format hinders both automated analysis and human comprehension. We

researcharxiv-cs-lg
4 Jun 2026
Safety

PerceptTwin: Semantic Scene Reconstruction for Iterative LLM Planning and Verification

DGX agent

arXiv:2606.04226v1 Announce Type: cross Abstract: Simulation environments are useful for both robot policy learning and planning verification and validation. Traditionally, the process of creating a s

safetyarxiv-cs-ai
4 Jun 2026
Local Ai

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

DGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

local-aiarxiv-cs-ai
4 Jun 2026
Applications

SSSD: Simply-Scalable Speculative Decoding

DGX agent

arXiv:2411.05894v3 Announce Type: replace-cross Abstract: Speculative Decoding has emerged as a popular technique for accelerating inference in Large Language Models. However, most existing approaches

applicationsarxiv-cs-ai
4 Jun 2026
Research

Teaching Robots to Say 'I Don't Know' : SENTINEL for Uncertainty-Aware SLAM

DGX agent

arXiv:2606.04853v1 Announce Type: new Abstract: Low-cost 2D LiDARs lack the intensity channel that higher-end sensors use to diagnose measurement failures, yet they are widely used on educational and

researcharxiv-cs-ro
4 Jun 2026
Applications

The Differentiable Auditory Loop (DAL): An ML Framework for Hyper-Personalized Hearing Aids

DGX agent

arXiv:2606.04103v1 Announce Type: cross Abstract: Conventional hearing aids rely on fixed, frequency-dependent amplification and compression to manage reduced sensitivity, which often fails to provide

applicationsarxiv-cs-ai
4 Jun 2026
Research

The future of quantum takes center stage at NY Tech Week

DGX agent

IBM Research highlighted quantum computing developments and applications at NY Tech Week, showcasing the technology's potential impact on industry and research. The event featured discussions on quant

researchibm-research
4 Jun 2026
Safety

TransTac: Visuo-Tactile Modality Transition via Ultraviolet-Encoded Transparent Elastomers

DGX agent

arXiv:2606.04477v1 Announce Type: new Abstract: Vision-based tactile sensors (VBTS) recover high-resolution contact geometry but typically rely on opaque elastomer layers that prevent visual transpare

safetyarxiv-cs-ro
4 Jun 2026
Applications

Uncertainty-Aware End-to-End Co-Design of Neural Network Processors: From Training and Mapping to Fabrication

DGX agent

arXiv:2606.04850v1 Announce Type: cross Abstract: Designing a neural network processor is an end-to-end co-design problem: network architecture and training budget determine the inference workload; ha

applicationsarxiv-cs-ai
4 Jun 2026
Local Ai

v0.30.5

DGX agent

v0.30.5 is a release candidate version of Ollama, a local large language model runner. The v0.30 series features improved compatibility and performance using llama.cpp and augments the MLX engine on A

local-aiollama-releases
4 Jun 2026
Agents

Vectorized Online POMDP Planning

DGX agent

arXiv:2510.27191v5 Announce Type: replace-cross Abstract: Planning under partial observability is an essential capability of autonomous robots. The Partially Observable Markov Decision Process (POMDP)

agentsarxiv-cs-ai
4 Jun 2026
Research

A Measurement-Driven Digital Twin Architecture for Plant-Level Biomass Estimation and Growth Forecasting in Hydroponic Systems

DGX agent

arXiv:2606.02796v1 Announce Type: new Abstract: Alternatives to soil-based horticulture, such as hydroponics, have been developed to respond to food distribution concerns for dense urban centers. A ne

researcharxiv-cs-ro
3 Jun 2026
Local Ai

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

DGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

local-aiarxiv-cs-ai
3 Jun 2026
Model Releases

ArrowFlow: Hierarchical Machine Learning in the Space of Permutations

DGX agent

arXiv:2604.04087v2 Announce Type: replace Abstract: We introduce ArrowFlow, a machine learning architecture that operates entirely in the space of permutations. Its computational units are ranking fil

model-releasesarxiv-cs-lg
3 Jun 2026
Safety

Assessing Region-Level EEG Contributions to Cognitive Workload Prediction

DGX agent

arXiv:2606.02598v1 Announce Type: new Abstract: Accurate and generalizable estimation of cognitive workload from electroencephalography (EEG) is critical for human-centered and safety-critical systems

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

ATLAS: A Large-Scale Evaluation Benchmark for Adversarial LiDAR Perception

DGX agent

arXiv:2606.02924v1 Announce Type: new Abstract: Autonomous driving perception is typically evaluated on clean benchmark data, yet real-world deployment requires robustness to rare, structured, and pot

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

AURA: Action-Gated Memory for Robot Policies at Constant VRAM

DGX agent

arXiv:2606.02775v1 Announce Type: new Abstract: The KV-cache is the right memory for datacenters but the wrong memory for robots. Datacenter inference batches many short requests and resets them, amor

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

b9491

DGX agent

b9491 is a release of llama.cpp , a C/C++ implementation of large language model inference that enables running LLMs locally with minimal dependencies. This release likely contains bug fixes, feature

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9493

DGX agent

B9493 is a release of llama.cpp, an LLM inference framework in C/C++ . This release includes updates to model support, such as centralized hidden activation mappings and additions for granite embeddin

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9494

DGX agent

The search results do not contain specific details about the b9494 release. Based on the context from the llama.cpp GitHub releases page, b9494 is an intermediate build release of llama.cpp, a C/C++ p

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9495

DGX agent

The search didn't return specific information about release b9495. Based on the available information about llama.cpp releases, here's a summary: llama.cpp enables LLM inference with minimal setup and

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9496

DGX agent

The search results do not contain specific information about release b9496. Based on the available context, b9496 is a build/release version of llama.cpp, a C/C++ implementation for efficient LLM infe

local-aillama-cpp-releases
3 Jun 2026
Safety

Bionic Human-Motion Style Transfer for Physically Executable Whole-Body Control of Humanoid Robots

DGX agent

arXiv:2606.03536v1 Announce Type: new Abstract: Expressive whole-body motion is important for humanoid robots operating in human environments, where robots are expected to move stably while presenting

safetyarxiv-cs-ro
3 Jun 2026
Model Releases

CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA

DGX agent

arXiv:2512.00360v2 Announce Type: replace Abstract: We study timestamped question answering over educational lecture videos under a single-GPU latency/memory budget. Given a natural-language query, th

model-releasesarxiv-cs-cl
3 Jun 2026
Safety

Dynamic Short Convolutions Improve Transformers

DGX agent

arXiv:2606.03825v1 Announce Type: cross Abstract: Transformers have become the dominant architecture for large language models, largely due to the scalability and flexibility of attention, feed-forwar

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

Evaluating LLMs' Effectiveness on Real-World Consumer Device Repair Questions

DGX agent

arXiv:2606.03331v1 Announce Type: cross Abstract: Consumer device repair is an important but underexplored testbed for large language models (LLMs). Repair tasks require reasoning over incomplete prob

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Fixed-Time Dynamic Landing of Quadrotors using Adaptive Unscented Kalman Filtering and Nonlinear Model Predictive Control

DGX agent

arXiv:2606.02658v1 Announce Type: new Abstract: This paper introduces an estimation and control framework for dynamic landing of multi-rotor uncrewed aerial vehicles on moving platforms. The proposed

researcharxiv-cs-ro
3 Jun 2026
Model Releases

FlashMLA-ETAP: Efficient Transpose Attention Pipeline for Accelerating MLA Inference on NVIDIA H20 GPUs

DGX agent

arXiv:2506.01969v3 Announce Type: replace-cross Abstract: Efficient inference of Multi-Head Latent Attention (MLA) is challenged by deploying the DeepSeek-R1 671B model on a single Multi-GPU server. T

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

DGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

model-releasesarxiv-cs-lg
3 Jun 2026
Safety

Latent Activation Editing: Inference-Time Refinement of Learned Policies for Safer Multirobot Navigation

DGX agent

arXiv:2509.20623v2 Announce Type: replace Abstract: Reinforcement learning has enabled significant progress in complex domains such as coordinating and navigating multiple quadrotors. However, even we

safetyarxiv-cs-ro
3 Jun 2026
Research

Link Prediction or Perdition: the Seeds of Instability in Knowledge Graph Embeddings

DGX agent

arXiv:2606.03365v1 Announce Type: new Abstract: Embedding models (KGEMs) constitute the main link prediction approach to complete knowledge graphs. Standard evaluation protocols emphasize rank-based m

researcharxiv-cs-lg
3 Jun 2026
Model Releases

LiveBand: Live Accompaniment Generation in the Audio Domain

DGX agent

arXiv:2606.03803v1 Announce Type: cross Abstract: We present LiveBand, a real-time system that generates high-fidelity music accompaniments to live audio input, respecting strict causal constraints. O

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

DGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

safetyarxiv-cs-cl
3 Jun 2026
Local Ai

NetKV: Network-Aware Decode Instance Selection for Disaggregated LLM Inference

DGX agent

arXiv:2606.03910v1 Announce Type: cross Abstract: Disaggregated LLM inference forces the KV cache to traverse the datacenter network before decoding begins, so transfer time enters directly into the T

local-aiarxiv-cs-ai
3 Jun 2026
Research

OASIS: Outlier-Aware LUT-Based GEMM with Dual-Side Quantization for LLM Inference Acceleration

DGX agent

arXiv:2507.23035v4 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated impressive capabilities across a wide range of applications, but demand substantial memory and comput

researcharxiv-cs-lg
3 Jun 2026
← Previous
1…7273747576…94
Next →