AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

local ai

GridTimelineEvolution
4,671 results
21 May 2026

OmniISR: A Unified Framework for Centralized and Federated Learning via Intermediate Supervision and Regularization

Local AiDGX agent

arXiv:2605.20276v1 Announce Type: new Abstract: The global deployment of edge intelligence operates across heterogeneous legal frameworks. While some regions permit centralized learning (CL) via cloud

On prem.

Local AiDGX agent

On prem. I'm excited about the new @amd Ryzen AI Halo because we need more local hardware for AI builders! There's something fun and exciting about building on your own machines rather than sending to

One-Step Distillation of Discrete Diffusion Image Generators via Fixed-Point Iteration

Local AiDGX agent

arXiv:2605.21484v1 Announce Type: new Abstract: Discrete diffusion models excel at visual synthesis but rely on slow, iterative decoding. Existing single-step distillation methods attempt to bypass th


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Online 3D Multi-Camera Perception through Robust 2D Tracking and Depth-based Late Aggregation

Local AiDGX agent

arXiv:2509.09946v2 Announce Type: replace Abstract: Multi-Target Multi-Camera Tracking (MTMC) is an essential computer vision task for automating large-scale surveillance. With camera calibration and

Optimized Federated Knowledge Distillation with Distributed Neural Architecture Search

Local AiDGX agent

arXiv:2605.21322v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative model training without centralizing data. However, real-world deployments must simultaneously address stat

OSGNet with MLLM Reranking @ Ego4D Episodic Memory Challenge 2026

Local AiDGX agent

arXiv:2605.20818v1 Announce Type: new Abstract: In this report, we present our champion solutions for the Natural Language Queries and GoalStep tracks of the Ego4D Episodic Memory Challenge at CVPR 20

PaintCopilot: Modeling Painting as Autonomous Artistic Continuation

Local AiDGX agent

arXiv:2605.20941v1 Announce Type: new Abstract: We present PaintCopilot, a co-creative neural painting assistant that models painting as an open-ended autoregressive artistic behavior conditioned on e

Quant.npu: Enabling Efficient Mobile NPU Inference for on-device LLMs via Fully Static Quantization

Local AiDGX agent

arXiv:2605.20295v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed on mobile devices, where Neural Processing Units (NPUs) necessitate fully static quantization for

R2AoP: Reliable and Robust Angle of Progression Estimation from Intrapartum Ultrasound

Local AiDGX agent

arXiv:2605.21099v1 Announce Type: new Abstract: Accurate estimation of the Angle of Progression (AoP) from intrapartum transperineal ultrasound is critical for objective assessment of labor progressio

RePCM: Region-Specific and Phenotype-Adaptive Bi-Ventricular Cardiac Motion Synthesis

Local AiDGX agent

arXiv:2605.21237v1 Announce Type: new Abstract: Cardiac motion over a cardiac cycle is crucial for quantifying regional function and is strongly affected by cardiovascular diseases. Since temporally d

Scale-Calibrated Median-of-Means for Robust Distributed Principal Component Analysis

Local AiDGX agent

arXiv:2605.20681v1 Announce Type: cross Abstract: Distributed principal component analysis (PCA) produces node-level estimates of both a mean vector and a principal subspace. Robustly aggregating thes

Semantic Granularity Navigation in Image Editing

Local AiDGX agent

arXiv:2605.21190v1 Announce Type: new Abstract: Despite the generative capabilities of diffusion and flow models, real-image editing remains constrained by a persistent trade-off between semantic edit

SpikeDet: Better Firing Patterns for Accurate and Energy-Efficient Object Detection with Spiking Neural Networks

Local AiDGX agent

arXiv:2501.15151v5 Announce Type: replace Abstract: Spiking Neural Networks (SNNs) are the third generation of neural networks. They have gained widespread attention in object detection due to their l

SpineContextResUNet: A Computationally Efficient Residual UNet for Spine CT Segmentation

Local AiDGX agent

arXiv:2605.20760v1 Announce Type: new Abstract: Automated segmentation of the vertebral column in Computed Tomography (CT) scans is a prerequisite for pathological assessment and surgical planning. Ho

Stable Audio 3.0 is now Day-0 supported in ComfyUI. Open-weight music models (fully licensed data)—from quick SFX and short tracks to longer…

Local AiDGX agent

Stable Audio 3.0 is now Day-0 supported in ComfyUI. Open-weight music models (fully licensed data)—from quick SFX and short tracks to longer, more musical pieces—inside the workflows you already use.

Sustainability Is Not Linear: Quantifying Performance, Energy, and Privacy Trade-offs in On-Device Intelligence

Local AiDGX agent

arXiv:2603.26603v2 Announce Type: replace-cross Abstract: The migration of Large Language Models (LLMs) from cloud clusters to edge devices promises enhanced privacy and offline accessibility, but thi

The General Theory of Localization Methods

Local AiDGX agent

arXiv:2605.20635v1 Announce Type: new Abstract: This paper proposes a general machine learning framework called the localization method, which is fundamentally built on two core concepts: localization

Unlock Exascale Performance on NVIDIA GB200 NVL72 with Slurm Topology-Aware Job Scheduling

Local AiDGX agent

The NVIDIA GB200 NVL72 is a rack-scale GPU supercomputer leveraging Blackwell architecture with NVLink switches for high-density computing , and topology-aware block scheduling in Slurm can align larg

v0.30.0-rc22

Local AiDGX agent

v0.30.0-rc22 is a pre-release version of Ollama that changes the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX

We have an arxiv paper up describing the work in more detail here: https://arxiv.org/abs/2605.20706. Also want to call out that there is eve…

Local AiDGX agent

We have an arxiv paper up describing the work in more detail here: https://arxiv.org/abs/2605.20706. Also want to call out that there is even more room for improvement, some recent updates to wllama b

What are the latest methods for face swapping? (Images & Videos)

Local AiDGX agent

Latest face swapping methods are expanding toward 3D-consistent video swapping and identity-preserving generation systems that control portraits and clips . Improvements in diffusion models and neural

You Don't Need Attention: Gated Convolutional Modeling for Watch-Based Fall Detection

Local AiDGX agent

arXiv:2605.20275v1 Announce Type: new Abstract: Existing deep learning approaches for wearable fall detection systems rely on self-attention mechanisms that impose quadratic computational overhead, di

20 May 2026

3D Modeling and Automated Measurement of Concrete Cracks via Segment Anything Refinement and Visual Inertial LiDAR Fusion

Local AiDGX agent

arXiv:2501.09203v2 Announce Type: replace Abstract: Visual-Spatial Systems has become increasingly essential in concrete crack inspection. However, existing methods often lacks adaptability to diverse

A glimpse into the fleet: Deep space patrol (4K Sci-Fi) [OC]

Local AiDGX agent

A user-generated AI image creation showcasing a deep space patrol scene rendered in 4K quality using Stable Diffusion, a popular text-to-image generation model. The post was shared on the r/StableDiff

AI Technologies in Language Access: Attitudes Towards AI and the Human Value of Language Access Managers

Local AiDGX agent

arXiv:2605.19234v1 Announce Type: cross Abstract: The rapid emergence of AI technologies is reshaping translation practices and theory across the board. This paper deals with the impact of AI in langu

Announcing the release of Stable Audio 3!

Local AiDGX agent

Stability AI announced the launch of Stable Audio 3, a family of three AI music models and one audio-based special effects model. Most of these releases are 'open weight' models trained on licensed tr

b9239

Local AiDGX agent

Build b9239 is a llama.cpp release that includes a fix for the --fit verbosity flag when used with --verbosity 4 . The release provides compiled binaries for multiple platforms including macOS (Apple

b9240

Local AiDGX agent

b9240 is a release of llama.cpp that includes a fix for the --help option related to the --verbosity flag . The release provides prebuilt binaries across multiple platforms including macOS (Apple Sili

b9244

Local AiDGX agent

b9244 is an intermediate build release of llama.cpp, a C/C++ implementation framework for running large language models with GGUF format support. The release includes pre-compiled binaries for multipl

b9245

Local AiDGX agent

llama.cpp release b9245, published on May 20, 2026, includes a CUDA optimization for RDNA3 Q6_K MMVQ performance tuning. The release provides pre-built binaries for multiple platforms including macOS,

b9251

Local AiDGX agent

B9251 is a build identifier for a release in the llama.cpp project, which enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware - locally and in the clo

b9253

Local AiDGX agent

b9253 is the latest version of llama.cpp, released on May 20, 2026. Llama.cpp is a project for LLM inference in C/C++. This build includes bug fixes, performance improvements, and features for running

BabyMamba-HAR: Lightweight Selective State Space Models for Efficient Human Activity Recognition on Resource Constrained Devices

Local AiDGX agent

arXiv:2602.09872v2 Announce Type: replace Abstract: Human activity recognition (HAR) on resource constrained devices requires high accuracy across diverse sensor setups. Selective state space models (

Code-Guided Reasoning for Small Language Models: Evaluating Executable MCQA Scaffolds

Local AiDGX agent

arXiv:2605.18827v1 Announce Type: cross Abstract: Multiple-choice QA benchmarks usually evaluate small language models (SLMs) as direct answerers, but deployed language-model systems increasingly rely

ComfyUI HiDream text->image and image-edit templates - multiple reference image facility. Discuss please.

Local AiDGX agent

HiDream-O1-Image is a unified image generative foundation model that supports text-to-image and instruction-based image editing with support for multiple reference images. The sampler accepts up to 12

CompoSE: Compositional Synthesis and Editing of 3D Shapes via Part-Aware Control

Local AiDGX agent

arXiv:2605.19350v1 Announce Type: cross Abstract: Creating and editing high-quality 3D content remains a central challenge in computer graphics. We address this challenge by introducing CompoSE, a nov

DAG-Based QoS-Aware Dynamic Task Placement for Networked Multi-Stage Control Pipelines

Local AiDGX agent

arXiv:2605.19887v1 Announce Type: cross Abstract: Current Physical AI (PAI) relies heavily on closed-loop visual-servoing pipelines, whose perception and planning stages may become computationally int

Descriptive versus Regulatory Uncertainty in Bounded Predictive Systems

Local AiDGX agent

arXiv:2605.18909v1 Announce Type: new Abstract: Any system that models the world under finite representational capacity must compress; any compression entails a prior; and the prior is the system's bi

Draft Less, Retrieve More: Hybrid Tree Construction for Speculative Decoding

Local AiDGX agent

arXiv:2605.20104v1 Announce Type: cross Abstract: Speculative decoding (SD) accelerates large language model inference by leveraging a draft-then-verify paradigm. To maximize the acceptance rate, rece

Dual-Channel Tensor Neural Networks: Finite-Sample Theory and Conformal Structure Selection

Local AiDGX agent

arXiv:2605.19122v1 Announce Type: cross Abstract: Tensor-valued data arise naturally in neuroimaging, genomics, climate science, and spatiotemporal networks, where multilinear dependencies across mode

Efficient coding along the visual hierarchy

Local AiDGX agent

arXiv:2605.19155v1 Announce Type: new Abstract: Biological visual systems learn from limited experience, unlike deep learning models that rely on millions of training images. What learning principles

Efficient Conditioning Why Pseudo Observation Batch Bayesian Optimization Works When It Does not

Local AiDGX agent

arXiv:2605.18819v1 Announce Type: new Abstract: Constant Liar (CL), Kriging Believer (KB), and fantasy models are widely used for batch selection in parallel Bayesian Optimization, yet a unified theor

FieldFormer: Locality-Aware Transformers for Spatio-Temporal Modeling on Sparse Sensor Networks

Local AiDGX agent

arXiv:2510.03589v2 Announce Type: replace Abstract: Spatio-temporal sensor data in real-world systems is often sparse, noisy, and irregular, making latent field reconstruction fundamentally underconst

FormalASR: End-to-End Spoken Chinese to Formal Text

Local AiDGX agent

arXiv:2605.19266v1 Announce Type: cross Abstract: Automatic speech recognition (ASR) systems are typically optimized for verbatim transcription, which preserves disfluencies, filler words, and informa

Gaussian Approximation and Multiplier Bootstrap for Federated Linear Stochastic Approximation

Local AiDGX agent

arXiv:2605.19629v1 Announce Type: cross Abstract: In this paper, we establish Berry-Esseen-type bounds for federated linear stochastic approximation (LSA). Our results provide the first federated Gaus

Going PLACES: Participatory Localized Red Teaming for Text-to-Image Safety in the Global South

Local AiDGX agent

arXiv:2605.19190v1 Announce Type: cross Abstract: Despite the global deployment of text-to-image (T2I) models, their safety frameworks are largely calibrated to a Western-centric default, creating sig

GRALIS: A Unified Canonical Framework for Linear Attribution Methods via Riesz Representation

Local AiDGX agent

arXiv:2605.05480v2 Announce Type: replace-cross Abstract: The main XAI attribution methods for deep neural networks -- GradCAM, SHAP, LIME, Integrated Gradients -- operate on separate theoretical foun

GRASP: Deterministic argument ranking in interaction graphs

Local AiDGX agent

arXiv:2605.19141v1 Announce Type: cross Abstract: Large language models are increasingly deployed as automated judges to evaluate the strength of arguments. As this role expands, their legitimacy depe

GRLoc: Geometric Representation Regression for Visual Localization

Local AiDGX agent

arXiv:2511.13864v2 Announce Type: replace Abstract: Absolute Pose Regression (APR) has emerged as a compelling paradigm for visual localization. However, APR models typically operate as black boxes, d

Hierarchical Schedule Optimization for Fast and Robust Diffusion Model Sampling

Local AiDGX agent

arXiv:2511.11688v3 Announce Type: replace-cross Abstract: Diffusion probabilistic models have set a new standard for generative fidelity but are hindered by a slow iterative sampling process. A powerf

I built a local Qwen2.5-VL desktop tool that lets you ask questions about any part of your screen (using Ollama + live overlays)

Local AiDGX agent

This project implements a local desktop application using Qwen2.5-VL (a vision-language model) and Ollama that enables users to interactively query visual content on their screen through a live overla

iDiff: Interpretable Difference-aware Framework for Pairwise Image Quality Assessment

Local AiDGX agent

arXiv:2605.19522v1 Announce Type: new Abstract: Pairwise image quality assessment (IQA) in professional photography requires a model not only to identify the preferred image between two candidates, bu

INSIGHTS: Demonstration-Based Summaries of Time Series Predictors

Local AiDGX agent

arXiv:2605.18849v1 Announce Type: cross Abstract: Explainability methods have progressed rapidly, but global explanations for time-series models remain underdeveloped, with most approaches focusing on

Is that a Fuji apple? 👀🔥

Local AiDGX agent

This post likely features a humorous or striking visual comparison involving Fuji apples, possibly showcasing ComfyUI's image generation or processing capabilities. The fire emoji suggests the content

KIO-planner: Attention-Guided Single-Stage Motion Planning with Dual Mapping for UAV Navigation

Local AiDGX agent

arXiv:2605.19703v1 Announce Type: new Abstract: Autonomous UAV flight in confined, wall-dense environments requires low-latency and reliable motion planning under strict safety constraints. Traditiona

Landscape-Awareness for Geometric View Diffusion Model

Local AiDGX agent

arXiv:2605.19865v1 Announce Type: new Abstract: Accurate camera viewpoint estimation under sparse-view conditions remains challenging, particularly in two-view scenarios. Recent approaches leverage di

Learning to Hand Off: Provably Convergent Workflow Learning under Interface Constraints

Local AiDGX agent

arXiv:2605.19140v1 Announce Type: new Abstract: We study workflow learning in a setting where specialized agents hand off control through a shared artifact, each agent observes only a local function o

Local LLM - privacy first - doctor

Local AiDGX agent

A discussion from the r/ollama community about using local large language models for privacy-sensitive applications, particularly for handling sensitive documents like medical records. Local LLMs proc

📽️ LTX 2.3 IC LoRAs - LipDub & Video Utilities 🎤 Hosts: Purz & Rob ⏲️ May 20th – 3pm PST / 6pm EST 📍 Live on YouTube, X, and Twitch Today…

Local AiDGX agent

📽️ LTX 2.3 IC LoRAs - LipDub & Video Utilities 🎤 Hosts: Purz & Rob ⏲️ May 20th – 3pm PST / 6pm EST 📍 Live on YouTube, X, and Twitch Today we're diving into LTX 2.3 IC LoRAs to check out LipDub and a b

LTX 2.3 IC LoRAs (LipDub and Video Utilities) https://x.com/i/broadcasts/1qKDzPZeLjVJV

Local AiDGX agent

LTX 2.3 IC introduces specialized LoRA (Low-Rank Adaptation) modules for video generation, including LipDub functionality for synchronized lip movements and Video Utilities for enhanced video processi

← Previous
1…4243444546…78
Next →