AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “local-ai”

GridTimelineEvolution
4,679 results
4 May 2026

Been noticing a lot of 'slow responses' today: models do not inherently more slow, rate limiting more likely.

Local AiDGX agent

A Reddit discussion from the Ollama community addresses reports of slow model responses, clarifying that the models themselves are not inherently slower but that rate limiting is a more likely cause o

Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking

Local AiDGX agent

arXiv:2605.00348v1 Announce Type: cross Abstract: Recent multi-bit watermarking methods for large language models (LLMs) prioritize capacity over reliability, often conflating decoding with detection.

Characterizing the Expressivity of Local Attention in Transformers

Local AiDGX agent

arXiv:2605.00768v1 Announce Type: new Abstract: The transformer is the most popular neural architecture for language modeling. The cornerstone of the transformer is its global attention mechanism, whi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cloud Is Closer Than It Appears: Revisiting the Tradeoffs of Distributed Real-Time Inference

Local AiDGX agent

arXiv:2605.00005v1 Announce Type: new Abstract: The increasing deployment of deep neural networks (DNNs) in cyber-physical systems (CPS) enhances perception fidelity, but imposes substantial computati

Comparative Analysis of Polygon-Based and Global Machine Learning Models for Bus Occupancy Prediction

Local AiDGX agent

arXiv:2605.00083v1 Announce Type: new Abstract: Accurate forecasting of bus ridership (passengers numbers) is crucial for efficient management and optimization of public transport systems. Traditional

Depth-Guided Privacy-Preserving Visual Localization Using 3D Sphere Clouds

Local AiDGX agent

arXiv:2605.00562v1 Announce Type: new Abstract: The emergence of deep neural networks capable of revealing high-fidelity scene details from sparse 3D point clouds has raised significant privacy concer

Did they shut down deep seek cloud for free users?

Local AiDGX agent

DeepSeek V3 and R1 API free tier includes 500M tokens per month , indicating free access remains available for API users. The search results focus on a major service outage in March 2026 and the recen

E^2DT: Efficient and Effective Decision Transformer with Experience-Aware Sampling for Robotic Manipulation

Local AiDGX agent

arXiv:2605.00159v1 Announce Type: new Abstract: In reinforcement learning (RL) for robotic manipulation, the Decision Transformer (DT) has emerged as an effective framework for addressing long-horizon

EGREFINE: An Execution-Grounded Optimization Framework for Text-to-SQL Schema Refinement

Local AiDGX agent

arXiv:2605.00628v1 Announce Type: cross Abstract: Text-to-SQL enables non-expert users to query databases in natural language, yet real-world schemas often suffer from ambiguous, abbreviated, or incon

Federated Distillation for Whole Slide Image via Gaussian-Mixture Feature Alignment and Curriculum Integration

Local AiDGX agent

arXiv:2605.00578v1 Announce Type: new Abstract: Federated learning (FL) offers a promising framework for collaborative digital pathology by enabling model training across institutions. However, real-w

FedKPer: Tackling Generalization and Personalization in Medical Federated Learning via Knowledge Personalization

Local AiDGX agent

arXiv:2605.00698v1 Announce Type: cross Abstract: Federated learning (FL) holds great potential for medical applications. However, statistical heterogeneity across healthcare institutions poses a majo

From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings

Local AiDGX agent

arXiv:2605.00225v1 Announce Type: cross Abstract: We show that pretrained acoustic embeddings classify elephant vocalisations at a level approaching that of end-to-end supervised neural networks, with

Gated Differential Linear Attention: A Linear-Time Decoder for High-Fidelity Medical Segmentation

Local AiDGX agent

arXiv:2603.02727v4 Announce Type: replace Abstract: Medical image segmentation requires models that preserve fine anatomical boundaries while remaining practical for clinical deployment. Transformers

Hyperspherical Forward-Forward with Prototypical Representations

Local AiDGX agent

arXiv:2605.00082v1 Announce Type: new Abstract: The Forward-Forward (FF) algorithm presents a compelling, bio-inspired alternative to backpropagation. However, while efficient in training, it has a co

i thought on device ai was stupid but now my local voice model turns 4x faster on my corporate m4 max laptop than my m2 max personal laptop …

Local AiDGX agent

i thought on device ai was stupid but now my local voice model turns 4x faster on my corporate m4 max laptop than my m2 max personal laptop now i want to have a beefier computer to run the transcripti

Learning the Helmholtz equation operator with DeepONet for non-parametric 2D geometries

Local AiDGX agent

arXiv:2605.00760v1 Announce Type: new Abstract: This paper deals with solving the 2D Helmholtz equation on non-parametric domains, leveraging a physics-informed neural operator network based on the De

Lucid-XR: An Extended-Reality Data Engine for Robotic Manipulation

Local AiDGX agent

arXiv:2605.00244v1 Announce Type: cross Abstract: We introduce Lucid-XR, a generative data engine for creating diverse and realistic-looking multi-modal data to train real-world robotic systems. At th

MSACT: Multistage Spatial Alignment for Stable Low-Latency Fine Manipulation

Local AiDGX agent

arXiv:2605.00475v1 Announce Type: cross Abstract: Real-world fine manipulation, particularly in bimanual manipulation, typically requires low-latency control and stable visual localization, while coll

NLPOpt-Net: A Learning Method for Nonlinear Optimization with Feasibility Guarantees

Local AiDGX agent

arXiv:2605.00260v1 Announce Type: new Abstract: Nonlinear Parametric Optimization Network (NLPOpt-Net) is an unsupervised learning architecture to solve constrained nonlinear programs (NLP). Given the

NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search

Local AiDGX agent

arXiv:2605.00751v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) scales poorly in cooperative multi-agent domains because expansion must consider an exponentially large set of joint acti

OneTrainer now supports Ernie LoRA

Local AiDGX agent

OneTrainer is a one-stop solution for all diffusion training needs. The tool now supports the Ernie Image model, which can be trained using LoRA (Low-Rank Adaptation) methods. This adds support for tr

Polaris: Coupled Orbital Polar Embeddings for Hierarchical Concept Learning

Local AiDGX agent

arXiv:2605.00265v1 Announce Type: new Abstract: Real-world knowledge is often organized as hierarchies such as product taxonomies, medical ontologies, and label trees, yet learning hierarchical repres

Prefer-DAS: Learning from Local Preferences and Sparse Prompts for Domain Adaptive Segmentation of Electron Microscopy

Local AiDGX agent

arXiv:2602.19423v3 Announce Type: replace Abstract: Domain adaptive segmentation (DAS) is a promising paradigm for delineating intracellular structures from various large-scale electron microscopy (EM

RadLite: Multi-Task LoRA Fine-Tuning of Small Language Models for CPU-Deployable Radiology AI

Local AiDGX agent

arXiv:2605.00421v1 Announce Type: new Abstract: Large language models (LLMs) show promise in radiology but their deployment is limited by computational requirements that preclude use in resource-const

Sentinel: an open-source local-first desktop app for AI coding

Local AiDGX agent

Sentinel is a local-first AI coding desktop application built with Rust and Tauri that automatically routes coding tasks to appropriate models based on task complexity. Instead of swapping between mul

Tempus: A Temporally Scalable Resource-Invariant GEMM Streaming Framework for Versal AI Edge

Local AiDGX agent

arXiv:2605.00536v1 Announce Type: cross Abstract: Scaling laws for Large Language Models (LLMs) establish that model quality improves with computational scale, yet edge deployment imposes strict const

Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation

Local AiDGX agent

arXiv:2605.00438v1 Announce Type: cross Abstract: Long-horizon robotic manipulation requires plans that are both logically coherent and geometrically grounded. Existing Vision-Language-Action policies

This April we shipped 14 new models ⬇️ → Seedance 2.0 → GPT Image 2 → Wan 2.7 → Ace Step 1.5 XL → Quiver SVG → Happy Horse → Ernie-Image → V…

Local AiDGX agent

This April we shipped 14 new models ⬇️ → Seedance 2.0 → GPT Image 2 → Wan 2.7 → Ace Step 1.5 XL → Quiver SVG → Happy Horse → Ernie-Image → Veo 3.1 Lite / Veo 3.1 Fast → Sonilo → SUPIR → RIFE & FILM in

TimesNet-Gen: Deep Learning-based Site Specific Strong Motion Generation

Local AiDGX agent

arXiv:2512.04694v3 Announce Type: replace Abstract: Effective earthquake risk reduction relies on accurate site-specific evaluations, which require models capable of representing the influence of loca

TokenWeave: Efficient Compute-Communication Overlap for Distributed LLM Inference

Local AiDGX agent

arXiv:2505.11329v5 Announce Type: replace-cross Abstract: Distributed inference of large language models (LLMs) using tensor parallelism can introduce communication overheads of 20% even over GPUs con

VkSplat: High-Performance 3DGS Training in Vulkan Compute

Local AiDGX agent

arXiv:2605.00219v1 Announce Type: new Abstract: We present VkSplat, a high-performance, cross-vendor 3D Gaussian Splatting (3DGS) training pipeline implemented fully in Vulkan compute, addressing perf

Would a 2nd hand custom built 2080 Ti 22GB vram be worth it? How usable would it be with ComfyUI? Or maybe even 2 pcs with NVLink? Can large models, like Wan2.2 be split with NVLink like it's 44GB, or will it always be 22GB for 1 model, and 22GB for another model (encoders, CLIP, anything else)?

Local AiDGX agent

This Reddit post discusses the viability of using second-hand custom-built RTX 2080 Ti GPUs with 22GB VRAM for running Stable Diffusion models in ComfyUI, including whether two cards could be linked v

3 May 2026

b9012

Local AiDGX agent

b9012 is a build version from the llama.cpp GitHub repository, which is an LLM inference implementation in C/C++ . The release likely contains updates, bug fixes, and improvements to the llama.cpp inf

Best Local Vision-Language Models?

Local AiDGX agent

This discussion thread explores lightweight vision-language models that can run locally, including options like Llama 3.2 Vision, Qwen2.5-VL, and SmolVLM2, optimized for tasks like OCR and visual ques

Built an open-source cognitive OS — persistent memory, 24/7 runtime, bring your own model

Local AiDGX agent

An open-source locally-run conversational AI that moves beyond simple request-response models by implementing persistent memory, belief, and self-reflection. The system stores all user profiles, memor

Comfy developers pushing important updates to fix broken workflows

Local AiDGX agent

ComfyUI developers released important updates addressing workflow compatibility issues, including fixes for model compatibility problems like HunYuan 3D 2.0 support and EasyCache input/output channel

FastSDCPU release v1.0.0-beta.301

Local AiDGX agent

FastSDCPU is an optimized fork of Stable Diffusion designed to run efficiently on CPUs and devices without dedicated GPUs by leveraging Latent Consistency Models and Adversarial Diffusion Distillation

v0.23.0

Local AiDGX agent

The search results show releases from v0.20.0 through v0.22.0, but do not contain specific information about v0.23.0. Based on the available release information from Ollama and the version timeline sh

zit vs zib settings question

Local AiDGX agent

ZIT (Z-Image Turbo) is a compact, fast distilled model tuned for photorealism that uses 8 inference steps with a fixed CFG of 1, which limits creativity . ZIB (Z-Image Base) is an undistilled model id

2 May 2026

【革命】動画から3Dモーションを抽出!ComfyUIの新機能が凄すぎる 動画から人物の動きを抜き出し、高品質な3Dアニメーションに変換できる「Mocap Surgeon」が登場しました!✨ 注目の神機能はこちら: ・Jitter Filtering:動きのガタつきを数理的に除去し…

Local AiDGX agent

【革命】動画から3Dモーションを抽出!ComfyUIの新機能が凄すぎる 動画から人物の動きを抜き出し、高品質な3Dアニメーションに変換できる「Mocap Surgeon」が登場しました!✨ 注目の神機能はこちら: ・Jitter Filtering:動きのガタつきを数理的に除去し、滑らかさを実現 ・Slerp(球面線形補間):関節のねじれを自然にブレンドして手動修正が可能 ・Time-Travel

Another example of greed. The PRO subscription!

Local AiDGX agent

Ollama Cloud launched in September 2025 with fixed-price subscription tiers (20/month Pro, 100/month Max) for cloud-hosted inference , while local deployment remains free with unlimited local usage .

b9000

Local AiDGX agent

b9000 is a build release of llama.cpp, a C/C++ implementation for LLM inference . The release follows the project's rapid development cycle where multiple releases can be published in a single day . A

b9002

Local AiDGX agent

B9002 is a build release of llama.cpp, a C/C++ implementation for LLM inference. The release includes compiled binaries and artifacts for multiple platforms and hardware configurations, as part of the

b9004

Local AiDGX agent

llama.cpp is an LLM inference framework in C/C++ that enables efficient local execution of large language models. Build b9004 is a recent release with optimizations and improvements for supporting var

b9008

Local AiDGX agent

B9008 is a build release of llama.cpp from May 2, 2026. Llama.cpp is a C/C++ implementation of LLM inference designed to enable large language model inference with minimal setup and high performance a

b9009

Local AiDGX agent

Build b9009 is an incremental release of llama.cpp from May 2026, continuing the rapid development cycle that characterized April 2026's updates with tensor parallelism, 1-bit quantization, and expand

b9010

Local AiDGX agent

b9010 is a build release of llama.cpp, an open-source project for LLM inference in C/C++ . As a numbered build tag in the llama.cpp release system, it represents a specific development build containin

Fast & clean face swap workflow for ComfyUI (FLUX + InsightFace) — ready to use

Local AiDGX agent

A ComfyUI face swap workflow that auto-detects and aligns source and target faces, composites them together, then refines the result using FLUX image generation . Uses InsightFace for face detection a

I built Aura: a local-first AI daemon that gives your tools persistent memory, claim verification, and MCP observability

Local AiDGX agent

Aura is a local-first AI daemon that enhances tools with persistent memory capabilities, claim verification features, and Model Context Protocol (MCP) observability. The project appears designed to ru

I feel dumb for asking, but how do I get WAI-illustrious-SDXL v17 to work on Comfy?

Local AiDGX agent

This post likely addresses technical setup and troubleshooting for running the WAI-illustrious-SDXL v17 model within ComfyUI, a node-based interface for Stable Diffusion. The discussion probably cover

I need testers. Ollama Cloud Chat android app

Local AiDGX agent

A developer is seeking beta testers for 'Ollama Cloud Chat,' an Android application that integrates Ollama's cloud models with a mobile chat interface. The post likely discusses features, how to parti

Most accurate Al model for generating videos from images while preserving text?

Local AiDGX agent

This post likely discusses the challenge of generating videos from images while preserving text, as text and fine details in AI-generated videos often appear garbled or distorted. Based on the subredd

RTX 5080 with 16 GB VRAM, 64 GB RAM best quantized model for programming?

Local AiDGX agent

For programming tasks with an RTX 5080 (16GB VRAM) and 64GB RAM, optimal quantized models include Qwen 3 14B at Q6 quantization, Llama 3.1 13B at Q8, or DeepSeek R1 Distill 14B Q4, all of which fit co

This is so sick!

Local AiDGX agent

This is so sick! 【革命】動画から3Dモーションを抽出!ComfyUIの新機能が凄すぎる 動画から人物の動きを抜き出し、高品質な3Dアニメーションに変換できる「Mocap Surgeon」が登場しました!✨ 注目の神機能はこちら: ・Jitter Filtering:動きのガタつきを数理的に除去し、滑らかさを実現 ・Slerp(球面線形補間):関節のねじれを自然にブレンドして手動修

Trooper v2.1 — when your cloud LLM quota runs out, falls back to your local Ollama with context compaction

Local AiDGX agent

Trooper v2.1 is a tool that provides automatic fallback functionality from cloud-based LLM services to local Ollama instances when cloud quota limits are exceeded, incorporating context compaction to

1 May 2026

A Collective Variational Principle Unifying Bayesian Inference, Game Theory, and Thermodynamics

Local AiDGX agent

arXiv:2604.27942v1 Announce Type: new Abstract: Collective intelligence emerges across biological, physical, and artificial systems without central coordination, yet a unifying principle governing suc

Attractor FCM

Local AiDGX agent

arXiv:2604.27947v1 Announce Type: cross Abstract: In this paper an attractor FCM is created, tested, and analyzed. This FCM is neither a hebbian based nor agentic, nor a hybrid; it rather is a gradien

b8995

Local AiDGX agent

llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Release b8995 is a build/versio

b8999

Local AiDGX agent

llama.cpp b8999 is a build release from the llama.cpp project, which provides LLM inference in C/C++ . This build represents an intermediate development version in the project's rapid release cycle. T

Can Tabular Foundation Models Guide Exploration in Robot Policy Learning?

Local AiDGX agent

arXiv:2604.27667v1 Announce Type: cross Abstract: Policy optimization in high-dimensional continuous control for robotics remains a challenging problem. Predominant methods are inherently local and of

← Previous
1…5758596061…78
Next →