AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “local-ai”

GridTimelineEvolution
4,679 results
26 May 2026

Real Lighting Control with Flux 2 Klein 9B with ControlLight

Local AiDGX agent

ControlLight is a technique for achieving precise lighting control when editing images with Flux 2 Klein 9B, enabling adjustments to lighting, backgrounds, and other visual elements while maintaining

Regional Condition Custom Node for Anima model

Local AiDGX agent

Anima is a 2 billion parameter text-to-image model focused mainly on anime concepts and styles, but also capable of generating other non-photorealistic content. A Regional Condition Custom Node for An

Retrieval-Augmented Detection of Potentially Abusive Clauses in Chilean Terms of Service

Local AiDGX agent

arXiv:2605.26019v1 Announce Type: cross Abstract: Online Terms of Service often function as contracts of adhesion, creating asymmetries that may expose consumers to potentially abusive clauses. In Chi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Signs Beat Floats: Low-Rank Double-Binary Adaptation for On-Device Fine-Tuning

Local AiDGX agent

arXiv:2605.24058v1 Announce Type: cross Abstract: On-device adaptation of large language models commonly keeps a quantized base model frozen while training and deploying a small, task-specific LoRA ad

Spatio-temporal, multi-field deep learning of shock propagation in meso-structured media

Local AiDGX agent

arXiv:2509.16139v5 Announce Type: replace Abstract: Predicting the extreme hydrodynamic response of porous and architected lattice materials is a fundamental challenge in high energy density physics,

SpecPrune-VLA: Accelerating Vision-Language-Action Models via Action-Aware Self-Speculative Pruning

Local AiDGX agent

arXiv:2509.05614v3 Announce Type: replace-cross Abstract: Pruning is a typical acceleration technique for compute-bound models by removing computation on unimportant values. Recently, it has been appl

Temporal Score Rescaling for Temperature Sampling in Diffusion and Flow Models

Local AiDGX agent

arXiv:2510.01184v2 Announce Type: replace Abstract: We present a mechanism to steer the sampling diversity of denoising diffusion and flow matching models, allowing users to sample from a sharper or b

The Model Parking Tax: Quantifying the Hidden Energy Cost of Always-On GPU Model Deployment

Local AiDGX agent

arXiv:2605.23918v1 Announce Type: cross Abstract: The AI inference industry keeps models loaded in GPU memory around the clock to avoid cold-start latency, implicitly treating idle power as a fixed co

Towards Verifiable Transformers: Solver-Checkable Circuit Explanations

Local AiDGX agent

arXiv:2605.24033v1 Announce Type: new Abstract: Mechanistic interpretability often identifies circuits inside Transformer models, but explanations of those circuits are usually validated through examp

Treatment Effect Estimation with Differentiated Networked Effect on Graph Data

Local AiDGX agent

arXiv:2605.24358v1 Announce Type: cross Abstract: Estimating individual treatment effect (ITE) from observational graph data is crucial for decision-making in the fields such as commerce and medicine.

Try out the models Below 👇 Nanobana Pro: https://links.comfy.org/4f5k5lN GPT Image 2: https://links.comfy.org/4dAi4Nq

Local AiDGX agent

ComfyUI promoted two AI models available for testing: Nanobana Pro and GPT Image 2, providing direct links for users to access and try out these models through their platform. This appears to be a soc

Unlocking Apple's Private Cloud Compute: An Analysis of Privacy-Preserving Artificial Intelligence

Local AiDGX agent

arXiv:2605.24239v1 Announce Type: cross Abstract: Many existing Artificial Intelligence (AI) solutions on mobile devices rely on an extensive collection of sensitive data, raising privacy concerns and

What Are We Actually Decoding? Source Attribution for Non-Invasive Brain-to-Language Retrieval

Local AiDGX agent

arXiv:2605.24524v1 Announce Type: cross Abstract: In non-invasive neural language decoding, results can be inflated by sources that are not stimulus-evoked neural evidence: decoder priors, embedding-b

Your Embedding Model is SMARTer Than You Think

Local AiDGX agent

arXiv:2605.24938v1 Announce Type: cross Abstract: Multimodal retrieval relies heavily on single-vector retrievers, which compress rich, sequential token sequences into one single global representation

25 May 2026

2 PhaaS 2 Furious: The Evolution of Chinese-language Phishing Services

Local AiDGX agent

Written by: Jamie Collier While Russian-speaking threat actors have historically dominated the phishing-as-a-service (PhaaS) landscape, a rival ecosystem is rapidly growing within the Chinese-language

Anatomy-Guided Vision-Language Learning with Angular Prototype Separation for Multi-Label Video Capsule Endoscopy Classification Under Class Imbalance

Local AiDGX agent

arXiv:2603.17879v2 Announce Type: replace-cross Abstract: This work presents a multi-label temporal event detection framework for video capsule endoscopy (VCE) that addresses the extreme class imbalan

Android app Ollama Talk now on playstore

Local AiDGX agent

Ollama Talk is an Android app that connects to an Ollama server, enabling conversations with AI models like Llama and Mistral . The app features an intuitive chat interface with real-time conversation

Apparently this clip is too spicy! So let's try it this way! Examples of Director with LTX 2.3 and a few different techniques.

Local AiDGX agent

This post discusses workarounds for content restrictions in Stable Diffusion, specifically showcasing examples of using the Director feature with LTX 2.3 model and alternative techniques to generate c

Approximate Machine Unlearning through Manifold Representation Forgetting Guided by Self Mode Connectivity

Local AiDGX agent

arXiv:2605.22871v1 Announce Type: cross Abstract: Machine unlearning is a fundamental mechanism that enforces the right to be forgotten. Existing unlearning studies that rely on label manipulation or

b9310

Local AiDGX agent

b9310 is a release build of llama.cpp, a C/C++ implementation for LLM inference . As an intermediate build number in the llama.cpp project's continuous release cycle, it represents incremental updates

b9311

Local AiDGX agent

b9311 is a release build version of llama.cpp, an open-source C/C++ framework for large language model inference. The project enables LLM inference in C/C++ , offering optimized performance across var

b9315

Local AiDGX agent

Build b9315 is a release version of llama.cpp, a C/C++ implementation for LLM inference. As an intermediate build in the llama.cpp development sequence, b9315 likely includes bug fixes, performance im

b9319

Local AiDGX agent

llama.cpp is a C/C++ implementation for LLM inference . Build b9319 is a specific commit/version release from the llama.cpp project repository, representing a particular point in the software's develo

Building a privacy-preserving Federated Recommender system for mobile devices

Local AiDGX agent

arXiv:2605.22924v1 Announce Type: new Abstract: Serving personalized content on mobile devices has traditionally required pooling sensitive user data on centralized servers, a practice increasingly at

Built a local MCP memory server that uses Ollama to give AI coding assistants persistent memory — no cloud, no API keys, Tool and Model Agnostic

Local AiDGX agent

A developer created a local Model Context Protocol (MCP) memory server that integrates Ollama to provide AI coding assistants with persistent memory capabilities while maintaining complete privacy and

CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs

Local AiDGX agent

arXiv:2605.23344v1 Announce Type: cross Abstract: Large Vision-Language Models have shown strong multimodal reasoning capabilities, yet they remain susceptible to object hallucinations when language p

CultivAgents: Cultivating Relationship-Centered Multi-Agent Systems for Personalized Gardening

Local AiDGX agent

arXiv:2605.23193v1 Announce Type: cross Abstract: Gardening is critical to support well-being, cultural continuity, and food autonomy, yet existing digital tools often provide generic advice that over

DART: Semantic Recoverability for Structured Tool Agents

Local AiDGX agent

arXiv:2605.23311v1 Announce Type: new Abstract: When a structured tool agent fails mid-execution, the runtime faces a dilemma: replaying the entire task is safe but wasteful, while restoring from a lo

Deja Vu in Plots: Leveraging Cross-Session Evidence with Retrieval-Augmented LLMs for Live Streaming Risk Assessment

Local AiDGX agent

arXiv:2601.16027v2 Announce Type: replace Abstract: The rise of live streaming has transformed online interaction, enabling massive real-time engagement but also exposing platforms to complex risks su

Did Ollama Cloud silently nerf the usage limits?

Local AiDGX agent

This discussion likely addresses concerns about whether Ollama has quietly reduced its cloud usage limits, as the platform has revised limits twice since launch. Ollama Cloud usage is governed by sess

Evaluating PhaseNet on Teleseismic Data with MsPASS

Local AiDGX agent

arXiv:2605.22837v1 Announce Type: cross Abstract: Numerous studies have shown that the machine-learning picker PhaseNet produces accurate P and S picks on local earthquake signals, but its performance

FederatedRSF : Federated Random Survival Forests for Partially Overlapping Medical Data

Local AiDGX agent

arXiv:2605.22954v1 Announce Type: new Abstract: Multi-center survival prediction can improve robustness and generalizability, yet privacy regulations and institutional governance often prevent pooling

From Activation to Causality: Discovery of Causal Visual Representations in the Human Brain

Local AiDGX agent

arXiv:2605.23895v1 Announce Type: new Abstract: Identifying which brain regions represent a visual concept in the human brain is a central challenge in neuroscience. Existing approaches have localized

GenRecon: Bridging Generative Priors for Multi-View 3D Scene Reconstruction

Local AiDGX agent

arXiv:2605.23888v1 Announce Type: new Abstract: We introduce a new approach to high-fidelity 3D scene reconstruction from multi-view RGB images that tightly couples reconstruction with a strong genera

HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation

Local AiDGX agent

arXiv:2605.23043v1 Announce Type: new Abstract: Agentic text-simulation systems write in sequence, with each item becoming possible context for later steps. That makes uncertainty path-dependent: an e

How Far Will They Go? Red-Teaming Online Influence with Large Language Models

Local AiDGX agent

arXiv:2605.22880v1 Announce Type: cross Abstract: As large language model (LLM)-based agents increasingly participate in online discourse, red-teaming their capacity to support political influence cam

ObjectCache: Layerwise Object-Storage Retrieval for KV Cache Reuse

Local AiDGX agent

arXiv:2605.22850v1 Announce Type: cross Abstract: Prefix KV caching has become a key mechanism in LLM serving: it reduces time to first token (TTFT) by avoiding redundant computation across requests t

Ollama com 3x de 3060 12vram em uma placa Machinist ,Xeon com 32gb de memória ram,

Local AiDGX agent

This Reddit post discusses a technical setup for running Ollama with three NVIDIA RTX 3060 GPUs (each with 12GB VRAM) on a Machinist motherboard paired with a Xeon processor and 32GB of system RAM. Th

OpenStudio - Hybrid local/cloud (openrouter) AI router

Local AiDGX agent

OpenStudio is a hybrid AI router that combines local model inference with cloud-based model access through OpenRouter, a unified API providing access to hundreds of AI models through a single endpoint

PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA

Local AiDGX agent

arXiv:2605.23559v1 Announce Type: cross Abstract: Whole-slide image visual question answering (WSI-VQA) frames pathology as an extreme-context search problem: to answer a free-form clinical query, a s

PixlStash 1.3: grid loading speed, JoyCaption and bulk tag selections with your chosen model

Local AiDGX agent

PixlStash 1.3 is a Python-based image management and tagging web app that improves grid loading performance and introduces JoyCaption integration for AI-powered image captioning. The update adds suppo

Preisach Attention: A Hysteretic Model of Sequential Memory

Local AiDGX agent

arXiv:2605.23603v1 Announce Type: cross Abstract: We introduce the Preisach Attention Layer (PAL), a novel sequence modelling architecture grounded in the classical Preisach hysteresis operator from m

SCOPE: Simulating Cross-game Operations in Playable Environments for FPS World Models

Local AiDGX agent

arXiv:2605.23345v1 Announce Type: new Abstract: Interactive world models for first-person shooter (FPS) games must resolve high-frequency overlapping control signals at every frame without disrupting

Spectral-inspired Operator Learning with Limited Data and Unknown Physics

Local AiDGX agent

arXiv:2505.21573v3 Announce Type: replace-cross Abstract: Learning PDE dynamics from limited data with unknown physics is challenging. Existing neural PDE solvers either require large datasets or rely

The Attribution Contract: Feature Attribution for Generative Language Models

Local AiDGX agent

arXiv:2605.23080v1 Announce Type: new Abstract: Feature attribution methods promise to identify which input features matter for a model output. In generative language models, however, it is often uncl

The CLI natively displays “Thinking…” Is there a way to access the raw reasoning stream?

Local AiDGX agent

The raw reasoning stream can be accessed through the message.thinking field in the API response or the thinking endpoint field, which contains the reasoning trace separately from the final answer. Use

v0.30.0-rc25

Local AiDGX agent

v0.30.0-rc25 is a pre-release version of Ollama that changes the architecture to directly support llama.cpp instead of building on top of GGML, and allows for compatibility with GGUF file format. MLX

Vision-Based Agile Landing on Turbulent Waters

Local AiDGX agent

arXiv:2605.23717v1 Announce Type: new Abstract: Autonomous landing of Unmanned Aerial Vehicles on maritime vessels is challenging due to the coupled motion of the vehicle and landing platform in open-

ZipMoE: Efficient On-Device MoE Serving via Lossless Compression and Cache-Affinity Scheduling

Local AiDGX agent

arXiv:2601.21198v2 Announce Type: replace-cross Abstract: While Mixture-of-Experts (MoE) architectures substantially bolster the expressive power of large-language models, their prohibitive memory foo

24 May 2026

300,000 AI builders filled their hardware profile on @huggingface and we're sharing the results: http://hf.co/hardware. Excited to see how i…

Local AiDGX agent

300,000 AI builders filled their hardware profile on @huggingface and we're sharing the results: http://hf.co/hardware. Excited to see how it evolves in the coming months especially with the explosion

b9301

Local AiDGX agent

llama.cpp is an open-source C/C++ project for LLM inference , and build b9301 is a specific release version in the project's development history. This build number represents an incremental developmen

b9305

Local AiDGX agent

Release b9305 was published on May 24, 2026 , featuring CMake UI build fixes and improvements for multiple platforms including macOS Apple Silicon and Linux architectures . The release provides prebui

ComfyUI_SamplingUtils plus Klein_9B for quick style change

Local AiDGX agent

ComfyUI_SamplingUtils with Klein_9B enables rapid image style changes within ComfyUI workflows, supporting text-to-image generation and reference-based editing for quick style transformation and conte

Dual 3090s?

Local AiDGX agent

A discussion about difficulties getting Ollama to fully utilize dual NVIDIA RTX 3090 GPUs when running various language models including 8B and 70B parameter models. Users report that despite Ollama r

LongCat-Video-Avatar 1.5 Release

Local AiDGX agent

LongCat-Video-Avatar 1.5 is an upgraded open-source framework that prioritizes extreme empirical optimization and production-readiness for audio-driven human video generation. The v1.5 release replace

My daily average local model token burn is 17M They have become tremendously useful.

Local AiDGX agent

Clem Delangue reports consuming approximately 17 million tokens daily when running local language models, indicating substantial usage of on-device AI inference. He expresses satisfaction with the uti

Tired of the Cloud Terminal Hassle? Building a Universal 'Stability Matrix' but for Cloud GPUs (RunPod, Vast...) 🚀

Local AiDGX agent

A Reddit discussion from the StableDiffusion community about challenges with managing cloud GPU services like RunPod and Vast, with someone proposing to build a unified 'Stability Matrix' tool to simp

v0.30.0-rc24

Local AiDGX agent

v0.30.0-rc24 is a pre-release version that changes Ollama's architecture to directly support llama.cpp instead of building on top of GGML, enabling compatibility with the GGUF file format. MLX is used

23 May 2026

A Mechanistic Explanatory Strategy for XAI

Local AiDGX agent

arXiv:2411.01332v5 Announce Type: replace Abstract: Despite significant advancements in XAI, scholars note a persistent lack of solid conceptual foundations and integration with broader scientific dis

AsymFLUX.2-klein-9B is all about textures

Local AiDGX agent

AsymFLUX.2-klein-9B is a pixel-space text-to-image model finetuned from FLUX.2-klein-base-9B using the AsymFlow method. The model is noted for producing sharp textures with strong adherence and compos

← Previous
1…4142434445…78
Next →