AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “local-ai”

GridTimelineEvolution
4,645 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Companies

CompanyAnthropic8 recent entries
26 Jul 2026Will prices finally go down?

I am seeing more and more videos as posts about how OpenAI is in complete financial ruin, Anthropic isn't much better. Their expenses go with the revenue they make etc etc. Meta made big investments i

→26 Jul 2026Karparthy removed Anthropic from his bio

Andrej Karpathy, a prominent advocate for open-source AI and a co-founder of OpenAI, appears to have removed Anthropic from his X bio, suggesting he may have left the company. Karpathy joined Anthropi

→27 Jul 2026
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
NYT: Protect America’s lead in the A.I. race.

“China is working hard to catch up, and the United States should take steps to keep its advantage. Most important, it should continue to prohibit American companies from selling the most advanced chip

→27 Jul 2026Give any Ollama-compatible client session memory + a shared knowledge wiki by swapping the chat URL

Hey folks — I built ContextMemory, an open-source agentic context gateway for apps that already talk to LLMs. The idea is simple: keep your existing POST /api/chat client (Ollama wire format), point i

→28 Jul 2026Now, this: 1,100 current/former frontier-AI employees sign a petition calling for US gov't to step in for 'pacing' frontier development

So, it appears that this is the week of open letters in AI🥲... an open letter signed by current and former employees of OpenAI, Anthropic and Google primarily - calling for a slow-down in frontier AI

→7 Aug 2026A visualization of LLM API costs to ask for local resources

I have not been successful with management to get funding for local resources despite bringing forth solid arguments about data sovereignty and related architectures. What actually succeeded in gettin

→9 Aug 2026US data center bans top 500, up from 300+ in late June, as New York and Texas join cities and counties pushing back against data center development (Shane Burke/The Information)

Shane Burke / The Information: US data center bans top 500, up from 300+ in late June, as New York and Texas join cities and counties pushing back against data center development — Local government re

→12 Aug 2026You can now use Ollama as a provider in GitHub Copilot for JetBrains. https://github.blog/changelog/2026-08-11-copilot-memory-and-ollama-in-…

GitHub announced on August 12 2026 that users can now integrate Ollama as a provider in **GitHub Copilot for JetBrains**. This update allows JetBrains developers to switch to or add locally‑hosted (or

CompanyOpenAI8 recent entries
5 Aug 2026SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

→5 Aug 2026None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + offi…

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + official quantized builds, Day 0 Diffusers — the reference Pytho

→8 Aug 2026Building a zero-dependency C inference engine for BitNet (1.58-bit) - lessons from hitting 36 tok/s on a Xeon CPU

Over the past few months I have been building a CPU-first inference engine from scratch in pure C99 (no Python, no CUDA, no BLAS, just GCC and make). The focus has been running 1.58-bit ternary models

→9 Aug 2026It took two years, but we finally have a 'local Sora'

Who remembers when OpenAI previewed Sora two years ago and the quality felt unreal? We had never seen anything like it. Back then, Sora 1 didn't even generate audio and was heavily censored. Prompt: i

→11 Aug 2026Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

→11 Aug 2026Multi-modal Interactive Control of Robotic Arm based on Offline Large Language Models

arXiv:2608.08183v1 Announce Type: new Abstract: Large Language Models (LLMs) have significantly revolutionized the modern society with numerous advanced interactions between humans and AI agents, wher

→11 Aug 2026AquiLLM: An Architecture for Supporting Tacit Knowledge Capture in Research Groups

arXiv:2608.08883v1 Announce Type: new Abstract: Recent advances in retrieval-augmented generation (RAG) and large language models (LLMs) enable researchers to integrate AI into scientific workflows. H

→12 Aug 2026How to do clean uninstall of chatgpt desktop app on windows?

ChatGPT desktop app will not download images even after a complete reinstall I am on Windows 11 and the Download button in the ChatGPT desktop app does nothing when I try to download generated images.

CompanyGoogle6 recent entries
10 Apr 2026OpenClaude com Ollama Cloud

OpenClaude is an open-source coding-agent CLI, forked from the Claude Code source, that adds an OpenAI-compatible provider shim enabling use of GPT-4o, DeepSeek, Gemini, Ollama local models, and 20...

→13 Apr 2026Tried Ollama Cloud, just realize only Kimi model accept images

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

→14 Apr 2026AI tool to analyze a video and generate a prompt?

This Reddit thread from r/StableDiffusion discusses the community's search for AI tools capable of analyzing an existing video and automatically generating a descriptive text prompt from it — essentia

→8 May 2026Chrome's 4GB AI model isn't new, but you're not wrong for being confused

Google has offered Gemini Nano for Chrome since 2024 as a lightweight, on-device model , but users reasonably expect the visible AI Mode to use the on-device model with queries staying local, when in

→10 Jun 2026Ideogram4 vs Flux.2 Dev vs GPT Image 2 vs Nano Banana Pro

A practical comparison of Ideogram 4.0, Nano Banana Pro, and GPT Image 2 across text rendering, design, photorealism, and real creator workflows. Ideogram 4.0 is the first open-weight challenger to ra

→23 Jul 2026Koopman Dreamer: Spectrally Constrained Latent Dynamics for Stable World-Model Imagination

arXiv:2607.19719v1 Announce Type: new Abstract: Latent world models improve sample efficiency in continuous control by optimizing policies over imagined latent trajectories, but common neural transiti

CompanyMeta8 recent entries
29 Jul 2026A Control System, a Dataset, and a Recipe for Making Frozen LLM Agents Learn a Domain

arXiv:2607.25415v1 Announce Type: new Abstract: Production LLM agents are increasingly assembled from a frozen model wrapped in a harness: a prompt template, a tool set, a memory/retrieval layer, a pl

→30 Jul 2026Conformal Changepoint Localization and Root Cause Analysis with Corrupted Observations

arXiv:2607.26481v1 Announce Type: new Abstract: Detecting when the statistical behavior of an engineered system changes, and identifying which component is responsible, are core problems in the monito

→3 Aug 2026Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

→3 Aug 2026MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language Models

arXiv:2607.29561v1 Announce Type: cross Abstract: Symbolic Regression (SR) aims to discover analytical equations from observational data and plays a central role in scientific modeling. While recent L

→3 Aug 2026Best Friends, Not Forever: Evaluating Long-Horizon Persona Collapse and Behavioral Drift in AI Companions

arXiv:2607.28818v1 Announce Type: new Abstract: As AI companions increasingly mediate repeated social interaction, users may rely on a stable role and shared history, yet locally acceptable replies do

→11 Aug 2026Agentic AI-driven Immersive Simulation: A Knowledge-Aware Virtual Training Platform forHigh Dose Rate (HDR) Brachytherapy

arXiv:2608.08163v1 Announce Type: new Abstract: The convergence of the Metaverse and Large Language Model (LLM)-based AI agent is catalyzing a shift toward autonomous, immersive, and personalized peda

→12 Aug 2026VidForensics-M1: Meta-Detection Reinforcement Learning with Verifiable Temporal Grounding for AI-Generated Video Forensics

arXiv:2608.11201v1 Announce Type: new Abstract: Recent advances in video generation models have significantly improved the realism of synthetic videos, blurring the boundary between generated and auth

→12 Aug 2026Quantum Coordination Advantages in AI State-Tracking Tasks: Semantic Compilation and Latent Memory

arXiv:2608.11066v1 Announce Type: cross Abstract: We prove inference-time quantum coordination advantages for specified AI state-tracking tasks. A solver compresses semantic history into a future-acce

CompanyMistral8 recent entries
14 Apr 2026Agents in Ollama and Langflow

This Reddit post from r/ollama likely discusses how to build and run AI agents locally by combining Ollama — which handles local model serving to keep data private — with Langflow's visual, drag-and-d

→4 May 2026OneTrainer now supports Ernie LoRA

OneTrainer is a one-stop solution for all diffusion training needs. The tool now supports the Ernie Image model, which can be trained using LoRA (Low-Rank Adaptation) methods. This adds support for tr

→4 May 2026b9014

llama.cpp enables LLM inference in C/C++ , and release b9014 represents a build revision or commit snapshot from the llama.cpp repository on GitHub. This release would include bug fixes, feature impro

→12 May 2026Uncensored LLM

Uncensored LLMs are architectures that have been modified or fine-tuned to remove standard safety alignment layers (guardrails) that limit a model's ability to discuss sensitive topics. Ollama offers

→25 May 2026Android app Ollama Talk now on playstore

Ollama Talk is an Android app that connects to an Ollama server, enabling conversations with AI models like Llama and Mistral . The app features an intuitive chat interface with real-time conversation

→5 Jun 2026What are the most capable LLM models I can run on my laptop?

A discussion on r/ollama exploring which high-performance LLM models can be effectively run locally on standard laptop hardware , likely covering model size comparisons, hardware requirements, and per

→5 Jun 2026b9523

b9523 is a release build of llama.cpp, an open-source tool for LLM inference in C/C++ that enables language model execution with minimal setup on a wide range of hardware locally and in the cloud. The

→9 Jul 2026https://ollama.com/blog/all-aboard-open-models

Ollama announced support for running multiple open-source language models locally, emphasizing accessibility and ease of deployment for users who want to use AI models without relying on cloud service

CompanyxAI8 recent entries
26 Jul 2026Will prices finally go down?

I am seeing more and more videos as posts about how OpenAI is in complete financial ruin, Anthropic isn't much better. Their expenses go with the revenue they make etc etc. Meta made big investments i

→28 Jul 2026Context-Aware Concept Distillation for Trustworthy Flood Prediction

arXiv:2607.23237v1 Announce Type: cross Abstract: Effective flood risk management relies on accurate forecasting, yet the 'black box' nature of stateof-the-art Deep Learning models creates a barrier t

→28 Jul 2026Beyond Local Inspection: Global, Guideline-Grounded Evaluation of Post-hoc XAI Methods for ECG Classification

arXiv:2607.24035v1 Announce Type: cross Abstract: Explainable AI (XAI) is used to assess whether artificial intelligence models rely on meaningful patterns, yet explanations that appear plausible for

→29 Jul 2026TaylorPODA: A Taylor Expansion-Based Method to Improve Post-Hoc Attributions for Opaque Models

arXiv:2507.10643v4 Announce Type: replace-cross Abstract: Post-hoc model-agnostic local attribution (LA) methods have been widely adopted to explain opaque AI models by quantifying feature-wise contri

→4 Aug 2026Paris as a 15-Minute City: An Explainable AI Perspective

arXiv:2608.00815v1 Announce Type: new Abstract: The 15-minute city promotes access to everyday services within a short walk or bicycle ride, but its relationship with observed mobility remains difficu

→4 Aug 2026I added a verify-before-load safety check for Ollama models

I maintain llm-checker, and I’ve added structural model-file validation for Ollama. Ollama stores downloaded models as local blobs. If one is truncated, malformed, or has invalid internal offsets, you

→10 Aug 2026Human-Centered Explainable AI for TinyML Edge Devices: A Pareto-Based Selection Framework with LLM-Guided Design

arXiv:2608.07091v1 Announce Type: cross Abstract: Edge Artificial Intelligence (Edge AI) enables the deployment of AI models directly on local edge devices, while such deployments are subject to stric

→11 Aug 2026xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally (Cade Metz/New York Times)

Cade Metz / New York Times: xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally — Igor Babuschki

CompanyDeepSeek8 recent entries
4 May 2026Sentinel: an open-source local-first desktop app for AI coding

Sentinel is a local-first AI coding desktop application built with Rust and Tauri that automatically routes coding tasks to appropriate models based on task complexity. Instead of swapping between mul

→4 May 2026Did they shut down deep seek cloud for free users?

DeepSeek V3 and R1 API free tier includes 500M tokens per month , indicating free access remains available for API users. The search results focus on a major service outage in March 2026 and the recen

→4 May 2026b9014

llama.cpp enables LLM inference in C/C++ , and release b9014 represents a build revision or commit snapshot from the llama.cpp repository on GitHub. This release would include bug fixes, feature impro

→5 May 2026Parllama -- a terminal UI for Ollama model management and multi-provider LLM chat

Parllama is a TUI (Text UI) application designed for easy management and use of Ollama-based LLMs that also works with major cloud-provided LLMs. It provides core model management features including f

→20 May 2026@ollama + @deepseek_ai v4 pro handled entire monthly dev reports on Eigent. github prs → word doc → slack message → sent to product-release …

@ollama + @deepseek_ai v4 pro handled entire monthly dev reports on Eigent. github prs → word doc → slack message → sent to product-release channel. in just one prompt. fully local. the full walkthrou

→29 May 2026b9414

b9414 is a release build of llama.cpp that includes improvements to CUDA PTX version checking , which helps prevent incorrect kernel dispatch on different GPU architectures. This build also adds suppo

→29 May 2026b9413

Release b9413 includes a CUDA fix that checks PTX version on the host side to guard PDL dispatch, addressing an issue where incorrect dispatching could occur on newer GPU architectures like sm_90/sm_1

→5 Jun 2026b9523

b9523 is a release build of llama.cpp, an open-source tool for LLM inference in C/C++ that enables language model execution with minimal setup on a wide range of hardware locally and in the cloud. The

CompanyNVIDIA8 recent entries
8 Aug 2026Has anyone here fiddled with TPUs for inference ?

I discovered recently that Google uses their own TPUs, like tiny ASIC cards like the toy ones that existed for bitcoin. And while it sounds inefficient the fact they use thousands of them because...th

→9 Aug 2026I Turned My Underused Gaming Laptop Into a Local AI Workstation

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

→10 Aug 2026v0.32.8

v0.32.8 is an Oct 10, 2023 release of the ollama repository on GitHub, following a pre‑release tag v0.32.8‑rc0. The update adds Muse Glimmer support for NVIDIA, AMD and additional platforms, with the

→10 Aug 2026Dual-Node NVIDIA DGX Spark over Tailscale: A Remote-Access Testbed for Distributed LLM Training and Cyber-Threat-Intelligence Fine-Tuning

arXiv:2608.07226v1 Announce Type: cross Abstract: Compact AI systems make local language-model experimentation increasingly accessible, yet practical evidence for multi-node training on desktop-class

→11 Aug 2026xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally (Cade Metz/New York Times)

Cade Metz / New York Times: xAI co-founder Igor Babuschkin's River AI raised $1B led by General Catalyst to build home or small business computer servers capable of running AI locally — Igor Babuschki

→11 Aug 2026Rumored 50-series Super refresh bumps everything +50% VRAM

leaked Super specs have the 5070 Ti and 5080 going 16GB to 24GB and the 5070 to 18GB, thanks to the new 3GB GDDR7 modules. 24GB on a Ti-class card is actually the number people here have been waiting

→11 Aug 2026NVIDIA and Local AI Community Fuel Open Source Models and Intelligent Agents

The open source ecosystem is making it easier for AI enthusiasts and developers to build, customize and run increasingly capable agents locally. Throughout August, NVIDIA is celebrating the partners a

→12 Aug 2026LiquidAI/LFM2.5-VL-3B · Hugging Face

LFM2.5-VL-3B is a multimodal variant of LFM2.5, a family of hybrid models designed for on-device deployment. It builds on LFM2-VL-3B with further mid- and post-training. LFM2.5-VL-3B can process both