AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
1,446 results
Local Ai

AnimaYume - Anima finetune.

DGX agent

AnimaYume is a text-to-image model fine-tuned from Anima, a 2-billion-parameter anime-focused image generation model developed by CircleStone Labs in collaboration with Comfy Org, which is itself buil

local-air-stablediffusion
13 Apr 2026
Local Ai

Face vs body Zit Lora

DGX agent

This Reddit post from r/StableDiffusion discusses a community comparison or showcase involving a 'Zit LoRA' — a Stable Diffusion LoRA model — examining how it performs differently when applied to face

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
local-air-stablediffusion
13 Apr 2026
Local Ai

NO MORE PAYING FOR API! NEW SOLUTION!

DGX agent

A Reddit post from the r/ollama community discussing a free alternative to paid AI API services, likely centered around using Ollama to run large language models locally. The post probably highlights

local-air-ollama
13 Apr 2026
Model Releases

Ollama / Mistral with MCP to Mempalace

DGX agent

This Reddit post from r/ollama discusses integrating Ollama-served Mistral with MemPalace — a free, locally-run AI memory system — via the Model Context Protocol (MCP). MemPalace runs entirely on a us

model-releasesr-ollama
13 Apr 2026
Local Ai

Can't get a good coding setup on Macbook Pro M3 Max 36GB

DGX agent

This Reddit thread from r/ollama discusses a user's difficulty achieving a satisfactory local AI coding assistant setup using Ollama on a MacBook Pro M3 Max with 36GB of unified memory. The discussion

local-air-ollama
12 Apr 2026
Local Ai

Greg Rutkowski Anima Lora from Circlestone Labs (Anima makers) with training params

DGX agent

This Reddit post discusses a LoRA trained to emulate the style of digital artist Greg Rutkowski, built on top of Circlestone Labs' Anima model — a 2 billion parameter text-to-image model created via a

local-air-stablediffusion
12 Apr 2026
Local Ai

Tile upscale controlnet with Z-Image-Base? Has anybody achieved good results?

DGX agent

This Reddit thread from r/StableDiffusion discusses community experiences using ControlNet Tile upscaling in combination with Z-Image-Base, a Stable Diffusion base model. ControlNet Tile models are of

local-air-stablediffusion
12 Apr 2026
Local Ai

Z-Image Turbo Checkpoint - Deedeemegadoodo Edition

DGX agent

The 'Z-Image Turbo Checkpoint - Deedeemegadoodo Edition' is a community-shared checkpoint on r/StableDiffusion based on Z-Image Turbo, a distilled version of Z-Image, a 6B image model developed by the

local-air-stablediffusion
12 Apr 2026
Local Ai

fine-tune LTX 2.3 with his own dataset?

DGX agent

This r/StableDiffusion thread discusses how to fine-tune the LTX-Video 2.3 model on a personal dataset, with the primary approach being LoRA (Low-Rank Adaptation), which fine-tunes a large AI model on

local-air-stablediffusion
11 Apr 2026
Local Ai

I got trolled

DGX agent

A Reddit post from the r/StableDiffusion community in which a user shares an experience of being trolled, likely related to AI image generation workflows, model recommendations, or settings advice. Th

local-air-stablediffusion
11 Apr 2026
Model Releases

I reduced my token usage by 178x in Claude Code!!

DGX agent

A Reddit post from r/ollama describing how a user dramatically reduced their Claude Code token consumption by 178x, likely by routing simpler or lower-stakes tasks to a locally-run model via Ollama in

model-releasesr-ollama
11 Apr 2026
Local Ai

SDXL workflow

DGX agent

This Reddit post from r/StableDiffusion discusses an SDXL (Stable Diffusion XL) image generation workflow, likely covering pipeline setup, prompting strategies, and tool configurations using interface

local-air-stablediffusion
11 Apr 2026
Model Releases

Advanced inpaint/edit Klein/Qwen workflows

DGX agent

A Reddit post on r/StableDiffusion discussing advanced ComfyUI workflows that combine the FLUX Klein and Qwen Image Edit models for precision inpainting and image editing tasks. FLUX Klein offers ...

model-releasesr-stablediffusion
10 Apr 2026
Local Ai

Bad news on Happy Horse from twitter

DGX agent

HappyHorse-1.0 is a pseudonymous AI video generation model that appeared on April 7, 2026, topping the Artificial Analysis Video Arena leaderboard in both text-to-video and image-to-video (no audio...

local-air-stablediffusion
10 Apr 2026
Local Ai

glm 5.1 is doing well

DGX agent

GLM-5.1 is Z.ai's next-generation flagship model for agentic engineering, built on a 754-billion parameter Mixture-of-Experts architecture with 40 billion active parameters per token, a 200,000-tok...

local-air-ollama
10 Apr 2026
Model Releases

Possible memory leak in Ollama when using Claude Code?

DGX agent

Users in the r/ollama community have reported a possible memory leak occurring in Ollama when it is used as a backend with Claude Code, with Ollama runner processes not always being properly termin...

model-releasesr-ollama
10 Apr 2026
Local Ai

Should we be optimizing for limited compute instead of more parameters? Thoughts?

DGX agent

"The search did not return the specific Reddit thread. However, I can provide a summary based on what the topic is broadly about within the local-AI/Ollama community context:

local-air-ollama
10 Apr 2026
Model Releases

Idea for a deepseek-v4-flash-0731 backed automated research workflow to be leveraged via qwen3.6/3.8 27b for difficult tasks that require highly technical, not easy to find information.

DGX agent

Sometimes you have tasks that are outside of your expertise and the idea is this workflow automation could be leveraged to manage to have local AI figure it out using research from his workflow gather

model-releasesr-localllama
12 Aug 2026
Model Releases

12GB VRAM gang, what's our plan?

DGX agent

Seems like we're limited to qwen finetuned MoEs for now. Looking at the current landscape - focus seems to be on dense models (muse glimmer 30b, qwen 3.8 27b) for smaller setups. Is upgrading to 24GB

model-releasesr-localllama
11 Aug 2026
Model Releases

Motif-Technologies/Motif-3 official realese

DGX agent

Motif-Technologies is one of the tech company participated South Korea's AI Foundation Model project.(독파모) Upstage(Solar Series), LG AI Research(EXAONE Series), and SKT(A.X Series) are the competitors

model-releasesr-localllama
10 Aug 2026
Model Releases

Please Share Your Experience About Muse Glimmer

DGX agent

I have a classic test for local LLM's. I asked for 8 ball pool game with only one HTML file and Muse Glimmer spend 21k Token(I m using full context so 128k) and only created a 220 lines of HTML and sa

model-releasesr-localllama
10 Aug 2026
Model Releases

Tested Muse Glimmer locally on coding with OpenCode & agentic work

DGX agent

Ran the model with quants (Q4) by Unsloth with latest (build from master) llama.cpp server. It takes ~20GB ram running on M5 Pro with 48GB at about 17t/s. Didn't do any reasoning loops/overthinking. O

model-releasesr-localllama
10 Aug 2026
Model Releases

AMD llama.cpp: reducing MTP buffer overhead gave me 64K → 149K context for Qwen 27B

DGX agent

Available context length with and without the patch: Model: QWEN 27B ROCm stock patched Vulkan stock patched IQ4_XS Pure, single 16GB GPU 19.456 76.032 68,352 78,592 Q6_K_L on 16GB + 12GB 64,256 149,2

model-releasesr-localllama
9 Aug 2026
Model Releases

DeepSeek v4 Flash 0731 locally on CPU

DGX agent

After seeing the benchmark results for the full release of DS v4 Flash 0731, I replaced my 2 x 16GB DDR4 ram sticks with 2 x 32GB DDR4 ram sticks to get a max supported of 128 GB RAM, in hope to be ab

model-releasesr-localllama
9 Aug 2026
Model Releases

M3 16GB running Ollama (Qwen 9B) is extremely slow (10-12 mins per task). Am I doing something wrong?

DGX agent

Hey everyone, I constantly see high praise for M3 and M4 Macs for local LLM inference, even the base/16GB models. However, my experience has been quite different, and I'm trying to figure out if I hav

model-releasesr-ollama
9 Aug 2026
Model Releases

Underestimated budget solution: radeon 780m iGPU

DGX agent

There are so many posts where people complaining about high prices and asking for solution <= 1000 EUR. So, there is one solution to consider: PC/mini PC/laptop on Ryzen 7 260/Ryzen 9 8945HX/etc CPU w

model-releasesr-localllama
9 Aug 2026
Model Releases

Claude Code in 9 lines python

DGX agent

I was wondering what a minimal coding agent implementation would look like that can be used like Claude Code or Codex Not feature-by-feature of course but basically stripping everything out that is no

model-releasesr-localllama
8 Aug 2026
Model Releases

Auto-fit vs tuned MoE offload: 564 → 1330 pp tok/s, unchanged decode (Qwen3.6-35B-A3B Q6 / RTX 3090)

DGX agent

TL;DR: On a Qwen3.6-35B-A3B Q6 setup sized for 64K context on a 24GB RTX 3090, spilling eight MoE expert layers to CPU freed enough VRAM to increase -b from 512 to 1024 and -ub from 128 to 512. Prompt

model-releasesr-localllama
6 Aug 2026
Model Releases

Best llama cpp flags to run Deepseek-flash 0731

DGX agent

Hi all. These are my system specs: dual xeon e5 2696 v2 , 160gb DDR3 ram ECC(1600mhz), 3 gpus: 3060 12gb, p100 16gb, 3050 6gb. And a 400gb nvme sdd RAID0, 3000 mb/s. The model is Deepseek-flash-0731 U

model-releasesr-localllama
6 Aug 2026
Model Releases

KV cache quantization benchmarks: 413 pairs tested on Qwen 3.6 27B, Gemma 4 31B. KLD with BeeLlama.cpp v0.4.0: KVarN 6-bit beats q8_0, precision tail 1024 dominates

DGX agent

Link to the article: KV Cache Quantization Benchmarks: KVarN, Precision Tail KLD benchmarks with BeeLlama.cpp v0.4.0, fork of llama.cpp with more KV cache quantization options. Models: Qwen 3.6 27B Q5

model-releasesr-localllama
6 Aug 2026
Model Releases

nvidia/NVIDIA-Nemotron-Parse-2.0 · Hugging Face

DGX agent

NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information. Given a Red, Green, Blu

model-releasesr-localllama
6 Aug 2026
Model Releases

A ultra-lightweight mini agent - zero framework and local/ollama first

DGX agent

https://github.com/mohsinkaleem/agent-mini.git A minimal 3k lines, local-first AI agent you can actually understand and extend. Optimized for smaller local models like qwen 3.6 4b or 9b pip install ag

model-releasesr-ollama
5 Aug 2026
Model Releases

Building a Fully Local PDF Read-Aloud & PDF-to-Audiobook Desktop App with Kokoro 82M, Qwen, and llama.cpp

DGX agent

Hey everyone, I’ve been building Speechfony - a desktop app for reading PDFs (and EPUBs) with offline text-to-speech. Open a document, listen sentence-by-sentence with highlighting, or export selected

model-releasesr-localllama
5 Aug 2026
Model Releases

DeepSeek V4 Flash 0731 at 10–17 t/s (nothink) on MacBook M5 Pro **64GB***, partly via SSD streaming

DGX agent

Inspired by a post from u/giveen I motivated claude (no patinence on my side to work through everything myself) to help me get DS running on my MacBook M5 Pro 64GB and it exceeded my expectations.. be

model-releasesr-localllama
5 Aug 2026
Model Releases

Stable Diffusion might actually be remembered in the history books, and I don’t think that’s an overstatement

DGX agent

Hear me out before you roll your eyes. We tend to only recognize turning points in hindsight. Nobody in 1993 thought the Mosaic browser would be a history book moment, but the web is. I think Stable D

model-releasesr-stablediffusion
5 Aug 2026
Model Releases

Thinking of buying more DRAM right now...

DGX agent

So I'm looking at https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF and I realize my 128GB of DRAM just isn't cutting it for this (incredibly powerful) model. If only I had another 64GB, I th

model-releasesr-localllama
5 Aug 2026
Model Releases

inclusionAI/Ling-3.0-flash · Hugging Face

DGX agent

The Ling-3.0-flash MoE is now open-weighted at 124B A5B params. I know the original announcements were before the Kimi K3, DeepSeek-V4-Flash and Qwen3.8 hype, but this model might still have a good ni

model-releasesr-localllama
4 Aug 2026
Model Releases

I compared MinerU, Granite-Docling, and PaddleOCR-VL on 12 PDF-parsing capabilities using 6 document types

DGX agent

I tested them by sending the 6 documents, each meant to represent a different document type, through my own webapp and comparing every output against the source. All ran on the same L4 GPU. The docume

model-releasesr-localllama
3 Aug 2026
Local Ai

I got tired of ad-filled mobile wrappers for Ollama, so I built PocketLLM Lite an open-source, offline Android client (Local GGUF, SKILL.md plugins, local RAG)

DGX agent

Hey, Like a lot of people here, I use local models via Ollama on my desktop/server and wanted a mobile client that actually felt responsive, worked offline, and respected privacy. Most apps on the Pla

local-air-ollama
3 Aug 2026
Model Releases

V4-Flash-0731 - vibes after first weekend of use

DGX agent

Spent way too much time with V4-Flash-0731 this weekend and wanted to share my vibes as briefly as possible. I sent it through a bit of real-work and some of my personal benchmarks. My quick thoughts

model-releasesr-localllama
3 Aug 2026
Model Releases

PSA: llama.app, Mac app and llama serve from llama.cpp

DGX agent

https://llama.app/ Been using llama.cpp for years now and im on here all the time (im a mod..), but somehow I totally missed that llama.app exists and its official from the HF/llama.cpp team. So posti

model-releasesr-localllama
2 Aug 2026
Model Releases

DS4 flash 0731 - Acquarium Panel Failure - Q3_K_XL Unsloth

DGX agent

https://preview.redd.it/1a39x4zivqgh1.png?width=1550&format=png&auto=webp&s=de591c039cc18782a6b5d8e402fdc1594be05132 start C:llmllamam5uildinllama-server.exe --model 'H:UD-Q3_K_XLDeepSeek-V4-Flash-073

model-releasesr-localllama
1 Aug 2026
Model Releases

Possible to create accurate medieval woodcut style art?

DGX agent

Wondering if it's possible to actually produce ai art works that are indistinguishable from authentic medieval woodcut illustrations like the one attached. All the AI attempts I've seen at re creating

model-releasesr-stablediffusion
1 Aug 2026
Model Releases

Qwen 3.6 27B Q5 on 3x2080ti: 55tps with llama.cpp. Can I squeeze out more?

DGX agent

CPU: Threadripper 3970X RAM: 128GB DDR4 GPUs: 3x2080ti 11GB The current best parameters to run it: llama-server --model Qwen3.6-27B-Q5_K_S.gguf --n-gpu-layers 999 --split-mode tensor --flash-attn on -

model-releasesr-localllama
1 Aug 2026
Model Releases

DeepSeek v4 Flash for DS4 (DwarfStar) GGUF w/ DSpark MTP Head

DGX agent

I'm an avid user of Deepseek v4 Flash via antirez's DS4 DwarfStar inference engine, and so when the new checkpoint dropped, the first thing I did was rent a cloud box and spin up a quantization for us

model-releasesr-localllama
31 Jul 2026
Local Ai

I built a hybrid Transformer–SSM LLM agent with a local CLI, active control, and run receipts

DGX agent

I’m one of the builders of LOLM, a hybrid Transformer–SSM model and agent system from Qira. The model separates surface token processing from persistent latent-state tracking. An NFET controller can s

local-air-ollama
31 Jul 2026
Model Releases

I predict DeepSeek V4 Flash 0731's Artificial Analysis score to be 57 ± 1 point (Kimi K3 Level)

DGX agent

Deepseek's new model V4 Flash 0731 is much better, I (Claude lol) did a bit of linear regression with a leave one out style verification to predict its AA Score, and that puts it at Kimi K3 level, whi

model-releasesr-localllama
31 Jul 2026
Model Releases

Minimum VRAM GPU to run DeepSeek-V4-Flash-0731 Q4_K_XL at around 30 t/s ?

DGX agent

Hello guys, I'm curious about running DeepSeek-V4-Flash-0731 locally. Since it’s a Mixture of Experts (MoE) model with only 13B active parameters, I was hoping the VRAM requirements might be manageabl

model-releasesr-localllama
31 Jul 2026
← Previous
1…1011121314…31
Next →