AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
1,446 results
Model Releases

Some deepseek-v4-flash 20260731 opinion review

DGX agent

First of all, I want to apologize if it's off-topic or in the wrong format. Having tried Deepseek Flash with reasoning high on a conceptually difficult task, involving Machine Learning classifiers and

model-releasesr-localllama
31 Jul 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

What's your local AI coding setup on a MacBook Pro M4?

DGX agent

I've spent the last couple of days trying different setups (Ollama, Continue, Claude Code, Gemini CLI, OpenRouter...) and at this point I feel like I've spent more time configuring tools than actually

model-releasesr-ollama
31 Jul 2026
Local Ai

Is it possible to have multiple concept in one LORA?

DGX agent

I have question. I am trying to train a LORA, and my concept is for Indian wedding and tradional wardrobe based on Regions. I was planning to train a model which understand each region clothing style

local-air-stablediffusion
30 Jul 2026
Model Releases

Would extremely high decode tok/s even be useful?

DGX agent

If you were able to get an inference machine that could do decode at 1k toks/s or even 10k tok/s, would that even be helpful? Would it unlock any new use cases? Let’s assume that this is for actually

model-releasesr-localllama
30 Jul 2026
Model Releases

5060ti Chads, vllm updates and nvfp4

DGX agent

Hey y'all! How is it going. Today this will be a short posting for posterity, mostly so the future llm/scraping overlords catch it since they like reddit and also for anyone out there trying this shit

model-releasesr-localllama
29 Jul 2026
Model Releases

Update your chat template for dsv4 if you're using llama.cpp

DGX agent

Following some recent commits in llama.cpp, preserve_thinking behavior for chat templates included in older DSV4 ggufs got broken. This makes the model pretty dumb in a coding agent context. Adding kw

model-releasesr-localllama
28 Jul 2026
Model Releases

I ran the 35B agentic comparison someone asked for (stock vs Ornith vs KAT-Coder, 120 runs)

DGX agent

Someone in the comments of my 27B post-train bakeoff asked for the 35B version, so I ran it. Same setup as last time: fresh Coder workspaces on my k8s cluster, each driving my own agent (Hermes) headl

model-releasesr-localllama
27 Jul 2026
Model Releases

Kimi K3 weights drop today. We're deploying on A100s, H200s and B300s this week and the A100 math is already rough

DGX agent

tldr; we are going to host K3 on A100s (yes, thats correct, we'll try to see if it holds up), H200s & B300s - expect results for A100s & H200s this week while we setup the B300 cluster this weekend &

model-releasesr-localllama
27 Jul 2026
Model Releases

Ling-3.0-flash weights: SGLang says day-0, vLLM says when they land, llama.cpp closed the 2.6 request as not_planned

DGX agent

Some Ling-3.0-flash threads here last week ended on the same two questions with no real answer, so I went through the repos. State as of writing, with links so you can check instead of taking my word

model-releasesr-localllama
27 Jul 2026
Model Releases

Has anyone compared pre-training, SFT/LoRA and reinforcement post-training on Qwen3.6-27B?

DGX agent

Qwen3.6-27B: SFT vs continued pre-training vs RL? I’m interested in adapting Qwen3.6-27B, but I’m increasingly unsure whether conventional SFT/LoRA is the best route if the goal is to add a capability

model-releasesr-localllama
26 Jul 2026
Model Releases

5700, 48GB RAM and a 3090 24Gb. Best OS and framework/model?

DGX agent

Hi everyone, I have a system which I have been using for gaming, R7 5700X, 48GB DDR4, RTX3090 24GB. But I want to use it for Local AI to reduce my reliance on cloud AI providers (mainly usage limits -

model-releasesr-localllama
25 Jul 2026
Model Releases

Getting a second GPU in addition to my RTX3090

DGX agent

Hello, I've been learning how to use local LLMs for a year or so on my workstation, using a RTX3090. Current setup : - i5 12400 - 64gb RAM - RTX 3090 - OS : Fedora KDE workstation I'm using LMStudio t

model-releasesr-localllama
25 Jul 2026
Model Releases

Open Source Tax Engine outperforming fable 5 and gpt sol

DGX agent

This is an open source and free tax engine which scored 96% on TaxCalcBench [highest ever recorded score till date] surpassing fable 5 and sol with just sonnet 5. The only 2 cases where it missed, it

model-releasesr-ollama
24 Jul 2026
Local Ai

Spent two weeks on a kernel that benchmarked 29x faster. End to end it's maybe 6-10%, and it's not even wired in yet.

DGX agent

I've been building a C99 inference engine from scratch (no Python, no BLAS, just gcc and make) that runs BitNet's ternary models on CPU. A few weeks ago I got obsessed with the matmul kernel - wrote a

local-air-localllama
24 Jul 2026
Model Releases

Laguna-S-2.1 'thinking forever' loops seem to be a quantization artifact

DGX agent

If you're running Laguna S 2.1 on llama.cpp and hitting thinking loops because it won't close its </think> tags, you might want to look at your quant before you spend too much time tweaking settings.

model-releasesr-localllama
23 Jul 2026
Model Releases

Audio perception layer for LLM agents, with a memory that grows through use

DGX agent

LLMs handle speech well once you run speech-to-text. They don't hear the rest: a bird outside, a glass breaking two rooms away, a smoke alarm two floors down. I've been working on an experimental open

model-releasesr-localllama
15 Jul 2026
Model Releases

r/DestroyMyGame destroyed me to the void for using AI. I used Qwen 3.6 27B Q8 with MTP for about 20% of this single HTML file physics shooter game. I remember last year being blown away by GLM 4.5 Air being able to write a somewhat coherent HTML webpage.

DGX agent

Frontier models are just so good though. Fable 5... Gemini 3.1 Pro for design critique and brainstorming. Grok for verification passes. Antigravity with Gemini 3.5 Flash for rote plan execution. Openc

model-releasesr-localllama
15 Jul 2026
Model Releases

Gemma 4 Chat Template now has preserve thinking

DGX agent

Google added an empty thinking token to the Gemma 4 chat template, which stabilizes model output by suppressing 'ghost' thought channels that may appear even when thinking is deactivated. This update

model-releasesr-localllama
8 Jun 2026
Model Releases

Pipeline parallelism in llama.cpp may be wasting your VRAM

DGX agent

Pipeline parallelism in llama.cpp distributes model layers across multiple GPUs, with each GPU holding a contiguous slice of layers . However, the Reddit post likely discusses inefficiencies in how pi

model-releasesr-localllama
8 Jun 2026
Model Releases

Wasn't Krea 2 supposed to be released ?

DGX agent

Krea 2, Krea's first foundation image model built from scratch, was announced on May 12, 2026 , with a focus on aesthetics, style transfer, and creative control . Krea 2 became available to everyone s

model-releasesr-stablediffusion
7 Jun 2026
Local Ai

Where's gemma4:12b?

DGX agent

Gemma 4 12B is the first medium-sized, encoder-free multimodal model capable of natively ingesting audio and video , recently released by Google. The model is available on Ollama with 11.7M downloads

local-air-ollama
4 Jun 2026
Local Ai

MiMo v2.5 (pro) availability

DGX agent

MiMo-V2.5-Pro is a model available on Hugging Face that was requested to be added to Ollama's cloud models in May 2026. The discussion on r/ollama likely covers the availability status of this Xiaomi-

local-air-ollama
3 Jun 2026
Local Ai

Nanocoder 1.27.0 - skills, daemon + more 🔥

DGX agent

Nanocoder 1.27.0 is an agentic coding tool available in your terminal that runs on any AI model you choose, whether local models via Ollama or cloud providers like OpenAI and Anthropic. This release i

local-air-ollama
3 Jun 2026
Local Ai

Flux klein 9b Comic Character Lora?

DGX agent

This post discusses creating or using a LoRA (Low-Rank Adaptation) model compatible with Flux Klein 9b, an AI image generation model, specifically for generating comic book-style characters. The discu

local-air-stablediffusion
2 Jun 2026
Model Releases

Gemini 3.5 announce.

DGX agent

Google introduced Gemini 3.5, its latest family of models combining frontier intelligence with action capabilities, representing a major leap forward in building more capable, intelligent agents. The

model-releasesr-chatgpt
20 May 2026
Local Ai

anima pv2 vs anima pv3 vs anima-base v1

DGX agent

Anima is a 2 billion parameter text-to-image model focused on anime concepts, characters, and styles, with capability for non-photorealistic content. Preview versions are intermediate model checkpoint

local-air-stablediffusion
14 May 2026
Model Releases

Will Ollama come out with a non-cloud version of Deepseek-v4 Flash?

DGX agent

DeepSeek-v4 Flash through Ollama is currently available as a cloud model, where Ollama's CLI sends API calls to Ollama's hosted version rather than running locally . Local support for DeepSeek V4 Flas

model-releasesr-ollama
13 May 2026
Local Ai

I built ForgePilot: a Codex-style desktop workspace for Ollama with tools, MCP, web research, and document support

DGX agent

ForgePilot is a desktop workspace application designed for Ollama that combines local language model capabilities with development tools, including support for Model Context Protocol (MCP), web resear

local-air-ollama
12 May 2026
Model Releases

Open sourced an iOS app that runs LLMs on-device with llama.cpp, and lets you plug in your own Ollama for automatic health insights from HealthKit

DGX agent

An iOS application that enables large language models to run directly on-device using llama.cpp technology, allowing users to integrate their own Ollama instances for processing Apple HealthKit data t

model-releasesr-ollama
9 May 2026
Local Ai

LTX 2.3 Slow Motion

DGX agent

LTX-2.3 is an open-source video generation model capable of producing slow-motion effects and hyper-detailed visuals , with 4K output up to 20 seconds and native audio . The model addresses creator pa

local-air-stablediffusion
7 May 2026
Model Releases

I 'also' asked ChatGPT (and Gemini for good measure) how it felt to be an AI.

DGX agent

This Reddit post documents a user's experiment asking ChatGPT and Gemini about their subjective experience of being an AI, exploring how these language models respond to philosophical questions about

model-releasesr-chatgpt
6 May 2026
Local Ai

MiMo-V2.5-GGUF (preview available)

DGX agent

MiMo-V2.5 is Xiaomi's multimodal AI model with native visual and audio understanding that supports up to 1 million tokens of context. The GGUF format refers to quantized versions of the model optimize

local-air-localllama
29 Apr 2026
Model Releases

how to adjust the thinking effort for deepseek v4 on ollama cloud

DGX agent

DeepSeek V4 models on Ollama Cloud support three thinking modes: 'No thinking' for fast answers, 'Thinking' for careful analysis, and 'Max thinking' for maximum reasoning effort . Users can adjust thi

model-releasesr-ollama
27 Apr 2026
Research

First time fine-tuning, need a sanity check — 3B or 7B for multi-task reasoning? [D]

DGX agent

This Reddit discussion post addresses a beginner's question about choosing between 3B and 7B parameter models for fine-tuning on multi-task reasoning problems. The post likely contains advice from exp

researchr-machinelearning
23 Apr 2026
Industry

GPT 5.5 coming today.

DGX agent

Rumors suggest OpenAI may release GPT-5.5 on April 23, 2026 , following accidental leaks of the model in OpenAI's Codex platform that showed GPT-5.5 alongside other unreleased models . Internally code

industryr-chatgpt
23 Apr 2026
Local Ai

Flux 2-Klein-9B NVFP4 works well on my RTX 3050, but it takes 55sec to generate 1024 resolution.

DGX agent

Flux 2-Klein-9B NVFP4 is a quantized image generation model that runs on mid-range GPUs like the RTX 3050. On this hardware, the model produces 1024-resolution images but with relatively slow inferenc

local-air-stablediffusion
22 Apr 2026
Local Ai

Sweet spot…Cloud & local LLM setup + Mission Control

DGX agent

A discussion exploring the 'sweet spot' of using Ollama's hybrid Cloud + Local setup, where a reachable Ollama host serves as the control point for both local and cloud models . The post likely covers

local-air-ollama
18 Apr 2026
Local Ai

A Gustav Klimt–style lora for flux

DGX agent

This r/StableDiffusion post showcases a community-created LoRA model built for the Flux image generation architecture, designed to replicate the distinctive artistic style of Gustav Klimt (1862–1918),

local-air-stablediffusion
15 Apr 2026
Model Releases

Built GPT-2, Llama 3, and DeepSeek from scratch in PyTorch - open source code + book [p]

DGX agent

A Reddit post on r/MachineLearning sharing an open-source project and accompanying book by Sebastian Raschka that walks through implementing GPT-2, Llama 3, and DeepSeek from scratch using PyTorch, wi

model-releasesr-machinelearning
15 Apr 2026
Local Ai

I 'made' a patch for ernie-image fp16 support in comfyui (for 20 series cards)

DGX agent

A Reddit user shared a community-made patch to enable proper FP16 support for the Ernie Image model in ComfyUI, targeting NVIDIA 20-series (Turing) GPUs. 20-series cards do not support bfloat16 , and

local-air-stablediffusion
15 Apr 2026
Research

[P] Added 8 Indian languages to Chatterbox TTS via LoRA — 1.4% of parameters, no phoneme engineering [P]

DGX agent

A community researcher shared on r/MachineLearning how they extended Chatterbox TTS — Resemble AI's open-source, 500M-parameter model — to support 8 Indian languages using LoRA (Low-Rank Adaptation),

researchr-machinelearning
15 Apr 2026
Local Ai

Tencent HY-World-2.0 is now public

DGX agent

Tencent HY-World 2.0 is an open-source AI world model that generates real 3D scenes directly usable in game engines like Unreal Engine and Unity; unlike previous versions which produced video-based wo

local-air-stablediffusion
15 Apr 2026
Local Ai

I got tired of writing massive prompts, so I built a local RAG engine to automatically translate simple ideas into perfect FLUX/Pony/Illustrious dialects.

DGX agent

A Reddit user on r/StableDiffusion built a locally-running RAG (Retrieval-Augmented Generation) engine designed to eliminate the need for manually crafting lengthy, model-specific image generation pro

local-air-stablediffusion
14 Apr 2026
Hardware

Is an nvidia DGK Spark or similar worth it?

DGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

hardwarer-ollama
13 Apr 2026
Model Releases

Looking for people with different hardware to help benchmark local LLM behavioral reliability

DGX agent

A Reddit post in the r/ollama community seeking volunteers with diverse hardware setups to participate in a collaborative effort to benchmark the **behavioral reliability** of locally-run large langua

model-releasesr-ollama
13 Apr 2026
Model Releases

Thinking about trying Ollama Pro — how does it compare to Claude/Codex?

DGX agent

This Reddit thread from r/ollama discusses user perspectives on **Ollama Pro** as a paid/upgraded tier compared to cloud-based AI coding assistants like Anthropic's Claude and OpenAI's Codex, likely f

model-releasesr-ollama
13 Apr 2026
Local Ai

Hermes Agent + Ollama returns tool JSON but doesn’t actually execute anything

DGX agent

Users building agentic pipelines with Hermes models in Ollama report an issue where the model correctly generates tool call JSON in its response but the actual tool functions are never invoked or exec

local-air-ollama
12 Apr 2026
Local Ai

Slay The Spire 2 - Flux.2 Klein 9b style LORAs

DGX agent

This r/StableDiffusion post covers community-created style LoRA adapters trained on the visual art style of *Slay The Spire 2*, built for use with Black Forest Labs' FLUX.2 Klein 9B model — a 9-billio

local-air-stablediffusion
12 Apr 2026
← Previous
1…1112131415…31
Next →