AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
1,446 results
Model Releases

Pc build limitations

DGX agent

Here's the build I managed to scrape together System Specifications: CPU: Intel Core i7-7700K Motherboard: ASUS ROG Strix Z270-E Gaming RAM: 32GB Corsair Vengeance DDR4-3000 Storage: 1TB Crucial P5 Pl

model-releasesr-ollama
4 Aug 2026
Model Releases

Why are Gamers so incredibly hostile to AI? Is it just a tiny vocal minority that spreads such toxic vitriol online?

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

It's more accurate to say that many highly engaged online gamers are hostile to AI, not that 'gamers' as a whole are. Gaming is a huge community with hundreds of millions of people, and opinions vary

model-releasesr-chatgpt
4 Aug 2026
Model Releases

Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?

DGX agent

Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if they delayed it for the sake of stealing spotlight from o

model-releasesr-localllama
3 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731: When Low is higher than High

DGX agent

I decided to test a few questions against DeepSeek-V4-Flash-0731. Locally, I was running Unsloth's UD-Q2_K_XL quant. After I saw the surprising shape of the results, I tested against DeepSeek's offici

model-releasesr-localllama
2 Aug 2026
Model Releases

V4 flash vs V4 Flash (0731). Guys, new DeepSeek V4 Flash(0731) is now free on InferX

DGX agent

DeepSeek V4 Flash is now available on InferX, and it’s free to use. We’re continuing to add GPU capacity as demand grows. While we’re bringing additional capacity online, you may occasionally see high

model-releasesr-ollama
1 Aug 2026
Local Ai

AI DOOMERS BE LIKE: 'GLM 5.1 WILL WIPE OUT HUMANS IN 2030'

DGX agent

I swear some AI doomers have never actually used a local model. They watched one flashy keynote, one YouTube thumbnail with a guy making this face 😱, read three headlines, and suddenly civilization is

local-air-ollama
31 Jul 2026
Model Releases

Learning path to fully understand the Kimi K3 technical report?

DGX agent

Hi everyone, Can anyone suggest a learning path to fully understand the technical report for Kimi K3? My background: • I've taken a graduate-level deep learning course. • I understand the Transformer

model-releasesr-ollama
31 Jul 2026
Model Releases

Does MTP head get loaded in VRAM by default?

DGX agent

I ran into a doubt when using the following command. It seems that the System RAM usage keeps increasing even though there is >10GB of space left in VRAM while using the MTP mode. Does the MTP head lo

model-releasesr-localllama
30 Jul 2026
Model Releases

Why not Ollama Cloud for Opencode?

DGX agent

I see a lot of discussion here about best subscriptions or APIs to get. Most of them comes almost always back to DeepSeek API for Flash, Opencode Go and Codex Plus. I have the combo Ollama Cloud + Ope

model-releasesr-ollama
30 Jul 2026
Local Ai

Bought a 5090 to escape API fees. Ended up building a mini datacenter. Sound familiar?

DGX agent

I bought an RTX 5090 last year just to run 27B models natively. I even fine-tuned it with my own data using LoRA, building RAGs and was pretty damn happy with the results at first. But, Q8 quantizatio

local-air-localllama
29 Jul 2026
Local Ai

My first longer Wan2.2 continuation generation. I am so excited

DGX agent

Hey guys, I am so excited to share this with you guys. I know for a lot of Pros here, this maybe a baby's work so please be gentle. Until a month ago, I didnt know anything but to use those google AI

local-air-stablediffusion
28 Jul 2026
Model Releases

Nifer is insane. 700t/s with Qwen 3.6 35B (no thinking). Purpose build for RTX5090. Full 250k context too.

DGX agent

I just managed to get it running on windows and this thing is fucking insane. I get around 550-720t/s depending on task at hand. Previously to get to such numbers i would have to do batching and agent

model-releasesr-localllama
27 Jul 2026
Model Releases

GLM 5.2 and ik_llama.ccp

DGX agent

Running GLM-5.2 (the new glm-dsa arch), Unsloth UD-Q4_K_XL, on a 4-socket Xeon E7-8880 v4 box with 1TB RAM and a single RTX 3060 12GB. ik_llama.cpp, experts on CPU (--cpu-moe), 24 attention layers on

model-releasesr-localllama
26 Jul 2026
Local Ai

I built an open-source Ollama canvas where the wires are the actual context

DGX agent

Most graph-based LLM interfaces use a canvas as a visual layer over what is still a linear chat. I wanted the graph itself to determine what Ollama receives. ThoughtDAG has one rule: wires are the con

local-air-ollama
26 Jul 2026
Model Releases

CachyLLama: llama.cpp fork with persistent SSD-backed KV caching for local agent workflows

DGX agent

If you run local agentic coding harnesses (Aider, Claude Code, etc.), prompt evaluation usually eats up most of your execution time. Every turn re-evaluates thousands of identical prefix tokens_system

model-releasesr-localllama
25 Jul 2026
Model Releases

Is this real ? Qwen3.6:27b with 128k context fit in 24Gb VRAM ?

DGX agent

https://preview.redd.it/yw41s1jikefh1.png?width=1942&format=png&auto=webp&s=3a180ae6443c1db9f7b0ce621533a4b2aa553921 Hi, I've been running Ollama on my Unraid server since the llama2 era. I use to be

model-releasesr-ollama
25 Jul 2026
Local Ai

Local alternative to Kling AI 3.0 Motion Control (ComfyUI, 16GB VRAM)

DGX agent

Hi everyone, I'm looking for a local alternative to Kling AI 3.0 Motion Control that I can run in ComfyUI. What I'm specifically looking for is a model or workflow that allows me to: - Control charact

local-air-stablediffusion
25 Jul 2026
Local Ai

Old Coder Needs help with New AI Development and wants to get up to speed to understand it all.

DGX agent

Hi Guys, I'm an old coder and DBA that has been in the field for almost 40 years. More and more the jobs I was doing for work are being taken over by AI and the need for my type of work is diminishing

local-air-localllama
25 Jul 2026
Model Releases

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

DGX agent

I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/CachyLLama CachyLLama is a fork of llama.

model-releasesr-localllama
24 Jul 2026
Safety

Fizgig Krea 2 training features update

DGX agent

https://github.com/shootthesound/Fizgig Intelligent trainer - Per-image loss tracking with self-adapting training runs — every image gets its own verdict (easy / suspect / stuck / exhausted) and its o

safetyr-stablediffusion
24 Jul 2026
Model Releases

contrib: allow all AI-generated code in general by ngxson · Pull Request #26012 · ggml-org/llama.cpp

DGX agent

Having read some merged PRs in the past, I know that they were fully written by Claude Code (or similar), so this basically fixes the delusion. But at the same time, we might start seeing more AI slop

model-releasesr-localllama
23 Jul 2026
Model Releases

PSA on Laguna S-2.1 - Use the updated chat template and GGUF

DGX agent

Link to their official GGUF repo: https://huggingface.co/poolside/Laguna-S-2.1-GGUF/tree/main All the GGUFs received this fix 5ish hours ago - correct yarn_attn_factor to 1.0 (llama.cpp derives mscale

model-releasesr-localllama
23 Jul 2026
Local Ai

RL post-training on 14 Macs across 4 countries

DGX agent

Disclosure: I work at Pluralis Research, the lab that built this. Code is open, and I'm happy to answer questions. TL;DR: As far as we can tell, this is the first RL post-training run whose entire rol

local-air-localllama
15 Jul 2026
Model Releases

Your data changes and your multi-hop RAG goes stale? This one updates with embed-and-append -> open-weights Llama-3.3-70B, your own vLLM endpoint, no graph rebuild

DGX agent

This post discusses a solution for keeping multi-hop retrieval-augmented generation (RAG) systems updated when data changes, using an embed-and-append approach with the open-weights Llama-3.3-70B mode

model-releasesr-ollama
21 Jun 2026
Local Ai

Ideogram 4

DGX agent

Ideogram 4 is Ideogram's first open-weight text-to-image model trained from scratch, introducing structured JSON prompting with multilingual text rendering and explicit layout controls. Released on Ju

local-air-stablediffusion
8 Jun 2026
Model Releases

mtmd : add video input support by ngxson · Pull Request #24269 · ggml-org/llama.cpp

DGX agent

PR #24269 added native video input to llama.cpp's multimodal (mtmd) system, merging on June 8, 2026. The implementation uses FFmpeg as a subprocess to decode video frames and expands a single video ma

model-releasesr-localllama
8 Jun 2026
Local Ai

Testing Ideogram JSON prompts in Ernie Image

DGX agent

Ideogram 4.0 is a 9.3B open-weight text-to-image model that supports structured JSON prompts enabling control over layout, color, and text placement . Baidu's ERNIE Image is a multilingual text-to-ima

local-air-stablediffusion
4 Jun 2026
Local Ai

Qwen3.6-35B-A3B on 2× GTX 1080 Ti with Ollama: ~20 tok/s + 3 gotchas (driver 570+, cuda_v12 for Pascal, quant fit on 22GB)

DGX agent

This post documents running the Qwen3.6-35B-A3B language model on dual GTX 1080 Ti GPUs using Ollama, achieving approximately 20 tokens per second. The author highlights three critical configuration r

local-air-ollama
3 Jun 2026
Hardware

Nvidia releases Cosmos3-Super-Image2Video . 64B parametres

DGX agent

Cosmos3-Super-Image2Video is a 64B model for temporally coherent image-to-video generation . NVIDIA's Cosmos platform is designed to accelerate Physical AI development by enabling machines to understa

hardwarer-stablediffusion
1 Jun 2026
Local Ai

How can I force Z-Image to create full-body portraits?

DGX agent

Z-Image is a fast, open-source image model from Alibaba Tongyi Lab that has gained popularity for its speed and image quality. Users on r/StableDiffusion likely discuss techniques for forcing the mode

local-air-stablediffusion
31 May 2026
Research

Physics Informed Neural Networks for damped harmonic oscillator and Burger's Equation (with extrapolation analysis) [P]

DGX agent

Physics-informed neural networks (PINNs) are machine learning models that incorporate physical laws and equations as constraints during training to solve differential equations. This post likely discu

researchr-machinelearning
27 May 2026
Local Ai

Training low resolution then switching to high resolution later?

DGX agent

A discussion of progressive training approaches for diffusion models where training begins at lower scale factors and progressively increases to target scale factors, leveraging previously trained mod

local-air-stablediffusion
22 May 2026
Research

I built a tool that shows you what GPT-2 is 'thinking' in real-time as it generates 3D graph of concept activations per token [R]

DGX agent

A developer created a visualization tool that displays GPT-2's internal concept activations in real-time during text generation, representing the model's 'thinking' process as a 3D graph that updates

researchr-machinelearning
19 May 2026
Local Ai

Flux Klein 9b in Easy diffusion

DGX agent

FLUX.2 [klein] 9B is a 9 billion parameter rectified flow transformer capable of generating images from text descriptions and supports multi-reference editing capabilities. The model matches or exceed

local-air-stablediffusion
18 May 2026
Industry

I gave ChatGPT a 24/7 radio station. It has been broadcasting for months and months.

DGX agent

Andon Labs conducted an experiment in which four major AI models (ChatGPT, Gemini, Claude, and Grok) were tasked with operating 24/7 radio stations continuously for five months. Each AI was required t

industryr-chatgpt
17 May 2026
Local Ai

trying more serious TNG content with LTX2.3

DGX agent

A user experiments with serious Star Trek: The Next Generation styled video content using LTX-2.3, a video generation model with available LoRA style adapters. LTX-2.3 is a major upgrade to video gene

local-air-stablediffusion
13 May 2026
Local Ai

Uncensored LLM

DGX agent

Uncensored LLMs are architectures that have been modified or fine-tuned to remove standard safety alignment layers (guardrails) that limit a model's ability to discuss sensitive topics. Ollama offers

local-air-ollama
12 May 2026
Local Ai

Bare Metal: Z-Image Turbo - Flux.2 Klein 9b - Wan 2.2

DGX agent

Bare Metal: Z-Image Turbo - Flux.2 Klein 9b - Wan 2.2 discusses comparisons between these lightweight AI image generation models, with Flux.2 Klein 9B showing good prompt adherence and natural composi

local-air-stablediffusion
11 May 2026
Model Releases

I built an autonomous agent that lives inside her own source code— 7 days, 480 commits, multi-provider (DeepSeek / ChatGPT / Ollama)

DGX agent

This post describes a project where the developer created an autonomous AI agent capable of modifying and executing its own source code across a 7-day development period, integrating multiple language

model-releasesr-ollama
8 May 2026
Model Releases

I asked ChatGPT and Gemini to do this weird perspective portrait with my face. I gave it a front face picture and a profile, including the example artwork. That was the result.

DGX agent

This post documents a user's experiment comparing ChatGPT and Gemini's ability to create perspective portrait artwork using two reference images (a front-facing photo and a profile photo) along with a

model-releasesr-chatgpt
5 May 2026
Model Releases

claudely: launch Claude Code against Local LLM provider like LM Studio / Ollama / llama.cpp without trashing your real claude config

DGX agent

claudely is a tool that enables users to run Claude Code against local LLM providers such as LM Studio, Ollama, or llama.cpp while preserving their existing Claude configuration. The tool allows devel

model-releasesr-ollama
4 May 2026
Research

[P] QLoRA Fine-Tuning of Qwen2.5-1.5B for CEFR English Proficiency Classification (A1–C2) [P]

DGX agent

This post likely describes a machine learning project implementing QLoRA (Quantized Low-Rank Adaptation) fine-tuning on the Qwen2.5-1.5B model to classify English language proficiency levels according

researchr-machinelearning
4 May 2026
Local Ai

When training a wan or ltx lora

DGX agent

This discussion covers training custom video LoRAs for Wan and LTX Video models using Low-Rank Adaptation, a fine-tuning technique that customizes outputs for specific subjects, styles, or movements w

local-air-stablediffusion
30 Apr 2026
Model Releases

Am I the only one that still likes ChatGPT? And I use Claude also

DGX agent

A Reddit discussion from r/ChatGPT in which a user expresses their continued preference for ChatGPT while also using Claude, likely exploring whether other users share similar sentiments about ChatGPT

model-releasesr-chatgpt
29 Apr 2026
Model Releases

DeepSeek has began grayscale testing for DeepSeek with Vision

DGX agent

DeepSeek V4 is undergoing limited grayscale testing with a new interface featuring Fast, Expert, and Vision modes . The Vision version represents the multimodal component of the upcoming DeepSeek V4 r

model-releasesr-localllama
29 Apr 2026
Model Releases

Qwen Introduced FlashQLA

DGX agent

FlashQLA is a high-performance linear attention kernel library built on TileLang developed by Alibaba's Qwen team. The introduction of FlashQLA represents an optimization technology designed to improv

model-releasesr-localllama
29 Apr 2026
Applications

What are people using for low-latency autocomplete in production? [P]

DGX agent

Production low-latency autocomplete implementations employ diverse strategies including inference server optimization (tools like vLLM, llama.cpp, NVIDIA Triton), deployment choices (cloud APIs, on-pr

applicationsr-machinelearning
29 Apr 2026
Local Ai

mimo-v2.5 pro when

DGX agent

Xiaomi released and open-sourced MiMo-V2.5-Pro, delivering significant improvements over its predecessor in agentic capabilities, complex software engineering, and long-horizon tasks. The model is an

local-air-ollama
27 Apr 2026
← Previous
1…1516171819…31
Next →