AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
1,433 results
Model Releases

New wave of miniboss models you can run on dual DGX Spark

DGX agent

Two DGX Spark and a Connect-X7 cable give you about 250GB of usable memory for 7000 8000 USD. This allows using some interesting models at 4-bit. For what seemed like an eternity, the only serious mod

model-releasesr-localllama
15 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

InvokeAI 6.13 just released, its largest community-driven release ever. Adds full support for Anima & Qwen Image, support for API models (like GPT Image), support for Prompt Expansion & Image To Prompt, lasso & polygon tools, overhauled docs website and more

DGX agent

InvokeAI 6.13 is the largest community-driven release of the software, adding full support for Anima & Qwen Image models, API model integration (such as GPT Image), and new features including Prompt E

model-releasesr-stablediffusion
27 May 2026
Model Releases

TabPFN-3 just released: a pre-trained tabular foundation model for up to 1M rows [R][N]

DGX agent

TabPFN-3 is a pre-trained tabular foundation model that supports datasets up to 1,000,000 rows × 200 features , representing a significant scaling improvement for the TabPFN family. The model delivers

model-releasesr-machinelearning
12 May 2026
Local Ai

Best Ollama model for n8n workflows (RAG, file handling, reasoning) + hardware requirements?

DGX agent

Models like Qwen3 and Llama 3.2 are commonly used for n8n RAG workflows , with selection depending on use case requirements. Running Ollama models locally requires at least 16 GB of RAM on your device

local-air-ollama
17 Apr 2026
Model Releases

Google’s DeepMind just released new 4B and 27B MedGemma models!

DGX agent

Google DeepMind released MedGemma, a collection of medical vision-language foundation models based on Gemma 3 in 4B and 27B parameter sizes, demonstrating advanced medical understanding and reasoning

model-releasesr-ollama
16 Apr 2026
Local Ai

Just installed ForgeNeo and I'm facing this issue *failed to recognize model type*

DGX agent

The `'Failed to recognize model type!'` error in ForgeNeo (stable-diffusion-webui-forge) is a `ValueError` raised by the backend loader (`backend/loader.py`) when the application cannot identify th...

local-air-stablediffusion
11 Apr 2026
Model Releases

struggling choosing one edit model from klein 9b or qwen 2511.

DGX agent

This r/StableDiffusion thread discusses the community debate around choosing between FLUX.2 [klein] 9B and Qwen Image Edit 2511 as an image editing model, two strong open-source contenders in the spac

model-releasesr-stablediffusion
11 Apr 2026
Local Ai

What is the 'Unload Models and Execution Cache' from the ComfyUI menu doing that all the other model and cache-clearing nodes I've tried don't do?

DGX agent

I was unable to retrieve the content of the specified Reddit URL directly, as my web search tool does not fetch raw URLs or Reddit threads directly, and I've exhausted my search attempts for this t...

local-air-stablediffusion
10 Apr 2026
Local Ai

Using Ollama Gemma4 models via OpenWebUI on my phone and it’s been a good experience

DGX agent

Users are running Google's Gemma 4 models locally via Ollama and accessing them on their phones through Open WebUI, reporting a positive experience. Since Ollama doesn't run natively on iOS or And...

local-air-ollama
9 Apr 2026
Model Releases

Anyone Using (Koreas) 'Solar Open 2' (250B, 15B) Model?

DGX agent

I just heard of this model. Seems to be a competitor to DeepSeek V4 Flash. About the same size and active parameters. Anyone tested it compared to V4 Flash? Link: https://huggingface.co/upstage/Solar-

model-releasesr-localllama
11 Aug 2026
Local Ai

Anyone already used a model imported directly in the ollama cloud

DGX agent

Ollama allons you to import model but have you ever tried doing so ? Like running model imported from hugging face or you own model ? Any use case you wanna share ? Very curious about that submitted b

local-air-ollama
9 Aug 2026
Local Ai

I Turned My Underused Gaming Laptop Into a Local AI Workstation

DGX agent

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

local-air-ollama
9 Aug 2026
Model Releases

Scotoma-2: Gemma4, but with less annoying slop and better writing.

DGX agent

GGUFs here: https://huggingface.co/ReadyArt/gemma-4-31B-it-scotoma-2-GGUF Disclaimer: By slop, we are specifically talking about specific tics with the model(sentence structures), but this doesn't inc

model-releasesr-localllama
6 Aug 2026
Local Ai

For MoE models the arithmetic splits in two: capacity follows total params, speed follows active

DGX agent

A few people asked for this after the bandwidth thread, so here it is on its own instead of buried in a comment. The dense rule was simple: every token reads every weight, so tokens/sec ≈ bandwidth ÷

local-air-ollama
1 Aug 2026
Local Ai

What I learned using Ollama on a real Paperless archive: model choice was not the main problem

DGX agent

I maintain Tagvico, an open-source companion for Paperless-ngx. I added Ollama because document text is exactly the kind of data many people do not want to send to a hosted model. The surprising failu

local-air-ollama
24 Jul 2026
Local Ai

why does ollama unloads models automatically?

DGX agent

Ollama automatically unloads inactive models from memory based on inactivity parameters, with models being removed after a specified duration of non-use . By default, Ollama unloads a model after 5 mi

local-air-ollama
9 Jun 2026
Hardware

Nvidia releasesCosmos3-Super-Text2Image model . 64 billion paramteres

DGX agent

Cosmos3-Super-Text2Image is a 64 billion parameter omnimodal world model capable of generating high-quality images from text inputs as part of NVIDIA's Cosmos 3 foundation model platform. The model is

hardwarer-stablediffusion
1 Jun 2026
Local Ai

Are there any good story writer models that I can ruj with a 5080 16gb?

DGX agent

This Reddit post from r/ollama asks about story-writing language models that can run on a 5080 GPU with 16GB of VRAM . The discussion likely covers recommended open-source or quantized models suitable

local-air-ollama
23 Apr 2026
Local Ai

How to Disable Thinking mode of Ollama Models Using Copilot CLI?

DGX agent

A guide on disabling thinking mode in Ollama models using the CLI by running models with the `--think=false` flag or using `/set nothink` followed by a prompt. The post likely discusses how to configu

local-air-ollama
17 Apr 2026
Local Ai

What does 'Run <number> cloud models at a time' in Ollama Cloud Subscription mean?

DGX agent

The Ollama Cloud Subscription includes a feature described as 'Run cloud models at a time,' which refers to how many AI models a user can have simultaneously loaded and running in the cloud at any giv

local-air-ollama
15 Apr 2026
Research

You can decompose models into a graph database [N]

DGX agent

This Reddit post from r/MachineLearning discusses the concept of decomposing machine learning models into a graph database representation, treating a model's components — such as layers, weights, and

researchr-machinelearning
14 Apr 2026
Local Ai

I’m looking for advice on setting up a local AI model that can generate Word reports automatically.

DGX agent

This r/ollama thread discusses community advice on configuring a locally-run AI model (via Ollama) to automatically generate Word documents or reports, covering topics such as model selection, scripti

local-air-ollama
13 Apr 2026
Local Ai

OstrisAI-Toolkit Lora --> Anima model.

DGX agent

This Reddit post discusses using the Ostris AI-Toolkit to train a LoRA (Low-Rank Adaptation) model targeting the 'Anima' Stable Diffusion model. The Ostris AI Toolkit is a powerful framework that allo

local-air-stablediffusion
12 Apr 2026
Local Ai

Models randomly becoming corrupted?

DGX agent

I was unable to retrieve the specific Reddit thread at the provided URL, and the search results did not return content from that exact post. The search results returned related but distinct issues ...

local-air-stablediffusion
10 Apr 2026
Local Ai

Recommended Model for a 4060ti 8gb and 16gb ram

DGX agent

For users running Ollama on an NVIDIA RTX 4060 Ti with 8GB VRAM and 16GB system RAM, the community consensus recommends 7B–8B parameter models (such as Llama 3.1 8B, Mistral 7B, or Qwen 8B) using Q...

local-air-ollama
10 Apr 2026
Model Releases

swiss-ai/Apertus-v1.5 70B/8B

DGX agent

https://huggingface.co/swiss-ai/Apertus-v1.5-70B https://huggingface.co/swiss-ai/Apertus-v1.5-8B Apertus 1.5 is a family of 8B and 70B parameter language models designed to advance the state of multil

model-releasesr-localllama
24 Jul 2026
Model Releases

[Paper] SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

DGX agent

Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, including severe memory pressure, non-overlappe

model-releasesr-localllama
23 Jul 2026
Model Releases

Gemma 4:e4b offloads to RAM despite having just half of VRAM used.

DGX agent

Users on the r/ollama subreddit reported that the **Gemma 4 E4B** model in Ollama offloads layers to RAM even when GPU VRAM is only partially utilized. This behavior is linked to how Ollama and lla...

model-releasesr-ollama
10 Apr 2026
Model Releases

Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)

DGX agent

Needs a lot of requests compared to Qwen (almost twice) and Gemma (almost x3). Final score is fine, even though it is 'not a coding model' https://wonderrico.github.io/local_llm_benchmark/benchmark-ma

model-releasesr-localllama
11 Aug 2026
Local Ai

i just spent weeks rewriting my webUI from scratch, getting rid of all AI slop within the codebase and switching it over to a proper lightweight framework (alpine.js). i am now comfortable suggesting it as an alternative to openwebUI, librechat and the like! it is made for local models

DGX agent

[Fully open source under GPL3, made from the ground up for use with local models, no subscriptions, no corporate backing] When i first started this, it was meant to be a fully lightweight, extremely m

local-air-localllama
6 Aug 2026
Model Releases

I benchmarked the 4 models I had pulled. The 1.1GB one beat the 2GB one at math and lost badly at extraction.

DGX agent

152 generations, deterministic grading (exact number/string/JSON/regex), no LLM judge on my 16GB laptop. task type | deepseek-r1:1.5b (1.1GB) | llama3.2:3b (2.0GB) | gemma:2b | codellama 7b arithmetic

model-releasesr-ollama
4 Aug 2026
Model Releases

Ling-3.0-flash is another potential model to test before qwen3.8 27b

DGX agent

I tested Ling-3.0-flash with hard bugs and it fixed bugs that qwen3.6-27b could not. This models speed faster than deepseek v4 flash but almost the same level as (old) deepseek v4 flash. Note: hard bu

model-releasesr-localllama
3 Aug 2026
Model Releases

I have trained a model to predict my blood sugar [P]

DGX agent

It's an encoder-only transformer that consumes past(blood glucose + carbs + insulin) and future(carbs + insulin) and predicts future blood glucose for the next 2 hours. Announced meals and boluses/bas

model-releasesr-machinelearning
31 Jul 2026
Model Releases

We've gotten some great medium sized models lately (DSV4 Flash 0731, Inkling Small, Laguna S 2.1, Step 3.7 Flash) but does anybody else want to see some new 70-80b contenders?

DGX agent

I can run the mediums, but sometimes I want a faster option that's smarter than Qwen 27B/35B. On my hardware I get like 500 to 800 tok/s prefill and 16 to 22 tok/s gen on ~120B class models, which is

model-releasesr-localllama
31 Jul 2026
Model Releases

Current smallest usable coding model

DGX agent

I've been seeing a lot of news about the latest gemma 4 and qwen 3.6 being really good and the current go-to models but those are out of reach for my GPU at the moment. With 4GB VRAM and 40 GB RAM, I

model-releasesr-localllama
27 Jul 2026
Model Releases

Simple desktop GUI for multiple local TTS models (Tkinter)

DGX agent

https://preview.redd.it/qoogv5m1skfh1.png?width=688&format=png&auto=webp&s=3f06fb28ceec3cd3ab95bcebd0c71374332aac9b Built a simple desktop GUI (Tkinter) that supports multiple local TTS engines (curre

model-releasesr-localllama
26 Jul 2026
Model Releases

4x 3090, 96gb vram what Model to drive Hermes?

DGX agent

3 year lurker, now i finally got my server up and running. dont know which model to choose. llama.cpp or vllm, what makes more sense? mainly single user with maybe 2-3 more additional users in family,

model-releasesr-localllama
25 Jul 2026
Model Releases

🇦🇹 Austria is rolling out a government AI-platform using Mistral models and Open WebUI

DGX agent

This is a surprisingly large real-world deployment: 'GovGPT' is part of Austria’s Public AI initiative, running on sovereign infrastructure (in their BRZ - federal datacenter) with Mistral open-weight

model-releasesr-localllama
22 Jul 2026
Model Releases

I released a softmax-free attention model at GPT-2 Medium scale (~354M params, 11.5B tokens): structural sparsity + tile-skipping kernels for long-context VRAM savings. Open weights + custom Triton kernels [R]

DGX agent

A researcher released an open-source softmax-free attention model at GPT-2 Medium scale (354M parameters trained on 11.5B tokens) that uses structural sparsity and tile-skipping kernels to reduce VRAM

model-releasesr-machinelearning
21 Jun 2026
Local Ai

What are the most capable LLM models I can run on my laptop?

DGX agent

A discussion on r/ollama exploring which high-performance LLM models can be effectively run locally on standard laptop hardware , likely covering model size comparisons, hardware requirements, and per

local-air-ollama
5 Jun 2026
Model Releases

A tool to get Claude Code-style reliability from fully local models

DGX agent

Ollama exposes an Anthropic-compatible Messages endpoint , allowing developers to run powerful open-source AI models locally with no API costs and pair them with Claude Code for a capable local AI cod

model-releasesr-ollama
26 May 2026
Research

Call for Papers - Workshop on Unlearning and Model Editing U&ME at ECCV 2026 [R]

DGX agent

This call for papers announces a workshop focused on the growing need for efficient and effective techniques for editing trained models, especially large generative models. The workshop solicits paper

researchr-machinelearning
25 May 2026
Model Releases

Detailed review and guide from my testing of local ollama setup with DeepSeek models (Ryzen APU's only)

DGX agent

This post provides a detailed review and practical guide for setting up and testing Ollama with DeepSeek models specifically on Ryzen APU systems. It likely covers performance benchmarks, configuratio

model-releasesr-ollama
10 May 2026
Model Releases

Nemotron-3-Nano-Omni-30B-A3B-Reasoning, New model?

DGX agent

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding for enterprise Q&A, summarization, transcription, and document intelligence, w

model-releasesr-localllama
28 Apr 2026
Model Releases

Anthropic locked Claude Code to native apps in Jan 2026. Are we still comparing models or just ecosystems

DGX agent

Anthropic's Claude Code desktop app is strictly optimized for Anthropic's models , creating a 'walled garden' effect that restricts users to Claude exclusively. The redesigned Claude Code desktop app

model-releasesr-chatgpt
19 Apr 2026
Local Ai

Running a 31B model locally made me realize how insane LLM infra actually is

DGX agent

A Reddit post from r/ollama in which a user shares their experience running a 31B parameter model locally using Ollama, reflecting on the surprisingly demanding hardware and infrastructure requirement

local-air-ollama
15 Apr 2026
Local Ai

Which cloud model do you use for coding? Which one got better reasoning?

DGX agent

This r/ollama thread is a community discussion where users share their preferred cloud-based AI models for coding tasks and compare their reasoning capabilities. The conversation likely highlights pop

local-air-ollama
15 Apr 2026
Local Ai

Tried Ollama Cloud, just realize only Kimi model accept images

DGX agent

A Reddit user exploring Ollama Cloud noted that, at the time of their post, only the Kimi model supported image (vision/multimodal) inputs among the available cloud models. Kimi K2.5 is a native multi

local-air-ollama
13 Apr 2026
← Previous
12345…30
Next →