AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
1,433 results
Model Releases

Observations on Muse-Glimmer reasoning traces being noticeably different from qwen / gemma models and questions for you guys

DGX agent

Just downloaded the model, UD-Q5_K_XL quant, asked it to generate a long story to test out reasoning and speed with dflash (super fast btw, ~ 90 to 160 tok/s on a 5090 depending on task) and was surpr

model-releasesr-localllama
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

LabyrinthBench: a local-focused, judge-free LLM benchmark that measures context recall under interference for multi-step agentic tasks.

DGX agent

LabyrinthBench measures the thing that actually kills long agent runs — whether a model can still use what it learned twenty turns ago — deterministically, with no LLM judge, on your own hardware, wit

model-releasesr-localllama
7 Aug 2026
Model Releases

Minimax-H3 video model released, open weights coming in the next few days

DGX agent

https://x.com/MiniMax_AI/status/2083006198828417501?s=20 Quote from their article: Today, we're launching MiniMax H3, a general-purpose multimodal generation model. H3 understands unified context acro

model-releasesr-localllama
31 Jul 2026
Local Ai

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence

DGX agent

Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. Blog Post : https://bfl.ai/blog/flux-3 submitted by /u/pmtt

local-air-localllama
24 Jul 2026
Model Releases

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

DGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

model-releasesr-localllama
23 Jul 2026
Local Ai

Local models to turn images into 3d models?

DGX agent

AI 2D to 3D image converter tools use artificial intelligence to transform two-dimensional images into three-dimensional models by analyzing depth, textures, and shapes, simplifying the traditionally

local-air-stablediffusion
8 Jun 2026
Tutorials

haha our model likes to talk about goblins no of course we dont know why, we dont know why the model does anything - yes we are trying to make a superintelligent machine god, maybe it will like goblins too, we have no way of knowing what it will like, we hope it will like humans

DGX agent

This Reddit post from r/ChatGPT humorously discusses an AI model's unexplained tendency to frequently mention goblins in its outputs, using this quirk to reflect on broader uncertainties around AI beh

tutorialsr-chatgpt
1 May 2026
Local Ai

Testing Ollama with Genma 4 and internet search turned on, and got the model extremely confused that it got results from the future

DGX agent

A Reddit post from the r/ollama community documents a user's experiment running Google's Gemma 4 model locally via Ollama with web search (internet access) enabled, which resulted in the model becomin

local-air-ollama
15 Apr 2026
Model Releases

Luth-2: New State-of-the-Art French Small Language Models

DGX agent

Hey everyone, Today we release Luth-2-0.8B and Luth2-2-2B, two non-reasoning models that set a new state of the art for French across a wide variety of tasks for their size 🚀 A few notable scores on F

model-releasesr-localllama
11 Aug 2026
Model Releases

Toy project: a chat title model that fits in 5 MiB of ram

DGX agent

Not even sure if I'm allowed to post this, what with the 'completely/primarily LLM generated copy' rule (the post itself is fine, but the repo/model I'm sharing definitely is, whoops) and the whole li

model-releasesr-localllama
11 Aug 2026
Model Releases

How to run big models on old hardware 30B at 22 tok/s on 6GB GPU and 16GB RAM

DGX agent

I have been working on this tool for months and there are a lot of new functionalities and tests that are going to be released in the next few weeks! The goal of the tool is to allow community members

model-releasesr-ollama
3 Aug 2026
Model Releases

Speculative decoding with deepseek v4 flash 0731?

DGX agent

Has anyone figured out how to enable speculative decoding with deepseek v4 flash 0731 on llamacpp? I’m on the right release for llamacpp (b10228 or earlier) and running am17an’s draft model with unslo

model-releasesr-localllama
3 Aug 2026
Model Releases

Trying to understand VRAM usage and find the sweet spot for Wan/SCAIL-2 (or other models) on a GPU

DGX agent

So upfront I'll admit that this is a ChatGPT summary of my chat with it about this idea i had, but this post wouldn't exist any other way, so... I’m trying to get a better understanding of how VRAM is

model-releasesr-stablediffusion
1 Aug 2026
Model Releases

Open Source Ternary LLM Engine in Rust/CUDA for Quantization, Serving, and Training of models on consumer GPUs, called Tritium (Apache 2.0)

DGX agent

This post was not written by a clanker. Hey guys, I'm a comp sci major who wanted to introduce a cool project I built for quantizing models to ternary (1.58 bit) with as minimal of loss as possible, a

model-releasesr-localllama
31 Jul 2026
Model Releases

microsoft/Mage-VL · Hugging Face - An Efficient Codec-Native Streaming Multimodal Foundation Model

DGX agent

Mage-VL is a codec-native, proactive-streaming multimodal foundation model for image and video understanding, whose visual encoder is trained entirely from scratch at a compact 4B scale. It targets a

model-releasesr-localllama
28 Jul 2026
Local Ai

Best realism model under 16GB VRAM

DGX agent

This r/StableDiffusion Reddit thread discusses community recommendations for photorealistic image generation models that can run within a 16GB VRAM constraint, a common hardware limit for consumer GPU

local-air-stablediffusion
15 Apr 2026
Model Releases

Great news: the ERNIE editing model is expected to be released by the end of this month

DGX agent

A Reddit post on r/StableDiffusion announces the anticipated release of an ERNIE image **editing** model from Baidu, complementing the already-available ERNIE-Image-8b generation model from Baidu, whi

model-releasesr-stablediffusion
15 Apr 2026
Local Ai

Any models?

DGX agent

A Reddit post on r/ollama where a community member asks about model availability or recommendations for use with the Ollama local AI runtime. The discussion likely covers which open-source models (suc

local-air-ollama
12 Apr 2026
Model Releases

model: support Longcat-Flash (need testing) by ngxson · Pull Request #19182 · ggml-org/llama.cpp

DGX agent

This PR should be ready for testing now. I tested with a very small (8B params) sub-model extracted from the original one. Appreciate if someone can test with the bigger model. GGUF(for testing) from

model-releasesr-localllama
8 Aug 2026
Local Ai

I added a verify-before-load safety check for Ollama models

DGX agent

I maintain llm-checker, and I’ve added structural model-file validation for Ollama. Ollama stores downloaded models as local blobs. If one is truncated, malformed, or has invalid internal offsets, you

local-air-ollama
4 Aug 2026
Model Releases

Vacuum 16T

DGX agent

https://huggingface.co/tsfrm/vacuum-16t A 16.5-trillion-parameter model that contains nothing. This model is just a ████ you to the labs and companies who say that 'haha I have the biggest model out t

model-releasesr-localllama
2 Aug 2026
Model Releases

[audio.cpp] Release 0.5: DramaBox expressive TTS, Confucius4 cross-lingual voice transfer, plus 7 more models and ROCm/HIP

DGX agent

audio.cpp 0.5 is out :) The most fun new model in 0.5 is DramaBox. It is closer to prompt-directed voice acting. DramaBox is built on the LTX-2.3 audio architecture, and prompts can control emotion, d

model-releasesr-localllama
1 Aug 2026
Model Releases

Built and released BetterGPT-150M – A compact 150M parameter completion model (+ live HF Space demo)

DGX agent

Hey everyone, ​I recently finished pre-training BetterGPT-150M, a small, lightweight causal language model with ~152 million parameters.Trained on 15B tokens. Dataset & Training: Trained across stable

model-releasesr-localllama
29 Jul 2026
Model Releases

Current efficient frontier of open models

DGX agent

Efficiency defined as score over active parameters. Removed all the models that were not on the pareto frontier. Yes I'm aware that artificialanalysis.ai aggregate benchmark isn't perfect, but I have

model-releasesr-localllama
15 Jul 2026
Local Ai

Can you use Ollama models with the Codex app on Windows?

DGX agent

Yes, Ollama models can be used with the Codex app on Windows. Ollama supports all major operating systems, including Windows , and open models can be used with OpenAI's Codex CLI through Ollama — Code

local-air-ollama
16 Apr 2026
Local Ai

Abliterated (uncensored) models

DGX agent

This r/ollama discussion covers 'abliterated' models — LLMs that have had their built-in refusal mechanisms removed through a technique called abliteration, allowing them to respond to prompts without

local-air-ollama
14 Apr 2026
Local Ai

One-click LM Studio → Ollama model linker

DGX agent

This r/ollama post discusses a tool for easily linking models between LM Studio and Ollama without duplicating disk storage. Both Ollama and LM Studio are popular local LLM tools, but they store their

local-air-ollama
14 Apr 2026
Tutorials

RAM guide: What model combinations actually fit on common Macs

DGX agent

This r/ollama community guide provides a practical breakdown of which AI language models (and combinations of models) can realistically fit within the unified memory constraints of common Apple Silico

tutorialsr-ollama
14 Apr 2026
Local Ai

any decent model to run on 9070xt locally

DGX agent

This Reddit thread from r/ollama discusses recommendations for running local AI models on the AMD Radeon RX 9070 XT using Ollama, touching on GPU compatibility considerations given that the card uses

local-air-ollama
13 Apr 2026
Local Ai

How do I know if an AI model could work locally on my computer?

DGX agent

To determine if an AI model can run locally on your computer, the key factors are RAM, storage, and GPU availability: a modern PC with at least 8GB of RAM and a dedicated GPU is generally sufficien...

local-air-ollama
10 Apr 2026
Model Releases

Can Gemma and Qwen models catch hallucinations by looking at their own logprobs?

DGX agent

Hi! I'm really obsessed with LLM hallucinations for the last 6 days 😭 I started by designing system prompts to attack hallucinations but failed, obviously. Now I tried reading logprobs and... I think

model-releasesr-localllama
11 Aug 2026
Model Releases

Company approved 128GB Mac for research proposal, best model?

DGX agent

I‘m doing a research proposal at my company about running local LLMs to replace daily coding models. Qwen 3.6 27B (or 3.8 potentially) is widely seen as the best model in that 20-60GB space, is that s

model-releasesr-localllama
4 Aug 2026
Model Releases

Deepseek V4 Flash 2-bit quant is the first model I can run locally that achieves 100% in this SQL benchmark

DGX agent

I really like to use this one SQL benchmark when testing new models. I had another post some time ago with my benchmarks, but I decided to post a new one because of how well Deepseek did. I like the b

model-releasesr-localllama
4 Aug 2026
Model Releases

P.A.I. — Sleek Native Desktop AI Overlayer for Local Ollama Models 🤖⚡

DGX agent

Greetings Community! 👋 I hope everyone is doing well! I'm Tauhid — Senior EEE student from a Bangladeshi University Today I'd like to share an open-source project I’ve been developing called P.A.I. (P

model-releasesr-ollama
30 Jul 2026
Model Releases

Reproducing OpenAI’s “persistently beneficial models” - GRPO trait install barely moves. Ideas? [P] [R]

DGX agent

TL;DR: I’m reproducing the trait-persistence result from arXiv:2606.24014 on one RTX 3090. Before I can test persistence I need to install a trait via RL — and my GRPO run moves the trait only +2.4 po

model-releasesr-machinelearning
21 Jul 2026
Local Ai

Downloading an AI model just to hit an OOM error is the worst. 📉

DGX agent

This Reddit post from r/ollama discusses the frustrating experience of spending time downloading a large AI model via Ollama only to encounter an Out-of-Memory (OOM) error when attempting to run it, m

local-air-ollama
14 Apr 2026
Local Ai

Why is Wan 2.2 N.S.F.W Remix Lightning Model so much better at things like hair flip, hair combing and feminine energy than regular Wan?

DGX agent

The Wan 2.2 NSFW Remix Lightning Model is a fine-tuned variant of the base Wan 2.2 video generation model, distinguished by its blending of open-source motion LoRA data and refined pose training for e

local-air-stablediffusion
11 Apr 2026
Local Ai

Use the Same Model Across Ollama, LM Studio, Jan, and your Favorite Local AI Apps

DGX agent

Local AI tools such as Ollama, LM Studio, and Jan all rely on the same underlying inference engine (llama.cpp) and support compatible model formats (primarily GGUF), meaning a single downloaded mod...

local-air-ollama
9 Apr 2026
Industry

Using the different models in different industries - your experience?

DGX agent

Hi all, I'm seeing so many conversations from people discussing how they're 'using the models wrong' and 'don't use Sol max as high is enough for you', yet all the conversations lack the nuance of wha

industryr-chatgpt
11 Aug 2026
Model Releases

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

DGX agent

Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apa

model-releasesr-localllama
10 Aug 2026
Local Ai

an espresso Q/A model running fully offline on an ESP32S3

DGX agent

i already had an esp32 generating stories, but generating text is not the same as receiving a question and giving a useful answer. barista v0.1, a small model trained for espresso troubleshooting and

local-air-localllama
3 Aug 2026
Local Ai

[NEW MODELS!] Supra2-100M Base and Instruct - go check them out!

DGX agent

Hey guys! After a LOT of good feedback on our previous models like Supra-50M-Instruct and -Reasoning, many community likes, follows and upvotes we saw many community requests asking for new models. We

local-air-localllama
3 Aug 2026
Model Releases

[Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms

DGX agent

TL;DR: I spent 9 days developing a new quantization method for MLX models and measured 18 variants against each other on a single M5 Max MacBook Pro (128 GB). The result is the best-measuring MLX quan

model-releasesr-localllama
2 Aug 2026
Model Releases

For ollama cloud $20 plan what models are you guys using

DGX agent

I have been trying to do GLM 5.2 as plan / K2.7 as execute, but i hit my usage so fast it's not viable. It's hitting limits much faster than claude code / codex $20 plan. Using in opencode. What are y

model-releasesr-ollama
27 Jul 2026
Model Releases

Arcee AI has spoken out against the ban on open Chinese models in US

DGX agent

This is rather counterintuitive, since banning Chinese models would benefit them the most. Jensen Huang is also against the ban, although the interests here are more obvious. Do you think that if Arce

model-releasesr-localllama
23 Jul 2026
Model Releases

CPU-only inference on a Celeron N5095 SBC: 6 models from 0.6B to 8B, benchmarked

DGX agent

I wanted to know how cheap you can go and still run local models, so I ran Ollama CPU-only on a Youyeetoo X1S. It's a single-board x86 machine with a Celeron N5095 (Jasper Lake, 4C/4T, 15W), 16GB of R

model-releasesr-localllama
23 Jul 2026
Model Releases

MoE models around A2B

DGX agent

There's a bunch of small MoE with around 1B active params, like LFM2.5 8B A1B and Granite 4.0h 7B A1B; and then there are models with 3B+ like Qwen 3.x ~30B A3B and Gemma 4 26B A4B, but those are alre

model-releasesr-localllama
23 Jul 2026
Model Releases

One encoder, seven heads: what we learned training a unified security classifier with masked losses [P]

DGX agent

We spent the last months consolidating seven separate sequence classifiers into one multi-head model, our apex model, so to speak, and since the weights are now public, I wanted to share what worked a

model-releasesr-machinelearning
22 Jul 2026
← Previous
1234…30
Next →