AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
1,433 results
Tutorials

How to train loras in One Trainer for Z Image using Civitai models?

DGX agent

This r/StableDiffusion post likely discusses the community challenge of using OneTrainer — a locally-installed, GUI-based LoRA training tool — to train LoRAs specifically compatible with Z Image (Tong

tutorialsr-stablediffusion
12 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

KIV: 1M token context window on a RTX 4070 (12GB VRAM), no retraining, drop-in HuggingFace cache replacement - Works with any model that uses DynamicCache [P]

DGX agent

KIV is a project shared on r/MachineLearning presenting a drop-in replacement for HuggingFace's `DynamicCache` that enables up to 1 million token context windows on consumer hardware with only 12GB of

model-releasesr-machinelearning
12 Apr 2026
Local Ai

What's model should I run?

DGX agent

A Reddit discussion from the r/ollama community where a user seeks advice on which AI language model to run locally using Ollama. Responses likely include hardware-based recommendations (such as RAM a

local-air-ollama
11 Apr 2026
Local Ai

kugel-2 model (VibeVoice finetune) repo is gone. Does anyone know why?

DGX agent

The kugel-2 model is a community fine-tune of Microsoft's VibeVoice, a text-to-speech model — and its disappearance is rooted in Microsoft's own removal of the base VibeVoice repository. Microsoft...

local-air-stablediffusion
10 Apr 2026
Local Ai

VoxCPM TTS model + LoRa training abilities right in Comfy

DGX agent

ComfyUI-VoxCPM is a custom node that integrates VoxCPM — a novel tokenizer-free Text-to-Speech system that models speech in a continuous space — directly into ComfyUI's visual workflow environmen...

local-air-stablediffusion
10 Apr 2026
Model Releases

FlowInOne - A new Multimodal image model . Released on Huggingface

DGX agent

FlowInOne is a vision-centric multimodal image generation framework that reformulates multimodal generation as a purely visual flow, converting all inputs into visual prompts and enabling a clean i...

model-releasesr-stablediffusion
9 Apr 2026
Model Releases

I tested the CMP170HX

DGX agent

Lots of rumor and misinfo bouncing around, so I put some of these old mining cards to the test. I used 4 of the 8GB cards, set to 64GB each. Lots of models fit entirely on a single card, and you can a

model-releasesr-localllama
11 Aug 2026
Model Releases

Chat UIs with native audio input for multimodal models?

DGX agent

I've been running Gemma 4 E4B with oMLX and I can't find any chat interfaces that directly send the audio file to the model instead of running the audio through a separate STT layer. I can confirm the

model-releasesr-localllama
10 Aug 2026
Model Releases

Kimi K3 (Unsloth) IQ2-XXS from 711GB down to 478GB!!! Only Multi-language was removed to trim the size

DGX agent

Firstly a big thanks to the poster 'hellohazine', he basically only removed the multi-lingual fat of the model and just kept the English language intact. It is the exact model, and the rest of the mod

model-releasesr-localllama
8 Aug 2026
Local Ai

Best open-source harnesses for combining cloud and local AI model orchestration?

DGX agent

Looking for best current solutions for combining cloud models and local models seamlessly inside a harness' orchestration Edit: Right now, we don't have harnesses (that I'm aware of) that are blending

local-air-localllama
6 Aug 2026
Model Releases

Kimi K3 full model running on 16x GB10 cluster at 20+tps

DGX agent

Kimi K3 full model running on 16x GB10 cluster at 20+tps average (llama-benchy coherent corpus) 38tps peak, 750tps prefill. This is the first run of full k3 with dspark on my cluster. I will be doing

model-releasesr-localllama
4 Aug 2026
Model Releases

I built an open-source LLM Gateway to route and fallback between local Ollama models and cloud APIs

DGX agent

Hey r/ollama 👋 If you run Ollama locally alongside cloud endpoints for agent workflows, Cursor/Windsurf, or custom scripts, managing API switching, failover logic, and context limits can get messy fas

model-releasesr-ollama
2 Aug 2026
Model Releases

Deepseek V4 Flash is now ~#2 open weight model to Kimi K3 and >50x cheaper

DGX agent

https://preview.redd.it/h7zv5tb3tmgh1.png?width=2854&format=png&auto=webp&s=507380e8f862c18f10f7c5c84da9e8d1c59139b0 Deepseek's new flash model is unexpectedly cheap and high-performing across useful

model-releasesr-localllama
31 Jul 2026
Model Releases

Why are AI model tests always the same generic prompts?

DGX agent

Okay, hear me out. Why is it that every time a new model comes out, all the tests I see are 'make a car game,' 'make a website,' or something equally generic, usually from a prompt that's barely a lin

model-releasesr-localllama
31 Jul 2026
Model Releases

Agenta: an open-source Claude Cowork alternative where you can use self-hosted models (and any harness)

DGX agent

Hey r/LocalLLaMA, I’m Mahmoud from Agenta. We built a self-hosted, more flexible, alternative to Claude Cowork . This short video shows how it works. I use it to build AI coworkers for my startup, lik

model-releasesr-localllama
28 Jul 2026
Model Releases

Might need math+code benchmark for frontier model(LLMs Silently Replace Math)[D]

DGX agent

Hello guys. I found some problems in current frontier models. And want to share. # math_code_hallucination > Record of a failure caused by combining mathematics and code in a single prompt. --- ## Cas

model-releasesr-machinelearning
28 Jul 2026
Model Releases

What would it take for the frontier labs to open the weights of their old, deprecated proprietary models?

DGX agent

Anyone thought about this? What do you think needs to happen for them to release the old weights? I’d love to see models like Gemini-2.5, OAI o3, 4o, 4.1 being open one day. In Oct 2025 Scam Altman sa

model-releasesr-localllama
28 Jul 2026
Safety

White-hat hacking IS the defense to black-hat hacking. The techniques are the same. How does Dario expect companies to do it if their models refuse?

DGX agent

You patch security holes by intentionally finding them. If the models refuse to do it, how can companies protect themselves against rogue AIs, whether they are Chinese or OpenAI/Anthropic themselves?

safetyr-localllama
28 Jul 2026
Model Releases

Is it worth getting 128GB MacBook Pro? Will it ever be comparable to today’s frontier models for coding?

DGX agent

I am a long time iOS app developer. In the last year I have been using Cursor+Claude/others to assist with app development. I am concerned that the current low pricing will disappear eventually. I am

model-releasesr-localllama
25 Jul 2026
Model Releases

Honest take on Laguna S2.1 and its uses (from actual use)

DGX agent

So I've taken some time to actually test laguna on a few of my own projects. I wanted to share as I feel most peoples comments at this point have just been about getting it running or saying it doesnt

model-releasesr-localllama
24 Jul 2026
Local Ai

More than 20 companies including NVIDIA, Meta, Microsoft, Palantir, and Hugging Face have signed a letter urging policymakers to avoid premature restrictions on open weight models.

DGX agent

The Open Letter was initiated by Microsoft and published today: “Open Weights and American AI Leadership”. It argues against broad or premature restrictions on open-weight models and explicitly says p

local-air-localllama
24 Jul 2026
Model Releases

Cactus Hybrid: We taught Gemma 4 to know when it's wrong

DGX agent

Hey HN, Henry & Roman here from Cactus. A small, on-device model is fast and private, but sometimes wrong, but frontier models are getting expensive pretty fast. So, we post-trained Gemma 4 E2B post-t

model-releasesr-localllama
22 Jul 2026
Model Releases

My OCR model mislabels section titles as body text. Is a CRF the right fix, or am I overcomplicating it? [P]

DGX agent

Hi everyone, I'm working on extracting the hierarchical structure of long PDF documents (legal/regulatory text, lots of numbered sections) and would like to gather some feedback on my approach before

model-releasesr-machinelearning
21 Jul 2026
Local Ai

I took Andrej Karpathy's LLM Council concept to the next level (Docker, MCP, Skill, Search, local (Ollama)/cloud model support and much more)

DGX agent

A developer expanded on Andrej Karpathy's LLM Council concept by implementing an enhanced system with Docker containerization, Model Context Protocol (MCP) integration, skill modules, web search capab

local-air-ollama
10 Jun 2026
Hardware

Xiaomi just claimed 1,000+ tps on a 1T model using a standard 8-GPU server

DGX agent

Xiaomi achieved over 1,000 tokens per second output from a 1 trillion-parameter model using a single standard 8-GPU commodity node through extreme model-system codesign . The approach combines FP4 qua

hardwarer-localllama
8 Jun 2026
Local Ai

How to add specific knowledge to an ollama model?

DGX agent

Adds knowledge to Ollama models using Retrieval-Augmented Generation (RAG) , where users create a knowledge base directory with reference files like PDFs, text files, or CSVs . A custom model can be c

local-air-ollama
2 Jun 2026
Model Releases

Bernini released. Unified Video generation and editing model. Built on Wan-2.2

DGX agent

Bernini is a unified framework for video editing and video generation , built using Wan2.2-A14B as its renderer . The model covers complementary task families that demonstrate its capabilities as a un

model-releasesr-stablediffusion
1 Jun 2026
Local Ai

Stable Diffusion model recommendations for faster and cleaner outputs in 2026?

DGX agent

A Reddit discussion seeking Stable Diffusion model recommendations for faster and cleaner image outputs, addressing the reality that no single best model exists as the right choice depends on hardware

local-air-stablediffusion
31 May 2026
Local Ai

New LFM2.5 8b A1b model!!

DGX agent

Liquid AI released LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and lightweight server-side use-cases. The model is a fast, memory-

local-air-ollama
29 May 2026
Local Ai

New to OLLAMA, how to install best model for my Mac?

DGX agent

Ollama can be installed on Mac by downloading the application and placing it in the Applications folder, after which you use Terminal to run commands that download and launch models . The best model c

local-air-ollama
23 May 2026
Research

Can liveness detection models generalise to synthetic media generation techniques they were never trained on? [D]

DGX agent

Liveness detection and synthetic media detection models often fail to generalize across unseen data and struggle with content from different models. Understanding how factors like data source diversit

researchr-machinelearning
21 May 2026
Local Ai

Tencent is about to release an anime video model (AniMatrix).

DGX agent

Tencent launched Hunyuan Video in December 2024, an open-source AI video generation model with 13 billion parameters that supports text-to-video and image-to-video conversion. The model brings unique

local-air-stablediffusion
6 May 2026
Local Ai

Been noticing a lot of 'slow responses' today: models do not inherently more slow, rate limiting more likely.

DGX agent

A Reddit discussion from the Ollama community addresses reports of slow model responses, clarifying that the models themselves are not inherently slower but that rate limiting is a more likely cause o

local-air-ollama
4 May 2026
Local Ai

Need suggestions for Rag model

DGX agent

A Reddit discussion in the r/ollama community seeking recommendations for RAG (Retrieval-Augmented Generation) models, likely addressing model selection for local LLM-based document retrieval systems.

local-air-ollama
1 May 2026
Local Ai

Klein 9B Distilled vs. five different cloud API models

DGX agent

The Klein 9B Distilled is a distilled image generation model enabling sub-second image generation with text-to-image and image-to-image editing capabilities in a single unified model, designed for rea

local-air-stablediffusion
24 Apr 2026
Local Ai

Stop burning GPU time on the wrong model. Manifest now supports Ollama Cloud 🦙🦚

DGX agent

Manifest is a tool or platform that has added support for Ollama Cloud, enabling users to route AI model inference to cloud-hosted Ollama instances rather than relying solely on local GPU resources. T

local-air-ollama
14 Apr 2026
Local Ai

Does anyone know which model and potentially Lora was used to create these?

DGX agent

This Reddit thread from r/StableDiffusion is a community-driven reverse-identification request, where a user shares AI-generated images and asks fellow community members to help determine which Stable

local-air-stablediffusion
13 Apr 2026
Local Ai

I made a playable ping pong game where every frame is ai generated. This is my interactive diffusion model I made from scratch.

DGX agent

A Reddit user on r/StableDiffusion showcased a fully playable ping pong game in which every individual frame is rendered in real time by a custom-built interactive diffusion model, rather than using t

local-air-stablediffusion
13 Apr 2026
Model Releases

Inpaint workflows for z-image, qwen and flux fill onereward

DGX agent

This Reddit post from r/StableDiffusion shares ComfyUI inpainting workflows for several modern AI image models, including Z-Image, Qwen Image/Edit, and Flux-series models . Flux Fill is a dedicated in

model-releasesr-stablediffusion
13 Apr 2026
Local Ai

llama4 108b

DGX agent

This Reddit thread on r/ollama discusses running Meta's Llama 4 Maverick — a ~108B parameter model — locally using Ollama. Llama 4 models are natively multimodal AI models supporting text and image un

local-air-ollama
13 Apr 2026
Model Releases

Native Long Video Understanding Models locally?

DGX agent

I've been building a personal project and wanted to check with the community on multi-modal inputs since I can't find a lot of material around this online. Ultimately I'm trying to build something tha

model-releasesr-localllama
10 Aug 2026
Local Ai

AMA: MiniMax H3 Team — Ask us anything about our open video generation model, training, and future plans

DGX agent

https://preview.redd.it/kihat320ashh1.png?width=1672&format=png&auto=webp&s=a7ccc40ba3fb229ac7ebf57e8e6a314e0ee45646 Hi r/StableDiffusion! u/New-Requirement1419 -> dacongya (Head of H3 Researcher) u/A

local-air-stablediffusion
6 Aug 2026
Model Releases

The death of SLMs?

DGX agent

I love to see these impressive models coming out that compete with the giants from companies like Z.ai, Moonshot, Alibaba, etc. A win for the open source/weight community is always welcome. While I am

model-releasesr-localllama
6 Aug 2026
Model Releases

Why are Chinese models better* at Frontend than the western top labs?

DGX agent

I use A LOT both openAI and Anthropic products. When I need some frontend work (pure web dev) (or answer that feel less verbose and more to the point) I use Anthropic. For multimodality openAI feels b

model-releasesr-localllama
4 Aug 2026
Model Releases

Encrypted Clouds?

DGX agent

I love the progress happening on open models but I feel like it is kind of getting clear that hardware to run good sized models is completely unaffordable for me right now. I know that you all love Qw

model-releasesr-localllama
2 Aug 2026
Model Releases

Best C++ Local Model? (July 24th 2026 Edition :-P)

DGX agent

I apologize that this question has been asked in various flavors over time, but I couldn't find anything in the posts before that matches the options I have. I have a PC and a mac, both available in m

model-releasesr-ollama
25 Jul 2026
Local Ai

which model to use on local 24gb mac mini M4 pro

DGX agent

So, i have been building some apps that should run on the local every user system, tried gemma4 although its fast and great at reasoning its not as good in instructions following and tool calling. tri

local-air-ollama
24 Jul 2026
Hardware

Haven't seen much about the Nvidia Cosmos 3 video model that dropped, what's up with that?

DGX agent

Nvidia Cosmos 3 is an open physical AI foundation model built on a mixture-of-transformers architecture that combines vision reasoning, world generation, and action prediction for reasoning, simulatio

hardwarer-stablediffusion
9 Jun 2026
← Previous
123456…30
Next →