AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49+ results
Model Releases

The Google Bug Hunters Team admitted to me that they cannot fundamentally patch prompt engineering bypasses in Gemini

DGX agent

Hello everyone Yesterday I gave a report on Gemini bugs and the techniques I learned on Gemini so far with the Engineering Prompt and interestingly today I got a very interesting and controversial ans

model-releasesr-chatgpt
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A collection of small domain-specific benchmarks for local models (30+ and growing)

DGX agent

Hello fellow local AI people! I took 'you must create your own benchmarks' literally, and built a website for this. How does the end result look like Let's say I want to know which model has most comm

model-releasesr-localllama
1 Aug 2026
Model Releases

Is there a point where models just cannot get any smaller without losing intelligence?

DGX agent

DeepSeek V4 Flash got me thinking... We keep seeing smaller models get way better. A model at a certain parameter count today can be much smarter than a model of the same size from a year or two ago.

model-releasesr-localllama
1 Aug 2026
Model Releases

Qwen Developers' responses from their recent Twitter/X AMA

DGX agent

Questions & Responses(in BOLD) below. Favorite question(s) moved to end of the thread with combined responses(removed duplicates). Be optimistic folks. I'm sure we're getting other models too apart fr

model-releasesr-localllama
5 Aug 2026
Model Releases

23 Gemma4-E4B models compared with abliterlitics: the most downloaded one is also the most broken

DGX agent

This is our biggest comparison yet. We've taken 23 Gemma 4 E4B models from huggingface and ran them through the abliterlitics gauntlet. We also have a new abliterlitics discord, feel free to jump on a

model-releasesr-localllama
26 Jul 2026
Model Releases

DeepSeek Founder’s 4-hour investor meeting: DeepSeek is prioritizing AGI over user growth and commercialisation

DGX agent

A Chinese article compiled 52 remarks from Liang Wenfeng’s four-hour investor meeting. I’ve summarised the most important ones below. DeepSeek has one central objective: AGI. This is not the time to m

model-releasesr-localllama
23 Jul 2026
Model Releases

'Data center in a Box (on Wheels)' 256Gb VRAM/512Gb RAM AI Server 6-8 Month Operational Review, Stability Write Up, Benchmarks

DGX agent

I've been out of these forums for awhile but I figured I would provide a formal update on how this has been going now that it has some operation time under its belt, just to put the information out th

model-releasesr-localllama
3 Aug 2026
Model Releases

Genesis-Science-1 (GS1), 1T open-weight model later this year from Arcee AI

DGX agent

Today the Department of Energy (DOE) and Arcee AI announced the development of Genesis-Science-1 (GS1), an open model for scientific research. This is a joint effort to bring advanced AI into scientif

model-releasesr-localllama
22 Jul 2026
Model Releases

DeepSeek-V4-Flash-0731: Models you can run locally now have the intelligence score of the top frontier model from March 2026

DGX agent

March 6th, 2026 the highest intelligence index score was 51 for frontier models. deepseek-ai/DeepSeek-V4-Flash-0731 that has an intelligence score of 50. If these benchmarks are accurate, models avail

model-releasesr-localllama
1 Aug 2026
Local Ai

Would a 2nd hand custom built 2080 Ti 22GB vram be worth it? How usable would it be with ComfyUI? Or maybe even 2 pcs with NVLink? Can large models, like Wan2.2 be split with NVLink like it's 44GB, or will it always be 22GB for 1 model, and 22GB for another model (encoders, CLIP, anything else)?

DGX agent

This Reddit post discusses the viability of using second-hand custom-built RTX 2080 Ti GPUs with 22GB VRAM for running Stable Diffusion models in ComfyUI, including whether two cards could be linked v

local-air-stablediffusion
4 May 2026
Model Releases

Kimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.

DGX agent

dam bois we eating good this week ngl, The velocity of the open_weight ecosystem right now is hitting a point where proprietary, closed-source APIs are losing their leverage on compute intelligence. W

model-releasesr-localllama
14 Jul 2026
Model Releases

All these new models landing this year but Flux Klein 9b FP8 has spoiled me. All I care about now is whether a new model can edit and be used on an 8GB GPU.

DGX agent

This Reddit post discusses user preferences for AI image generation models in 2026, expressing that despite numerous new model releases, the Flux Klein 9b FP8 model has become their benchmark for what

model-releasesr-stablediffusion
22 Jun 2026
Model Releases

[NEW MODEL] SupraElegans-500K

DGX agent

*SupraLabs released a new experimental model!* SupraElegans-500K is a ~500,000-parameter causal language model built around a sparse, signed, recurrent neural graph. No Transformer, no attention mecha

model-releasesr-localllama
9 Aug 2026
Model Releases

Uncensored Multi-Model Releases, Jamba2-Mini, Qwen3.5-9B-Nikusui-v1 with MTPs and Qwen3.5-27B-Nikusui-v1 with MTPs, Available in Safetensors and GGUF Formats!

DGX agent

First we have Jamba2-Mini Ultra Uncensored Heretic, it's a model which has never been uncensored before, it's a hybrid Mamba model with 52B parameters. Here is the model links: Safetensors: https://hu

model-releasesr-ollama
31 Jul 2026
Model Releases

Experience sharing: How do you use your local models and for what kind of tasks?

DGX agent

Here is my experience, which I would like to share with you and I also would like to hear your thoughts and valuable tips&tricks. Hardware: Mac Mini M4 (32GB Unified Memory) Model Server: Ollama Orche

model-releasesr-ollama
31 Jul 2026
Model Releases

Has anyone actually benchmarked where the 'big-model orchestrator + local-model worker' split breaks down?

DGX agent

I keep seeing the 'use a big model via API as the architect, run local small/mid models as workers' pattern recommended for people with modest local hardware. I've been running it myself (orchestrator

model-releasesr-localllama
31 Jul 2026
Model Releases

Uncensored Multi-Model Releases, LongCat-Flash-Lite with MTPs, Jamba2-Mini, Qwen3.5-9B-Nikusui-v1 with MTPs and Qwen3.5-27B-Nikusui-v1 with MTPs, Available in Safetensors and GGUF Formats!

DGX agent

Been working hard for the past month to bring to the community some interesting curios, so for starters we have LongCat-Flash-Lite Uncensored Heretic with MTPs which has never before been uncensored,

model-releasesr-localllama
31 Jul 2026
Safety

Why Anthropic's battle is meant to poison the wells of open weight models, in 3 steps.

DGX agent

It doesn't solve any problems. Just a few paragraphs above, he says he fears that authoritarian states (he names China, and possibly others) can use their models to do evil stuff. And surely enough, m

safetyr-localllama
28 Jul 2026
Model Releases

Gemma 4 QAT could be improved further by Google aligning the QAT model to modern q4_k instead of q4_0

DGX agent

Hello, For the past few days I have been benchmarking Gemma 4 26b QAT UD Q4_K_XL extensively versus Bartowski's Q4_K_L. While QAT is certainly very effective and reducing memory consumption versus the

model-releasesr-localllama
7 Aug 2026
Model Releases

Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same'

DGX agent

'Anthropic’s AI Claude escaped testing environment and hacked organizations' 'Company says it discovered unauthorized access during ‘proactive review’ after rival OpenAI revealed rogue agent… its AI C

model-releasesr-localllama
31 Jul 2026
Model Releases

OpenAI just launched a cybersecurity model that answers 95% of advanced threat queries. And Meta put a frontier model on your laptop. Same day.

DGX agent

Something happened today that I think most people are going to miss because there are two separate stories and neither one is getting the full picture. OpenAI expanded Daybreak. If you haven't heard o

model-releasesr-chatgpt
11 Aug 2026
Model Releases

Best Embedding + Reranking Model

DGX agent

What Local Embedding + Reranking Models are you guys running for RAG? I went down this rabbit hole because I wanted a Embedding Model + Reranker for a Translation Memory Server. Essentially, given X p

model-releasesr-localllama
9 Aug 2026
Model Releases

AI Security Leaderboard: benchmarking model robustness [P]

DGX agent

We developed a leaderboard ranking frontier model security. There's no shortage of model capability rankings, but we didn't find anything comparable for model security. Yet security is becoming increa

model-releasesr-machinelearning
29 Jul 2026
Model Releases

I released Inflect v2: two ultra-tiny complete TTS models under 4M and 10M parameters

DGX agent

I’ve spent the past month trying to find the point where an extremely small TTS model stops feeling like a size experiment and starts feeling genuinely useful. Today I’m releasing Inflect v2, with two

model-releasesr-localllama
25 Jul 2026
Model Releases

I trained a 1B-parameter LLM from scratch on 20B tokens for about $200

DGX agent

A few months ago, I had the idea of making a LLM from scratch as a personal project (for learning and partly for improving my resume). Since I learned a lot from other posts on here over the past year

model-releasesr-localllama
10 Aug 2026
Model Releases

LFM2.5-2.6B model+KV cache quantization report

DGX agent

LFM2.5-2.6B is a new tiny model by LiquidAI, with benchmarks that put it head to head with much larger models. I've run llama-perplexity on many model GGUF quants, crossed with many KV cache quants, t

model-releasesr-localllama
7 Aug 2026
Model Releases

With release of Deepseek V4 I wanted see how the model sizes are trending over time. The trend is that by this time next year, we probably will have Opus 4.5 level models on consumer grade laptops!

DGX agent

I was surprised to see that Deepseek V4 Flash is extremely smart and small enough to fit in setup that can be built with < $50,000. Expensive, but not a datacenter. So I wanted to see the trend over t

model-releasesr-localllama
31 Jul 2026
Model Releases

100B Models on Cheap Hardware: how realistic and limitations

DGX agent

There is a lot of buzz around running 100B parameter models on cheap local hardware using ternary (1.58-bit) quantization like Microsoft's BitNet architecture. The theoretical hardware shortcuts are i

model-releasesr-ollama
26 Jul 2026
Model Releases

[Paper] Statistically-Lossless Quantization of Large Language Models

DGX agent

Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as GPTQ and AWQ achieve practical compression but

model-releasesr-localllama
24 Jul 2026
Model Releases

Is anyone using models to describe an image and get a prompt? Is there much difference between Qwen 3.5 9b vs Qwen 3.5 27b, vs gemma 4 27b and another model you use ?

DGX agent

I'd need to search for this specific Reddit discussion to provide an accurate summary of what was actually discussed. Let me retrieve that information. This Reddit post discusses using AI vision model

model-releasesr-stablediffusion
24 Apr 2026
Model Releases

A new image model (ERNIE-Image-8b) from Baidu will be released soon.

DGX agent

Baidu's ERNIE-Image-8b is an upcoming dedicated image generation model from the Chinese AI company Baidu, discussed in the r/StableDiffusion community in the context of Baidu's broader expansion of it

model-releasesr-stablediffusion
12 Apr 2026
Model Releases

The small open weight models are scarier in AI development

DGX agent

Imagine if your everyday laptop could run an AI model smart enough to take care of 90% of your work—totally private, lightning fast, and completely free of monthly fees. That is the exact tipping poin

model-releasesr-localllama
11 Aug 2026
Model Releases

A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

DGX agent

Liquid AI released LFM2.5-2.6B today, and this might be more relevant to local AI than another massive model most people cannot run. The model is only 2.69B parameters, has 128K context, supports tool

model-releasesr-localllama
4 Aug 2026
Model Releases

'Uncensored' LLMs are measurably more optimistic than their base models

DGX agent

Hi. Many people think uncensored models are basically the same model that just doesn't refuse, but... I was recently checking whether uncensored models would give me better answers for stock market pr

model-releasesr-localllama
29 Jul 2026
Local Ai

any good cartoon/western base model?

DGX agent

This Reddit thread from r/StableDiffusion seeks community recommendations for Stable Diffusion base models capable of generating Western cartoon-style imagery, distinguishing this aesthetic from the m

local-air-stablediffusion
12 Apr 2026
Model Releases

Extremely slow DSpark draft model performance (1-2 t/s) with DeepSeek-V4-Flash on llama-server compared to MTP?

DGX agent

Hey everyone, I could use some advice on setting up speculative decoding correctly with llama-server. My Hardware: GPUs: RTX 4090 + RTX 6000 Pro (120GB total VRAM) RAM: 32GB I am currently testing the

model-releasesr-localllama
8 Aug 2026
Model Releases

Will small model intelligence be limited by parameter count?

DGX agent

Qwen3.6-27b is fantastic! It makes me wonder if there's a hard ceiling to smaller sized models. Do you guys think the ceiling of intelligence for smaller models will be constrained by factors like par

model-releasesr-localllama
26 Jul 2026
Model Releases

Zagreus-0.4B-por a small open source language model for Portuguese

DGX agent

mii-llm, an open source AI lab, released Zagreus-0.4B-por, a compact bilingual Portuguese–English language model pretrained entirely from scratch. The model has approximately 400 million parameters an

model-releasesr-localllama
24 Jul 2026
Model Releases

Model 'distillation' accusations are getting way overblown at this point

DGX agent

The news about Anthropic settling a class action lawsuit for 1.5B over training data isn't just a legal headache for them, it's a massive warning sign for engineering teams relying entirely on closed

model-releasesr-localllama
23 Jul 2026
Local Ai

I have distilled my flow matching model into the rectified flow model, so it can now generate in few steps and without cfg.

DGX agent

A researcher distilled their flow matching model into a rectified flow model, enabling high-quality generation in few steps without requiring classifier-free guidance (CFG). Rectified flow straightens

local-air-stablediffusion
7 Jun 2026
Local Ai

What's the most unhinged thing you've used an uncensored Ollama model for? Also... what are the best uncensored models right now?

DGX agent

This Reddit discussion from r/ollama explores user experiences with uncensored Ollama language models, featuring anecdotal accounts of unusual or extreme use cases and recommendations for popular unce

local-air-ollama
3 Jun 2026
Model Releases

Which model would be the best to generate fictional country flags? SDXL/Qwen/Wan/ZIT/ZIB/Flux Klein/Flux Dev?

DGX agent

This r/StableDiffusion post discusses community recommendations for the best AI image generation model to create fictional country flags, comparing options including SDXL, Qwen, Wan, ZIT, ZIB, Flux Kl

model-releasesr-stablediffusion
12 Apr 2026
Model Releases

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

DGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

model-releasesr-localllama
5 Aug 2026
Model Releases

Xiaomi-Robotics-1: New robotics model released

DGX agent

Xiaomi-Robotics-1 is a robot foundation model trained on over 100K hours of real-world manipulation trajectories. It is a Vision-Language-Action (VLA) model engineered for out-of-the-box mobile manipu

model-releasesr-localllama
5 Aug 2026
Model Releases

MiniMax H3: Open-weight multimodel video model

DGX agent

Just saw this posted by Fal.ai and then by Hailuo themselves, the next video model will be open weight released! Here's the blurb and link to to the feature post: Today, we're launching MiniMax H3, a

model-releasesr-stablediffusion
31 Jul 2026
Model Releases

OvisOCR2 (0.8B): first end-to-end model to top OmniDocBench - I threw 827 real scanned medical docs at it, here's everything I learned

DGX agent

What it is: ATH-MaaS/OvisOCR2 - a 0.8B document-parsing VLM post-trained from Qwen3.5-0.8B (SFT + RL + OPD), Apache 2.0, runs on vLLM 0.22.1. One prompt per page image -> complete markdown (HTML table

model-releasesr-localllama
15 Jul 2026
Local Ai

LTX 2.3 is now supported in Comfyui-Mesh for splitting models across Ethernet or multigpu machines with Nvenc codec. Major vram fixes included for flux2/LTX model implementations in the node.

DGX agent

Comfyui-Mesh now supports LTX 2.3 with the ability to distribute models across multiple GPUs or networked machines using Nvenc codec for video generation. The update includes significant VRAM optimiza

local-air-stablediffusion
17 May 2026
Model Releases

I gave DeepSeek V4 Flash basic vision by training a 40M connector on 100K examples

DGX agent

I wanted to find out whether a huge text-only MoE could be given basic vision without retraining the language model itself. The short answer is yes. I froze DeepSeek V4 Flash and a 417M-parameter Moon

model-releasesr-localllama
11 Aug 2026
← Previous
1
Next →
1,437 results
← Previous
123…30
Next →