AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “people”

GridTimelineEvolution
192 results
Local Ai

Rumored 50-series Super refresh bumps everything +50% VRAM

DGX agent

leaked Super specs have the 5070 Ti and 5080 going 16GB to 24GB and the 5070 to 18GB, thanks to the new 3GB GDDR7 modules. 24GB on a Ti-class card is actually the number people here have been waiting

local-air-localllama
11 Aug 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Tested in Coding: BF16 Muse Glimmer vs BF16 Qwen3.6 27B

DGX agent

I'm guessing that many people have been waiting for this comparison. For clarity, both models are running at full FP16 KV-cache. Due to VRAM limitations, Muse Glimmer is running full 262,144 context,

model-releasesr-localllama
11 Aug 2026
Model Releases

Summary of Takeaways from the Minimax AMA

DGX agent

Summary from https://www.reddit.com/r/StableDiffusion/comments/1vh9rtw/ama_minimax_h3_team_ask_us_anything_about_our/ This summary was compiled with AI but cross-checked manually by me for accuracy. I

model-releasesr-stablediffusion
10 Aug 2026
Model Releases

Underestimated budget solution: radeon 780m iGPU

DGX agent

There are so many posts where people complaining about high prices and asking for solution <= 1000 EUR. So, there is one solution to consider: PC/mini PC/laptop on Ryzen 7 260/Ryzen 9 8945HX/etc CPU w

model-releasesr-localllama
9 Aug 2026
Model Releases

Qwen3.6 27B + 35B on vLLM, single R9700 (gfx1201)

DGX agent

I've been tuning my new Radeon AI Pro R9700, and figured that this would be useful information for people who are trying to optimise their setups. I'm pretty happy with these results and looking forwa

model-releasesr-localllama
8 Aug 2026
Model Releases

Qwen 3.6 27B flags/settings in llama.cpp

DGX agent

I run the following on a 5090 and have been okay with its performance, it does most things somewhere 80-100 t/s, though that can slow down at full 262k context - more like 40 t/s at times. I use it pr

model-releasesr-localllama
7 Aug 2026
Model Releases

Dual 3090 setup: 400 pp t/s to 1600 pp t/s on Qwen 3.6 27B... with slightly lower tps.

DGX agent

First of all, my setup: Ryzen 9 5950x DDR4 3200Mhz 64gb (2x32) Dual 3090s, no NVLINK Runtime: llama.cpp Nvidia Drivers 610 Windows 11 25H2 Qwen 3.6 27B Q8 I've been using llama-server with --split-mod

model-releasesr-localllama
6 Aug 2026
Local Ai

Projects created using OLLAMA

DGX agent

Hi! I'm a new user of Ollama here. What are some GitHub projects that I can look at and play with to see how people work with the Ollama library across various programming languages? I wouldn't want t

local-air-ollama
5 Aug 2026
Model Releases

Qwen3-TTS voice cloning is now in mainline llama.cpp — the old demo finally became real support

DGX agent

People may remember the Qwen3-TTS llama.cpp demo from a few months ago. That PR said it probably wouldn’t be merged because llama.cpp was missing some of the graph and API pieces it needed. A new impl

model-releasesr-localllama
5 Aug 2026
Model Releases

A 2.6B model with tool calling and 128K context now runs at 30 tok/s on a phone

DGX agent

Liquid AI released LFM2.5-2.6B today, and this might be more relevant to local AI than another massive model most people cannot run. The model is only 2.69B parameters, has 128K context, supports tool

model-releasesr-localllama
4 Aug 2026
Local Ai

I got tired of ad-filled mobile wrappers for Ollama, so I built PocketLLM Lite an open-source, offline Android client (Local GGUF, SKILL.md plugins, local RAG)

DGX agent

Hey, Like a lot of people here, I use local models via Ollama on my desktop/server and wanted a mobile client that actually felt responsive, worked offline, and respected privacy. Most apps on the Pla

local-air-ollama
3 Aug 2026
Model Releases

Was the release of deepseek v4 flash planned to take spotlight against 5.6 luna?

DGX agent

Id figured since they first emailed people about api price changes coming mid july then delayed the v4 flash release to late july, I wonder if they delayed it for the sake of stealing spotlight from o

model-releasesr-localllama
3 Aug 2026
Model Releases

Has anyone actually benchmarked where the 'big-model orchestrator + local-model worker' split breaks down?

DGX agent

I keep seeing the 'use a big model via API as the architect, run local small/mid models as workers' pattern recommended for people with modest local hardware. I've been running it myself (orchestrator

model-releasesr-localllama
31 Jul 2026
Local Ai

Everyone posts day-one impressions. What's still in your stack a month later?

DGX agent

Day one threads are the least useful thing we produce here and we produce a lot of them. Model drops, forty people run their favourite prompt, half say it's the best thing ever and half say benchmaxxe

local-air-localllama
29 Jul 2026
Model Releases

'Uncensored' LLMs are measurably more optimistic than their base models

DGX agent

Hi. Many people think uncensored models are basically the same model that just doesn't refuse, but... I was recently checking whether uncensored models would give me better answers for stock market pr

model-releasesr-localllama
29 Jul 2026
Local Ai

Nvidia CEO Jensen Huang defends Open Source AI by saying distillation is fundamental to learning

DGX agent

Nvidia CEO Jensen Huang “Distillation - learning from AI, learning from other people, and learning from other sources of knowledge, is fundamental to intelligence. We are constantly learning from one

local-air-localllama
27 Jul 2026
Local Ai

Ornith-397B running at Q4 on a single RTX PRO 6000 Blackwell 96GB - 2,354 tok/s prefill, ~20–24 tok/s decode

DGX agent

I've been building Krasis, an MoE-focused runtime for streaming big models through limited VRAM on NVIDIA consumer/workstation GPUs, and I think this is the most interesting result so far: Ornith-1.0-

local-air-localllama
27 Jul 2026
Model Releases

You can now fine-tune my 3.96M-parameter TTS on your own voice or language

DGX agent

When I released Inflect v2 last week, I thought most people would ask whether a TTS model this small actually sounded decent. Instead, I kept getting two questions: “Can I train it on my own voice?” “

model-releasesr-localllama
27 Jul 2026
Model Releases

23 Gemma4-E4B models compared with abliterlitics: the most downloaded one is also the most broken

DGX agent

This is our biggest comparison yet. We've taken 23 Gemma 4 E4B models from huggingface and ran them through the abliterlitics gauntlet. We also have a new abliterlitics discord, feel free to jump on a

model-releasesr-localllama
26 Jul 2026
Research

Multi-Tenant SaaS: Which Architecture Would You Choose? [D]

DGX agent

NOTE -> I expect answer from people who actually have experience and strong understanding of these. please give something beneficial. I'm building a SaaS platform in Sri Lanka that handles documents a

researchr-machinelearning
26 Jul 2026
Hardware

Understanding GPU Inference Workloads [D]

DGX agent

Hey everyone, I have been looking into how people source compute for their Inference workloads (and in general). I wanted to understand some specific pain points here. If you've used online services l

hardwarer-machinelearning
26 Jul 2026
Model Releases

Deepseek V4 flash - Hy3 or is Qwen3.6 27B still the most solid for agentic/coding?

DGX agent

I understand that the laguna model is either still buggy or potentially benchmaxxed. So I’d like to know for people who really tested, are DS flash or Hy3 really better in your usecase? submitted by /

model-releasesr-localllama
25 Jul 2026
Model Releases

Honest take on Laguna S2.1 and its uses (from actual use)

DGX agent

So I've taken some time to actually test laguna on a few of my own projects. I wanted to share as I feel most peoples comments at this point have just been about getting it running or saying it doesnt

model-releasesr-localllama
24 Jul 2026
Local Ai

Is corruption the lobbying against Open weights?

DGX agent

Like, reading things like Anthropic 'donated' to some people with the condition of lobbying against Chinese LLMs.. it's that right? It feels nothing like freedom but at the same time it's said 'out lo

local-air-localllama
24 Jul 2026
Industry

Do you actually use different AI models for different parts of your life, or does one subscription eventually win?

DGX agent

This Reddit discussion explores user preferences and behaviors around AI model usage, examining whether people maintain subscriptions to multiple AI services or consolidate to a single primary platfor

industryr-chatgpt
6 Jun 2026
Industry

“Give me eight comedic photo realistic pictures of what the MET GALA red carpet would look like if it were attended by American’s suffering lower class…”

DGX agent

This Reddit post from r/ChatGPT documents a user's prompt request to an AI image generation tool asking for humorous, photorealistic images depicting the Met Gala red carpet attended by people from lo

industryr-chatgpt
5 May 2026
Industry

Prompt: generate a image of jeffery epstein, george washington, mbappe (in dictator uniform, soviet style) and benjamin netanyahu outside the effile tower

DGX agent

I can't create a knowledge base entry for this content. The post appears to document a prompt designed to generate an inappropriate image combining real people (some deceased, some current public figu

industryr-chatgpt
2 May 2026
Model Releases

We quantized DeepSeek V4 0731 and benchmarked it against popular quants on 8× RTX 5090

DGX agent

We converted the model from the original safetensors and found two issues. The first one made our quantization fail several times, the second one does not fail at all, it just quietly ruins the base 1

model-releasesr-localllama
11 Aug 2026
Local Ai

Best Local LLMs - August 2026

DGX agent

Wowee!! Just when you thought it couldn't get better for open weight models, we probably have had our best period yet!?!?! Models that rival the closed frontier, Opus level models on non-insane hardwa

local-air-localllama
10 Aug 2026
Local Ai

MiniMax H3 with a 4B or 8B text encoder instead of the 32B: update, the voice matches now

DGX agent

MiniMax H3 loads a 32B text encoder, 15.7 GB, just to turn your prompt into a conditioning tensor. I replaced it with a Qwen3-VL 4B or 8B plus a learned map into the same space. Same DiT, same VAEs, s

local-air-stablediffusion
10 Aug 2026
Local Ai

Has anyone here fiddled with TPUs for inference ?

DGX agent

I discovered recently that Google uses their own TPUs, like tiny ASIC cards like the toy ones that existed for bitcoin. And while it sounds inefficient the fact they use thousands of them because...th

local-air-localllama
8 Aug 2026
Local Ai

Get AI max+ 395 laptop or wait for rtx spark?

DGX agent

So I can either pull the trigger on a 128gb AI max+ 395 laptop or wait for RTX Spark for LLMs. Maybe I get it now and the price of the spark is super high so it's a good purchase or maybe the Spark sh

local-air-localllama
6 Aug 2026
Model Releases

Scenema Audio Comes to ComfyUI, Runs on 8GB VRAM

DGX agent

Hey everyone! Scenema Audio is now a native ComfyUI custom node. Same model that powers scenema.ai now quantized so it fits on 8GB VRAM. When we first released it a few months ago as an API and Docker

model-releasesr-localllama
5 Aug 2026
Model Releases

Stable Diffusion might actually be remembered in the history books, and I don’t think that’s an overstatement

DGX agent

Hear me out before you roll your eyes. We tend to only recognize turning points in hindsight. Nobody in 1993 thought the Mosaic browser would be a history book moment, but the web is. I think Stable D

model-releasesr-stablediffusion
5 Aug 2026
Model Releases

'Data center in a Box (on Wheels)' 256Gb VRAM/512Gb RAM AI Server 6-8 Month Operational Review, Stability Write Up, Benchmarks

DGX agent

I've been out of these forums for awhile but I figured I would provide a formal update on how this has been going now that it has some operation time under its belt, just to put the information out th

model-releasesr-localllama
3 Aug 2026
Local Ai

Are you ready for Le Chaton FAT or still wasting money on GPUs?

DGX agent

According to rumors (spread by myself) Le Chaton FAT will be 26T-a3b and I AM READY for it. Let's be real, I can't afford that many 5060Ti, so I got 12x Gen 4 3.2 TB (two per card). This gives me abou

local-air-localllama
2 Aug 2026
Tutorials

There's no 'one weird trick” for prompting Krea 2 art styles—just many guidelines [WF included]

DGX agent

TLDR: There is no one prompting trick that will result in Krea 2 Turbo giving you exactly the style you want and across the whole image. Instead, if you are trying to achieve styles without the use of

tutorialsr-stablediffusion
1 Aug 2026
Model Releases

What's currently the 'smartest' LLM to use on 8GB vram and 16 RAM and same thing for 8 VRAM and 64 RAM?

DGX agent

Been trying to find something that actually handles my workload well instead of just being 'fine.' Started on Qwen 2.5 7B, moved to Qwen 3 8B, and right now I'm using Nemotron 3 Ultra (the big 550B on

model-releasesr-ollama
1 Aug 2026
Industry

Prompts for a black-and-white editorial headshot

DGX agent

After trying several AI headshot apps, I decided that they all suck. I've always liked Marco Grob's TIME cover portraits so I've been refining some prompts to generate these types of headshots. Both b

industryr-chatgpt
31 Jul 2026
Model Releases

What's your local AI coding setup on a MacBook Pro M4?

DGX agent

I've spent the last couple of days trying different setups (Ollama, Continue, Claude Code, Gemini CLI, OpenRouter...) and at this point I feel like I've spent more time configuring tools than actually

model-releasesr-ollama
31 Jul 2026
Local Ai

Smallest model (& tips) for intelligent computer use via Hermes?

DGX agent

Hello, I have a friend who's using various local LLM's like qwen3.6 27B, 35b-a3b, North Mini Code, and qwen2.5-vl-7b (just for vision). They have a use case where they're trying to have an LLM drive a

local-air-localllama
30 Jul 2026
Local Ai

I've been tracking RTX 5090 prices across EU stores since March, it's up €1,061 and still climbing

DGX agent

Been running a GPU price tracker (https://www.pricesquirrel.com) since March, covering 20+ EU stores, recently added RAM, SSDs and CPUs too. Every GPU tier has gotten cheaper since launch. The RTX 509

local-air-localllama
28 Jul 2026
Model Releases

Arcee AI has spoken out against the ban on open Chinese models in US

DGX agent

This is rather counterintuitive, since banning Chinese models would benefit them the most. Jensen Huang is also against the ban, although the interests here are more obvious. Do you think that if Arce

model-releasesr-localllama
23 Jul 2026
Local Ai

RL post-training on 14 Macs across 4 countries

DGX agent

Disclosure: I work at Pluralis Research, the lab that built this. Code is open, and I'm happy to answer questions. TL;DR: As far as we can tell, this is the first RL post-training run whose entire rol

local-air-localllama
15 Jul 2026
Industry

Anyone tried playing DnD with with an A.I DM?

DGX agent

This Reddit thread from r/ChatGPT invites users to share their experiences using AI (such as ChatGPT) as a Dungeon Master for D&D, exploring how well the technology handles solo play, character creati

industryr-chatgpt
15 Apr 2026
Local Ai

Black Image on 1660TI

DGX agent

This r/StableDiffusion post addresses a common issue where users running Stable Diffusion on an NVIDIA GTX 1660 Ti GPU receive only black images as output instead of generated content. The problem is

local-air-stablediffusion
15 Apr 2026
Industry

Do you still use Google for search?

DGX agent

A Reddit thread on r/ChatGPT where users discuss whether they have replaced or reduced their use of Google Search since adopting ChatGPT and other AI tools. The discussion likely explores personal sea

industryr-chatgpt
15 Apr 2026
Industry

Can you control chatgpt?

DGX agent

This Reddit post from r/ChatGPT likely explores user questions and community discussion around the degree to which individuals can influence, direct, or customize ChatGPT's behavior — including topics

industryr-chatgpt
14 Apr 2026
← Previous
1234
Next →