AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
Human
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
1,931 results
Model Releases

Ollama Cloud Quota Benchmark

DGX agent

Recently I bought an Ollama Cloud sub and accidently spent my whole 5h quota upon using DeepSeek V4 Pro... but why? isnt it supposed to be a cheap model? Youd think there would be a correlation betwee

model-releasesr-ollama
25 Jul 2026
Model Releases

Ollama Qwen3.6:35b randomly stops outputting tokens

DGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RTX 4070, 32gb system ram, Linux. NVIDIA-SMI 610.43.03, KMD Version: 610.43.03, CUDA UMD Version: 13.3 Systemd service modifications: [Service] Environment='OLLAMA_HOST=0.0.0.0:11434' Environment='OLL

model-releasesr-localllama
25 Jul 2026
Local Ai

OrangePi AI Studio Pro - Qwen3.5-122B-A10B

DGX agent

https://preview.redd.it/wbq8ullnbafh1.png?width=1409&format=png&auto=webp&s=e6d2fe2b1c87c724bc64003c25f917dcee53260f I finally got round to tweaking this, with a bit of help from GLM5.2. The trick to

local-air-localllama
25 Jul 2026
Local Ai

PSA: DO NOT use Intel consumer platforms for multi-GPU setups

DGX agent

Since a lot more people are trying to build their own multi-GPU machines, I thought I should help to prevent a common mistake people make with building multi-GPU machines. Which is using an Intel cons

local-air-localllama
25 Jul 2026
Hardware

SVDQuant + native INT8/W4A4 for Krea 2 on ComfyUI — up to 2x faster, works on any modern NVIDIA GPU

DGX agent

Quantized Krea 2 Turbo checkpoints for ComfyUI, up to 2x faster and about a third smaller than the usual FP8 version — no calibration dataset, no quality cliff. How to use it (short version): clone th

hardwarer-stablediffusion
25 Jul 2026
Local Ai

Who ONLY use local models?

DGX agent

Please be honest. I would love to hear about guys really dedicated to local AI and who really reject subscriptions (especially to openai and anthropic). What do you use your model for? submitted by /u

local-air-localllama
25 Jul 2026
Model Releases

[audio.cpp] Release 0.4: Higgs Audio v3 TTS 4B (10x real time)+ Fish Audio S2 Pro in C++/GGML, full GGUF loading, Q8 speed and VRAM gains

DGX agent

audio.cpp again :) Release 0.4 is out. The headline this time is new high-quality TTS coverage plus GGUF becoming a first-class across the project. What’s new: Added Higgs Audio v3 TTS 4B, Fish Audio

model-releasesr-localllama
24 Jul 2026
Local Ai

[BIG DATASET RELEASE] - SupraLabs/reasoning-corpus-4K-5M-v1 - Train your tiny SLMs to think!

DGX agent

https://preview.redd.it/b7ybs7nqx5fh1.png?width=3440&format=png&auto=webp&s=e6aaaa15cbe59debaae1ebb7fcd708167e86dc35 Hey r/LocalLLaMA ! We are back and we have something really amazing today. Our big

local-air-localllama
24 Jul 2026
Model Releases

CachyLLama’s: llama.cpp fork with persistent KV cache that makes long local-agent sessions much less painful

DGX agent

I’m not affiliated with this project, but I’ve been running it recently and I’m surprised it hasn’t received more attention here: https://github.com/fewtarius/CachyLLama CachyLLama is a fork of llama.

model-releasesr-localllama
24 Jul 2026
Model Releases

Can LLMs solve mazes?

DGX agent

https://reddit.com/link/1v5rvuq/video/bgmwc754i9fh1/player My goal was to create a benchmark to measure the spatial awareness and memory of models. Eventually, I came up with the simple idea of a maze

model-releasesr-localllama
24 Jul 2026
Local Ai

Does Ollama Cloud prompt caching even work?

DGX agent

I launched a new session and tasked GLM to create an implementation plan for a spec. The plan was on the bigger side, about 5k lines. I started with 0% 5h used and ended with 80% used. A few more twea

local-air-ollama
24 Jul 2026
Local Ai

Explain love in one sentence...

DGX agent

Love is an active commitment of deep affection where you find genuine joy in prioritizing someone else’s well-being as deeply as or beyond your own needs. VS Love is the deep, enduring connection betw

local-air-ollama
24 Jul 2026
Model Releases

Extened garlic to run Qwen3.5 35B A3B float8 at 55 tok/s on RTX 5060 Ti

DGX agent

In a previous post (https://www.reddit.com/r/LocalLLaMA/comments/1utefpr/running_qwen3_30b_a3b_at_50_toks_on_rtx_5060_ti/) there seemed to be great demand for bringing in Qwen3.5 35B. Some Gated Delta

model-releasesr-localllama
24 Jul 2026
Safety

Fizgig Krea 2 training features update

DGX agent

https://github.com/shootthesound/Fizgig Intelligent trainer - Per-image loss tracking with self-adapting training runs — every image gets its own verdict (easy / suspect / stuck / exhausted) and its o

safetyr-stablediffusion
24 Jul 2026
Local Ai

FLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence

DGX agent

Introducing FLUX 3. One multi-modal model for Image, Video, Audio and Action-Prediction. Creations are truer to life in every kind of style. Blog Post : https://bfl.ai/blog/flux-3 submitted by /u/pmtt

local-air-localllama
24 Jul 2026
Tutorials

For the first time, ChatGPT asked me a question instead of writing a wall of text

DGX agent

I copy-pasted this cooking recipe and accidentally hit enter before adding the actual prompt. Normally this would result in ChatGPT interpreting the text, trying to guess what I need, and giving a wal

tutorialsr-chatgpt
24 Jul 2026
Model Releases

Getting the most out of MTP

DGX agent

If you want to get the most out of MTP. You have to run some tests / benchmarks to do so. Turning it on with defaults will get improvements, but for many models and card combinations, you are leaving

model-releasesr-localllama
24 Jul 2026
Model Releases

Honest take on Laguna S2.1 and its uses (from actual use)

DGX agent

So I've taken some time to actually test laguna on a few of my own projects. I wanted to share as I feel most peoples comments at this point have just been about getting it running or saying it doesnt

model-releasesr-localllama
24 Jul 2026
Local Ai

Hugging Face releases The Stack v3 – largest open code dataset yet

DGX agent

From Anton Lozhkov on 𝕏: https://x.com/anton_lozhkov/status/2080254608639701222 Two ways in: stack-v3-train - near-deduplicated, quality-filtered, PII-redacted, contents inline. Point load_dataset at

local-air-localllama
24 Jul 2026
Model Releases

I built a compiler that turns computation graphs into the weights of a vanilla transformer — no training anywhere [P]

DGX agent

I've been chasing the question of what algorithms a transformer can actually express -- separate from what it can learn. So I built a compiler: define a computation graph in ordinary Python, and it pr

model-releasesr-machinelearning
24 Jul 2026
Model Releases

I built an open-source multi-agent SDLC harness that beats a cold Claude Code run on large repos, by learning the repo once. Real benchmarks (incl. where it loses) inside. [P]

DGX agent

Built an open-source AI coding agent that was 7%–75% cheaper than a cold 'claude -p' run on 6/6 well-localized tasks across repositories up to ~82k LOC. The biggest difference: Cold agent: 6.83, 207 t

model-releasesr-machinelearning
24 Jul 2026
Local Ai

Is corruption the lobbying against Open weights?

DGX agent

Like, reading things like Anthropic 'donated' to some people with the condition of lobbying against Chinese LLMs.. it's that right? It feels nothing like freedom but at the same time it's said 'out lo

local-air-localllama
24 Jul 2026
Local Ai

Is everyone training a single model together, based on the principle of the Tor network?

DGX agent

I just had a thought while scrolling. No idea if this already exists. What if users trained an AI model together—a bit like the Tor network or Bitcoin mining back in the day? - Participants download a

local-air-ollama
24 Jul 2026
Local Ai

It appears that the anti opensource AI lobby is far outgunned already

DGX agent

The earlier post on this subreddit by 20+ companies signing the petition including Microsoft, Meta, Nvidia, YC (https://www.microsoft.com/en-us/corporate-responsibility/topics/open-weight/) etc plus t

local-air-localllama
24 Jul 2026
Local Ai

More than 20 companies including NVIDIA, Meta, Microsoft, Palantir, and Hugging Face have signed a letter urging policymakers to avoid premature restrictions on open weight models.

DGX agent

The Open Letter was initiated by Microsoft and published today: “Open Weights and American AI Leadership”. It argues against broad or premature restrictions on open-weight models and explicitly says p

local-air-localllama
24 Jul 2026
Local Ai

No sé nada de Ollama, ni programación ni idea, pero estoy creando un agente evolutivo

DGX agent

Con ayuda de ChatGPT y con el modelo de Ollama, Qwen3:14b estoy creando un agente que corre local y tiene la iniciativa para pensar, investigar, aprender, generar propuestas y esperar mi autorización

local-air-ollama
24 Jul 2026
Model Releases

Nvidia releases Qwen-Image-Flash

DGX agent

'The NVIDIA Qwen-Image-Flash model generates images from text prompts using a four-step, DMD2-distilled version of Qwen/Qwen-Image. The distillation used DMD2 from NVIDIA FastGen, NVIDIA Model Optimiz

model-releasesr-stablediffusion
24 Jul 2026
Model Releases

Open Source Tax Engine outperforming fable 5 and gpt sol

DGX agent

This is an open source and free tax engine which scored 96% on TaxCalcBench [highest ever recorded score till date] surpassing fable 5 and sol with just sonnet 5. The only 2 cases where it missed, it

model-releasesr-ollama
24 Jul 2026
Model Releases

Optimizing an Ollama (Qwen:2.5) AI Agent: Fixing Search Aggregation, Context Bleed, and Query Extraction

DGX agent

I am building a domain-specific AI agent powered by Ollama (using the qwen:2.5 model). For data retrieval, the agent utilizes multiple search APIs: DuckDuckGo Search (DDGS), Tavily, Serper, and Google

model-releasesr-ollama
24 Jul 2026
Model Releases

[Paper] Statistically-Lossless Quantization of Large Language Models

DGX agent

Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as GPTQ and AWQ achieve practical compression but

model-releasesr-localllama
24 Jul 2026
Model Releases

People are using Minecraft farms as AI agent benchmarks

DGX agent

Someone modelled sugarcane farming as an integer program. See, sugarcane only grows next to water. Water costs one tile and can feed at most four cane tiles. The layout therefore becomes a coverage pr

model-releasesr-chatgpt
24 Jul 2026
Local Ai

Quick Demo of the new Auto-Control feature in my Open-Source App that monitors stuff on your screen using local LLMs, so you don't have to :))

DGX agent

TLDR: This is a demo of my open-source app which now auto-controls itself so you can monitor your downloads, renders, progress bars, or whatever's on your screen and camera :) Hey r/ollama !! I'm deve

local-air-ollama
24 Jul 2026
Local Ai

Spent two weeks on a kernel that benchmarked 29x faster. End to end it's maybe 6-10%, and it's not even wired in yet.

DGX agent

I've been building a C99 inference engine from scratch (no Python, no BLAS, just gcc and make) that runs BitNet's ternary models on CPU. A few weeks ago I got obsessed with the matmul kernel - wrote a

local-air-localllama
24 Jul 2026
Model Releases

swiss-ai/Apertus-v1.5 70B/8B

DGX agent

https://huggingface.co/swiss-ai/Apertus-v1.5-70B https://huggingface.co/swiss-ai/Apertus-v1.5-8B Apertus 1.5 is a family of 8B and 70B parameter language models designed to advance the state of multil

model-releasesr-localllama
24 Jul 2026
Tutorials

The 'distillation' claim is just ridiculous in nature

DGX agent

Even if China was distilling from US models (assuming all accusations are true), nothing about it makes it illegal. It is like saying you distilled knowledge from your professor in colleges and now he

tutorialsr-localllama
24 Jul 2026
Local Ai

What I learned using Ollama on a real Paperless archive: model choice was not the main problem

DGX agent

I maintain Tagvico, an open-source companion for Paperless-ngx. I added Ollama because document text is exactly the kind of data many people do not want to send to a hosted model. The surprising failu

local-air-ollama
24 Jul 2026
Industry

What if AI had access to classified files it was never allowed to quote but can make image?

DGX agent

What if an AI had seen fragments of classified material it could never describe directly? No files. No report names. No official explanations. Just images. That was the concept behind this series. I a

industryr-chatgpt
24 Jul 2026
Local Ai

What's the last model trained on human-data only?

DGX agent

From my understanding, most current LLMs are trained on trillions and trillions of tokens of mostly AI-generated data. Are there any recent models that are trained purely (or as close as possible) on

local-air-localllama
24 Jul 2026
Local Ai

which model to use on local 24gb mac mini M4 pro

DGX agent

So, i have been building some apps that should run on the local every user system, tried gemma4 although its fast and great at reasoning its not as good in instructions following and tool calling. tri

local-air-ollama
24 Jul 2026
Applications

Who has set up ChatGPT Finance?

DGX agent

It’s been out for Plus members for a little while now. Has anyone here connected it to their accounts? I myself haven’t done it, even though it’s only read access I’m seriously hesitant to hand over t

applicationsr-chatgpt
24 Jul 2026
Safety

Will there be Flux 3 Klein?

DGX agent

https://bfl.ai/blog/flux-3 “Over the next few weeks and months, we will make the following capabilities available, each after an early access phase for ensuring smooth rollout, collecting feedback and

safetyr-stablediffusion
24 Jul 2026
Model Releases

Zagreus-0.4B-por a small open source language model for Portuguese

DGX agent

mii-llm, an open source AI lab, released Zagreus-0.4B-por, a compact bilingual Portuguese–English language model pretrained entirely from scratch. The model has approximately 400 million parameters an

model-releasesr-localllama
24 Jul 2026
Local Ai

A caveman qwen3.6 27B

DGX agent

Just saw this on huggingface: https://huggingface.co/ProCreations/grug-27b The benchmarks claim that it's quite a bit better than qwen3.6 27B original and that they reduced the amount of necessary tok

local-air-localllama
23 Jul 2026
Local Ai

A local-first harness for multi-agent workflows

DGX agent

Hey all, I’ve been working on this in my spare time and finally feel ready to share it outside my own circles. Arbiter is a single binary for running agents locally. I originally built it because I wa

local-air-ollama
23 Jul 2026
Local Ai

Absurd claim: the distilled model outperforms the originals

DGX agent

As an AI community of LLM experts, are we really going to stay silent while US officials make absurd claims to push anti-consumer laws? Not only does the release timeline between Fable and K3 make hig

local-air-localllama
23 Jul 2026
Model Releases

AI9Stars released G9v3-3B

DGX agent

AI9Stars has released G9v3-3B an open weights language model designed to deliver strong reasoning capabilities within a lightweight 3 billion parameter size. It is released under the Apache 2.0 licens

model-releasesr-localllama
23 Jul 2026
Model Releases

Apple M5 isn't making full use of its matmul cores yet

DGX agent

At the moment MLX (and Llama.cpp for Macs) run 16bit activations everywhere. Despite this, the M5 generation silicon actually does support INT8 activations - it actually allows w4a8 d_type. It's just

model-releasesr-localllama
23 Jul 2026
Model Releases

Arcee AI has spoken out against the ban on open Chinese models in US

DGX agent

This is rather counterintuitive, since banning Chinese models would benefit them the most. Jensen Huang is also against the ban, although the interests here are more obvious. Do you think that if Arce

model-releasesr-localllama
23 Jul 2026
← Previous
1…1112131415…41
Next →