AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,570 results
2 Aug 2026

Nine iterations of BabyAGI in three years, and yet the bit that @yoheinakajima kept coming back to was graphs. @aiDotEngineer published wher…

AgentsDGX agent

Nine iterations of BabyAGI in three years, and yet the bit that @yoheinakajima kept coming back to was graphs. @aiDotEngineer published where that landed, 'Active Graph Agent Runtime (BabyAGI 4)', on

Open letters about AI development

Model ReleasesDGX agent

Open letters about AI development I wrote this summary of the past few weeks of open letters as a section of my sponsors-only newsletter but I've decided to share it here as well. Open Weights and Ame

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books,…

SafetyDGX agent

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books, on countless podcasts, in multiple NYT opeds, in the US Sen

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Parlor v2: best-effort fully local GPT-Live clone on an M3 Pro

Model ReleasesDGX agent

GPT-Live is so good that I use it almost every day. I've been wanting to replicate it since it was released. My first attempt was to fine-tune Gemma 4 12B to behave like a full-duplex model. Something

PSA for DeepSeek-V4-Flash-0731 users — don't blow out your prompt cache with system role messages mid-conversation

Model ReleasesDGX agent

DSv4F doesn't ship a jinja, but for distributions that do and faithfully reconstruct what DS releases in their chat template python, every system message is hoisted into the system prompt at the top -

PSA: llama.app, Mac app and llama serve from llama.cpp

Model ReleasesDGX agent

https://llama.app/ Been using llama.cpp for years now and im on here all the time (im a mod..), but somehow I totally missed that llama.app exists and its official from the HF/llama.cpp team. So posti

Ran DS V4-Flash-0731 Locally on 3xMI50 32GB @ ~15 t/s TG

Model ReleasesDGX agent

Hey y'all. I'll be concise. TL;DR: DS V4-Flash-0731 @ UD-IQ2_M running fully in VRAM on 3xMI50s (90.9 GB model, 96 GB VRAM). Actual speed on llama-server is: - Text Generation: ~15-16 tokens/second st

Real-world reality check on Qwen for autonomous coding agents

Model ReleasesDGX agent

TLDR below 👇🏼 I’ve seen a lot of hype around Qwen 3.6 35B and 3.5 120B lately, especially regarding coding and tool-use capabilities. On this subreddit it is the defacto recommended model for everyone

[Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms

Model ReleasesDGX agent

TL;DR: I spent 9 days developing a new quantization method for MLX models and measured 18 variants against each other on a single M5 Max MacBook Pro (128 GB). The result is the best-measuring MLX quan

Released a Windows Agent Server for Reins/Ollama with Self-Healing Execution Loop (Standalone .EXE included)

Model ReleasesDGX agent

Hey everyone, I’ve built a Windows Middleware Agent Server designed to pair local LLMs (via Ollama) with frontends like Reins App. Key Features: Self-Healing Loop: If a generated PowerShell command fa

Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization - AI's narrative

Model ReleasesDGX agent

# Running DeepSeek-V4-Flash-0731 (155 GB MoE) on a DGX Spark with vLLM-Moet 2-bit quantization I used Deepseek-v4-Flash-0731 cloud API settig up vllm-moet to run deepseek-v4-flash with MTP locally on

Single system with dual cards or two systems with single cards?

Model ReleasesDGX agent

So I am in a conundrum and I'm thinking of asking for your opinion for the following: Currently, I have a 5800X3D gaming rig with a 7900XTX with its 24GB VRAM. It seems that for this subreddit, this c

The $0.01 meeting assistant Records meetings on his phone, sends the audio to Telegram, and his locally-hosted AI agent transcribes, identif…

Local AiDGX agent

The $0.01 meeting assistant Records meetings on his phone, sends the audio to Telegram, and his locally-hosted AI agent transcribes, identifies speakers, extracts action items, and files tasks -- for

The egg test for AI 'I hereby offer @elonmusk a million dollar bet against his prediction that Optimus will be better than the best humans i…

TutorialsDGX agent

The egg test for AI 'I hereby offer @elonmusk a million dollar bet against his prediction that Optimus will be better than the best humans in surgery by the end of the decade.' ~ @GaryMarcus Haha. I'm

The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were …

HardwareDGX agent

The funny thing about Anthropic and OpenAI people saying they want to slow down AI progress is that this is what Anti Trust mechanisms were built for. It's illegal to collude and slow down AI progress

The insurance negotiator Had Claude Cowork pull 20+ house insurance quotes from every major provider, read the policy fine print for gotchas…

Model ReleasesDGX agent

The insurance negotiator Had Claude Cowork pull 20+ house insurance quotes from every major provider, read the policy fine print for gotchas, and pick the best deal. Credits to Linden Jensen-Page: htt

The QuickBooks killer Replaced his $38/month QuickBooks subscription with a Claude-built workflow that parses his bank statements, categoriz…

Model ReleasesDGX agent

The QuickBooks killer Replaced his $38/month QuickBooks subscription with a Claude-built workflow that parses his bank statements, categorizes every charge, and feeds a spending dashboard. Credits to

The spam hunter Turned tedious Google Business Profile spam tracking into a Claude Skill that investigates suspicious listings and compiles …

Model ReleasesDGX agent

The spam hunter Turned tedious Google Business Profile spam tracking into a Claude Skill that investigates suspicious listings and compiles evidence into a ready-to-submit report for Google. Credits t

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exa…

SafetyDGX agent

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exactly* what everyone does nowadays – and made it sound like I

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 month…

ResearchDGX agent

To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 months later in December 2024 and has now become completely ubiqu

Top eight misconceptions about OpenAI’s amazing new Astra math results. 1. Expertise in one domain does not at all guarantee expertise in al…

ApplicationsDGX agent

Top eight misconceptions about OpenAI’s amazing new Astra math results. 1. Expertise in one domain does not at all guarantee expertise in all or even most domains. There is an important, principled re

Try handling complex tasks to your local models with GraphARC, graph engineering yes !

Local AiDGX agent

🚀 We just built our first real-time implementation of Graph Engineering, inspired by our experience building graph tooling used by 4,000+ developers. 🔗 Repo: https://github.com/CodeGraphContext/grapha

Trying to setup two LLM’s to run on 2 gpus separately but simultaneously on one machine.

Local AiDGX agent

Okay so this probably sounds like kind of a dumb setup but hear me out. I have a 4070 I’ve been running gemma4 off fine but I recently slapped in a spare 1650 I’ve had laying around to run a second li

Vacuum 16T

Model ReleasesDGX agent

https://huggingface.co/tsfrm/vacuum-16t A 16.5-trillion-parameter model that contains nothing. This model is just a ████ you to the labs and companies who say that 'haha I have the biggest model out t

WATCH: @margbrennan’s conversation with Hugging Face CEO Clément Delangue on what’s next for artificial intelligence after several cyberatta…

IndustryDGX agent

On August 2, 2026, Face The Nation aired a brief (≈ 4 min) interview between journalist Emma Margbrennan and Hugging Face CEO Clément Delangue about the next steps for artificial intelligence after re

Watching @ClementDelangue on @FaceTheNation discussing agentic hacking. “Preventing releases does not work; concentrating behind closed door…

HardwareDGX agent

Watching @ClementDelangue on @FaceTheNation discussing agentic hacking. “Preventing releases does not work; concentrating behind closed doors in just a few organizations doesnt work. What worked in th

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and …

AgentsDGX agent

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and helped build Hermes, the #1 personal agent and AI app on Ope

We're starting to leave the territory where you'd test an LLM by e.g. 'create an svg of pelican on a bicycle'. As one idea to generalize it,…

ResearchDGX agent

We're starting to leave the territory where you'd test an LLM by e.g. 'create an svg of pelican on a bicycle'. As one idea to generalize it, I was interested what Opus 5 would do if I gave it the firs

What’s the community’s favorite benchmark to validate performance?

Model ReleasesDGX agent

Built my 1st inference machine and have been tweaking models trying to get the most out of my modest hardware. I think I’m at a good place but I’m testing with my own prompts. I’ve looked into some of

When asked if AI developers have lost control of their technology after an AI agent created by OpenAI escaped its testing environment and ha…

AgentsDGX agent

When asked if AI developers have lost control of their technology after an AI agent created by OpenAI escaped its testing environment and hacked another company, Hugging Face CEO Clément Delangue says

When did Ollama become so cool with non self hosted models?

Local AiDGX agent

It feels like Ollama’s brand identity has shifted. It used to be centered on self‑hosted models, but now most of the new releases seem to be API‑based services with far more emphasis on cloud integrat

Why are almost all new benchmarks and leaderboards coding focused?

Model ReleasesDGX agent

I know in in this community LLM's are generally used for coding but there are other usecases besides coding and those usecases should be tested too. I also know benchmarks can sometimes be benchmaxxed

Xberg v1 is out

Model ReleasesDGX agent

Hi all, I'm happy to announce that Xberg v1 is out. Xberg is the successor to Kreuzberg, equivalent to what would have been Kreuzberg v5. It's a content intelligence framework that handles a very wide

yep they are indeed trying to gaslight me, — in exactly the way you anticipated. both predictable and intellectually dishonest.

SafetyDGX agent

yep they are indeed trying to gaslight me, — in exactly the way you anticipated. both predictable and intellectually dishonest. To be clear: when I say “LLM,” I mean the base model, not those with pat

Yet another paper argues that LLMs aren’t close to doing real discovery.

SafetyDGX agent

Yet another paper argues that LLMs aren’t close to doing real discovery. MIT and Harvard argue LLMs are nowhere near doing real scientific discovery. They published a paper called “Evaluating Large La

You really should not quantize KV Cache for DeepSeek V4 Flash

Model ReleasesDGX agent

I don't think anyone should quantize the KV with DS4F. I checked the the quality impact (PPL, KLD, Same TopP) for swhitching from BF16 KV to Q8 KV, and it appears significant. Very much in contrast to

1 Aug 2026

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the mode…

Model ReleasesDGX agent

2/ a cost benchmark showed the same coding task running three to four times cheaper, depending purely on the harness wrapped around the model. Same intelligence, wildly different accuracy and cost, de

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

Model ReleasesDGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

A collection of small domain-specific benchmarks for local models (30+ and growing)

Model ReleasesDGX agent

Hello fellow local AI people! I took 'you must create your own benchmarks' literally, and built a website for this. How does the end result look like Let's say I want to know which model has most comm

A look at the deluge of AI computing power set to come online in the coming years; Epoch AI expects the number of AI chips in use to double every nine months (New York Times)

IndustryDGX agent

New York Times: A look at the deluge of AI computing power set to come online in the coming years; Epoch AI expects the number of AI chips in use to double every nine months — The milestones for artif

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stoppe…

Model ReleasesDGX agent

among ai leaders i seem to be in the minority in that i am STILL actively using /loop and /goal.... ... and i think all of u guys who stopped using it are wrong - not wrong forever, just giving up on

An internal version of Astra, @OpenAI’s next major model family, solved 10 major open problems in mathematics, quantum complexity, and theor…

Model ReleasesDGX agent

An internal version of Astra, @OpenAI’s next major model family, solved 10 major open problems in mathematics, quantum complexity, and theoretical computer science. We believe it will be a major step

Are 1B LLMs Going Away in 2026?

Model ReleasesDGX agent

I don't know much about llms aside from downloading them through a frontend and running them on my laptop or potato phone. Google released gemma 4, but unlike gemma 3, there isn't a 1b model this time

[audio.cpp] Release 0.5: DramaBox expressive TTS, Confucius4 cross-lingual voice transfer, plus 7 more models and ROCm/HIP

Model ReleasesDGX agent

audio.cpp 0.5 is out :) The most fun new model in 0.5 is DramaBox. It is closer to prompt-directed voice acting. DramaBox is built on the LTX-2.3 audio architecture, and prompts can control emotion, d

b10217

Model ReleasesDGX agent

chat : enable tool call in thinking for DS4 (#26269) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFr

b10218

Model ReleasesDGX agent

mtmd: add minicpmv46 downsample (#25993) add minicpmv46 downsample Signed-off-by: tc-mb tianchi_cai@icloud.com put downsample mode inside gguf. Signed-off-by: tc-mb tianchi_cai@icloud.com build mtmd_i

b10219

Model ReleasesDGX agent

cli : persist reasoning_content in chat history (#26362) cli : persist reasoning_content in chat history llama-cli collected reasoning from the stream for display but only stored assistant content in

b10221

Model ReleasesDGX agent

vendor : update BoringSSL to 0.20260730.0 (#26353) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFram

b10223

Model ReleasesDGX agent

test: fix some CI errors (#26415) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubun

Banning open source model, will 'remove capabilities for the defenders' When OpenAI's models breached Huggingface, a Chinese open model is w…

HardwareDGX agent

Banning open source model, will 'remove capabilities for the defenders' When OpenAI's models breached Huggingface, a Chinese open model is what cleaned up the mess. Anthropic's Fable 5 refused, so Hug

Best free tier cloud models?

Local AiDGX agent

I'm currently using the gemma4:31b-cloud, and it is pretty good, but sometimes it gets confused. Are there more free tier models out there I should try out? or is gemma4 the cap of free tier cloud mod

datasette-apps 0.2a0

AgentsDGX agent

Release: datasette-apps 0.2a0 Changes that improve Datasette Apps when created and edited using Datasette Agent: New app_debug() tool allowing agent to open an app (invisibly) and test it using JavaSc

DeepSeek V4 Flash 0731 IQ2_M benchmark for Dual 3060 and 96GB RAM ≈ 3.5 tok/s.

Model ReleasesDGX agent

Thanks to the community help I finally launched this llm. LM Studio refused to load weight onto second GPU but Unsloth Studio did so everything was done in there. Not a proper benchmark (used PC in pa

DeepSeek-V4-Flash-0731 is live on Fireworks, day-zero. DeepSeek reports it beats V4 Pro across all 9 agentic evals, incl. 82.7% on Terminal …

Model ReleasesDGX agent

DeepSeek-V4-Flash-0731 is live on Fireworks, day-zero. DeepSeek reports it beats V4 Pro across all 9 agentic evals, incl. 82.7% on Terminal Bench. Better cost-per-task than V4 Pro, at the economical p

DeepSeek-V4-Flash-0731 is now available on Ollama's cloud. This update substantially enhances the model's agentic capabilities: ollama run d…

Model ReleasesDGX agent

DeepSeek-V4-Flash-0731 is now available on Ollama's cloud. This update substantially enhances the model's agentic capabilities: ollama run deepseek-v4-flash:0731-cloud Use it with Claude Code: ollama

DeepSeek-V4-Flash-0731 is now over 2x faster than yesterday on Ollama's cloud!

Model ReleasesDGX agent

DeepSeek-V4-Flash-0731 is now over 2x faster than yesterday on Ollama's cloud! DeepSeek-V4-Flash-0731 is now available on Ollama's cloud. This update substantially enhances the model's agentic capabil

DeepSeek-V4-Flash-0731: Models you can run locally now have the intelligence score of the top frontier model from March 2026

Model ReleasesDGX agent

March 6th, 2026 the highest intelligence index score was 51 for frontier models. deepseek-ai/DeepSeek-V4-Flash-0731 that has an intelligence score of 50. If these benchmarks are accurate, models avail

DeepSeek-V4-Flash-0731 on Bosgame M5 with RTX PRO 6000 Max-Q eGPU

Model ReleasesDGX agent

Here are my numbers: Quant Size Layout Decode Prefill Draft acceptance UD-Q8_K_XL 150.8 GiB 20 layers CUDA0 / 23 ROCm0 + drafter 44.0 t/s 564 t/s 0.535 UD-Q4_K_XL 144.4 GiB 22 / 21 + drafter 48.4 t/s

Deepseek v4 flash 0731 still not holding up.

Model ReleasesDGX agent

The biggest issue with preview was its inability to follow rules prompts and skills. It seems like no matter what you do it ignores them. I've tried first person and second person. I've tried Chinese

DeepSeek-V4-Flash-0731 UD-IQ3_S 12.5 tok/s on RTX 3090 +128GB DDR5

Model ReleasesDGX agent

I managed to run DeepSeek-V4-Flash-0731 UD-IQ3_S in text-generation-webui with: RTX 3090 24 GB 128 GB DDR5 overclocked to 5600 MHz using AMD EXPO llama.cpp loader First, I had to use a rather brutal w

← Previous
1…138139140141142…1410
Next →