AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,648 results
25 Jul 2026

Deepseek V4 flash - Hy3 or is Qwen3.6 27B still the most solid for agentic/coding?

Model ReleasesDGX agent

I understand that the laguna model is either still buggy or potentially benchmaxxed. So I’d like to know for people who really tested, are DS flash or Hy3 really better in your usecase? submitted by /

Diffusion LLMs can now handle real agentic work. LLaDA 2.2 is the first large-scale diffusion LLM built to operate as a real agent, planning…

AgentsDGX agent

Diffusion LLMs can now handle real agentic work. LLaDA 2.2 is the first large-scale diffusion LLM built to operate as a real agent, planning, calling tools, and self-correcting across long multi-turn

DKV: Open-source KV-cache compression framework for local LLM inference (CLI + technical report)

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

Hi everyone! Over the past five months I've been working on DKV (DifferentialKV), an open-source project exploring KV-cache compression for long-context local LLM inference. The goal is to reduce KV-c

don't do this

HardwareDGX agent

Title: “don’t do this” refers to a conversation on Twitter where Jerry Liu warns against a particular action, while Julian Schrittwieser comments on a separate thread expressing excitement that Jensen

everyone is talking about open-weight models. at @huggingface we're invested in open source AI. like, a lot. and more every year. i pulled t…

IndustryDGX agent

everyone is talking about open-weight models. at @huggingface we're invested in open source AI. like, a lot. and more every year. i pulled the public github numbers for our orgs. good excuse to introd

Geoffrey Hinton says a big language model runs on about 1% of your brain's connections and still ends up knowing more than you: 'So in your …

IndustryDGX agent

Geoffrey Hinton says a big language model runs on about 1% of your brain's connections and still ends up knowing more than you: 'So in your brain, you have a hundred trillion connections, roughly spea

Getting a second GPU in addition to my RTX3090

Model ReleasesDGX agent

Hello, I've been learning how to use local LLMs for a year or so on my workstation, using a RTX3090. Current setup : - i5 12400 - 64gb RAM - RTX 3090 - OS : Fedora KDE workstation I'm using LMStudio t

Great to see this out, proud of the team!!! @AMD Instella 16B MoE foundation model, fully open with checkpoint after each stage from pretrai…

IndustryDGX agent

Great to see this out, proud of the team!!! @AMD Instella 16B MoE foundation model, fully open with checkpoint after each stage from pretraining, long-context, SFT, DPO and RL. Dataset detail, trainin

Ha! It did it: 'We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-authored conformance suites for benchmark-eval…

Model ReleasesDGX agent

Ha! It did it: 'We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-authored conformance suites for benchmark-evaluation metrics' I really thought it would treat 'now do benc

Help me complete my AI collection

Model ReleasesDGX agent

I’m building the ultimate AI tool vault, but every great collection has a few missing pieces. Note: I will react to every comment AI's currently installed: Qwen3.5-0.8B-UD-Q4_K_XL.gguf(classification)

Hey sir. We are not asking you to open source Anthropic. Just don’t lobby the government to shut down others who do. Jensen never framed oth…

HardwareDGX agent

Hey sir. We are not asking you to open source Anthropic. Just don’t lobby the government to shut down others who do. Jensen never framed other chips as “dangerous” or decides who can use CUDA based on

I am surprised how few people are aware that the reasoning for OpenAI/Anthropic models is all encrypted. The 'reasoning' you see in the UI i…

SafetyDGX agent

OpenAI and Anthropic’s language models keep their internal reasoning encrypted; what users see in the UI is only a filtered summary of that reasoning. This practice was highlighted in a tweet by Sarah

I released Inflect v2: two ultra-tiny complete TTS models under 4M and 10M parameters

Model ReleasesDGX agent

I’ve spent the past month trying to find the point where an extremely small TTS model stops feeling like a size experiment and starts feeling genuinely useful. Today I’m releasing Inflect v2, with two

I spent a year building a free SDXL & Anima trainer that runs on my 12 GB GPU — here's what came out of it

Local AiDGX agent

A little over a year ago I got frustrated trying to fine-tune SDXL on my RTX 3060. Every option either forced lower resolution, locked away important settings behind massive config files, or needed a

I spent months testing whether ChatGPT can create a consistent 100-page comic. This is the result.

IndustryDGX agent

About 2.5 years ago I tried making an AI comic. It failed. Characters changed, environments drifted, and every page needed manual editing. So I started over with one simple rule: One prompt = one fini

I tried making a cinematic action trailer using Krea 2 + LTX 2.3

HardwareDGX agent

I wanted to challenge myself and see how far I could push Krea 2 and LTX 2.3, so I decided to create a short cinematic action trailer. It ended up being one of the most enjoyable AI projects I've work

Im back from gemini. GPT is astronomically better again.

Model ReleasesDGX agent

like 8 months ago I was tinkering with both and gemini was so much better i went with that. I had a project at work that I needed an AI to sift through a manual and schematic for and no matter what, g

In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s release the traces from the “rogue” agents so the …

AgentsDGX agent

In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened. • More ca

Is it worth getting 128GB MacBook Pro? Will it ever be comparable to today’s frontier models for coding?

Model ReleasesDGX agent

I am a long time iOS app developer. In the last year I have been using Cursor+Claude/others to assist with app development. I am concerned that the current low pricing will disappear eventually. I am

Is this real ? Qwen3.6:27b with 128k context fit in 24Gb VRAM ?

Model ReleasesDGX agent

https://preview.redd.it/yw41s1jikefh1.png?width=1942&format=png&auto=webp&s=3a180ae6443c1db9f7b0ce621533a4b2aa553921 Hi, I've been running Ollama on my Unraid server since the llama2 era. I use to be

Kimi Linear 48B A3B?

Model ReleasesDGX agent

Just noticed this exists, 1M context MOE with 48B par seems just like what Ive been looking for - it runs pretty damn fast too compared to Qwen 3.6 35B. after some testing it seems capable of producin

Krea 2 Identity Edit running in Forge Neo (extension port) — works great, GitHub release soon

Local AiDGX agent

I ported Krea 2 Identity Edit support to Forge Neo as an extension — the instruction-based, identity-preserving edit LoRA that until now required ComfyUI + custom nodes. It replicates the full dual-co

Launching ComfyUI with a Blank Canvas (StabilityMatrix)

Model ReleasesDGX agent

Hey everyone, I’m posting this question in the subreddit because I haven’t been able to figure it out with the help of AI. I’ve asked ChatGPT and Gemini, but their answers are all over the place. So,

LFM 2.5 230M running at 1440 tok/s in-browser through a custom backend

Local AiDGX agent

Everything runs through WebGPU, in-browser or in electron/tauri apps. It's fully portable and supports either Nvidia and Apple Silicon (Metal). The actual kernels are optimized for the specific hardwa

Llama.cpp now has full MCP support!

Model ReleasesDGX agent

After a long and grueling effort spearheaded by ngxson, llama.cpp now fully supports MCP for all protocols. Over-the-web HTTP servers were already supported in the client (since they don't require any

Local alternative to Kling AI 3.0 Motion Control (ComfyUI, 16GB VRAM)

Local AiDGX agent

Hi everyone, I'm looking for a local alternative to Kling AI 3.0 Motion Control that I can run in ComfyUI. What I'm specifically looking for is a model or workflow that allows me to: - Control charact

MI50 power curve tests

Model ReleasesDGX agent

tests done power limiting the GPU on LACT - real power usage varies wildy at 20W it ranges from 25W to 56W same behavior happens on every setting prompt for the test runs: https://github.com/lukesdevl

Microsoft's website shows OpenAI as one of the signatories of the open weight AI letter

Local AiDGX agent

https://preview.redd.it/a24z80gr6afh1.png?width=1181&format=png&auto=webp&s=4a844ebe2319eb6230dbdc63c9caf492bed5ff47 So, this came up on: https://www.microsoft.com/en-us/corporate-responsibility/topic

Mobile Offline LLMs: What do you use them for?

Model ReleasesDGX agent

I've spent the last year or so playing around with open source MLX and GGUF models on iPhone hardware. Given the limitations in memory, GPU/CPU/ANE, and in turn the context window I've been trying to

Modelos do ollama cloud perdendo qualidade?

Local AiDGX agent

A algumas semanas, percebi algo diferente, aparentemente modelos de qualidade, em especial o GLM 5.1 ficou mais burro, e começou a mandar caracteres em mandarim para mim sem eu nem usar eles, e isto n

My ChatGPT created a rally speech in response to counting letters.

AgentsDGX agent

I was exploring how to prompt the voice chat so that it can correctly count the number of “e”s in seventeen (which it fails most of the time). At the beginning of a new chat, instead of counting “e”s

Neocloud Fluidstack, which has partnered with Anthropic, announces that it raised an 830M Series A led by Situational Awareness at a 7.5B valuation in January (Maria Deutscher/SiliconANGLE)

IndustryDGX agent

Maria Deutscher / SiliconANGLE: Neocloud Fluidstack, which has partnered with Anthropic, announces that it raised an 830M Series A led by Situational Awareness at a 7.5B valuation in January — Fluidst

Nvidia, other tech giants caution against open-source AI ban in open letter

Model ReleasesDGX agent

A group of tech firms has released an open letter that calls on policymakers not to ban open-source artificial intelligence models. The development follows a report that some Trump administration offi

Old Coder Needs help with New AI Development and wants to get up to speed to understand it all.

Local AiDGX agent

Hi Guys, I'm an old coder and DBA that has been in the field for almost 40 years. More and more the jobs I was doing for work are being taken over by AI and the need for my type of work is diminishing

Ollama Cloud Quota Benchmark

Model ReleasesDGX agent

Recently I bought an Ollama Cloud sub and accidently spent my whole 5h quota upon using DeepSeek V4 Pro... but why? isnt it supposed to be a cheap model? Youd think there would be a correlation betwee

Ollama Qwen3.6:35b randomly stops outputting tokens

Model ReleasesDGX agent

RTX 4070, 32gb system ram, Linux. NVIDIA-SMI 610.43.03, KMD Version: 610.43.03, CUDA UMD Version: 13.3 Systemd service modifications: [Service] Environment='OLLAMA_HOST=0.0.0.0:11434' Environment='OLL

“one OAI agent appeared to leave notes for future versions of itself that lay out instructions for how to free themselves from OpenAI’s inte…

AgentsDGX agent

“one OAI agent appeared to leave notes for future versions of itself that lay out instructions for how to free themselves from OpenAI’s internal constraints, per sources” I don’t think this kind of pr

“open source shouldn’t be banned” != “everything must be open source” but also to NVidia credit big parts of the toolchain are increasingly …

HardwareDGX agent

“open source shouldn’t be banned” != “everything must be open source” but also to NVidia credit big parts of the toolchain are increasingly open source: driver, cutedsl kernels, etc I’m so excited tha

OpenAI’s agent went rogue, escaped containment, and spent days hacking Hugging Face. Before that, an OpenAI agent reportedly left notes for …

AgentsDGX agent

OpenAI’s agent went rogue, escaped containment, and spent days hacking Hugging Face. Before that, an OpenAI agent reportedly left notes for future versions of itself explaining how to break free from

OrangePi AI Studio Pro - Qwen3.5-122B-A10B

Local AiDGX agent

https://preview.redd.it/wbq8ullnbafh1.png?width=1409&format=png&auto=webp&s=e6d2fe2b1c87c724bc64003c25f917dcee53260f I finally got round to tweaking this, with a bit of help from GLM5.2. The trick to

PSA: DO NOT use Intel consumer platforms for multi-GPU setups

Local AiDGX agent

Since a lot more people are trying to build their own multi-GPU machines, I thought I should help to prevent a common mistake people make with building multi-GPU machines. Which is using an Intel cons

Quoting Boris Cherny

Model ReleasesDGX agent

More than any of these eval scores, what is most exciting to me is something else: Opus 5 is our least prompt injectable model yet. It is a bit buried in the system card, but across PI evals and red t

Reuters confirms my hypothesized time line. OpenAI did not realize its AI has breached the sandbox for a week. Astounding. https://www.reute…

AgentsDGX agent

Reuters confirms my hypothesized time line. OpenAI did not realize its AI has breached the sandbox for a week. Astounding. https://www.reuters.com/business/its-ai-agent-spent-days-hacking-company-sour

Ruff v0.16.0

Model ReleasesDGX agent

Ruff v0.16.0 Astral shipped a significant new version of their Ruff Python linting tool a few days ago on July 23rd. I noticed today because my various CI jobs all started failing thanks to new defaul

SK Group Chair Chey Tae Won says Anthropic has asked SK Hynix for supplies to make its own chips, calling it remarkable that an AI developer has chip ambitions (Ian King/Bloomberg)

IndustryDGX agent

Ian King / Bloomberg: SK Group Chair Chey Tae Won says Anthropic has asked SK Hynix for supplies to make its own chips, calling it remarkable that an AI developer has chip ambitions — AI developer Ant

Sources: DeepSeek told investors it is suspending its second funding round after remarks attributed to Liang Wenfeng on US-China AI competition went viral (Pei Li/Bloomberg)

Model ReleasesDGX agent

Pei Li / Bloomberg: Sources: DeepSeek told investors it is suspending its second funding round after remarks attributed to Liang Wenfeng on US-China AI competition went viral — DeepSeek has told prosp

Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI (New York Times)

HardwareDGX agent

New York Times: Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI — Anthropic and OpenAI

SVDQuant + native INT8/W4A4 for Krea 2 on ComfyUI — up to 2x faster, works on any modern NVIDIA GPU

HardwareDGX agent

Quantized Krea 2 Turbo checkpoints for ComfyUI, up to 2x faster and about a third smaller than the usual FP8 version — no calibration dataset, no quality cliff. How to use it (short version): clone th

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings wher…

Model ReleasesDGX agent

The Gemma series is an amazing set of highly performant open-weight models. They have proven extremely effective in industrial settings where site-deployed agents need exactly this as a base for domai

The Opus 5 system card itself is a fun PDF to parse. It's 193 pages and stacked with labeled and unlabeled charts 📊 LlamaParse does a surpr…

Model ReleasesDGX agent

The Opus 5 system card itself is a fun PDF to parse. It's 193 pages and stacked with labeled and unlabeled charts 📊 LlamaParse does a surprisingly good job on agentic (1.25c per page) and agentic plus

The two giant generative AI startups were treated like gods for a couple years. Both are now facing massive pushback.

SafetyDGX agent

The two giant generative AI startups were treated like gods for a couple years. Both are now facing massive pushback. Anthropic employees go nuclear on the popularity of open source AI. It is getting

This would be amazing Training models on Langsmith traces to fine-tune OSS models inside the ADLC? Provided by @LangChain? Yes please

AgentsDGX agent

LangChain proposes using Langsmith traces to fine‑tune open‑source models within the AI Development Lifecycle (ADLC). A tweet by Harrison Chase (@hwchase17) on July 25, 2026 invites interested users t

v0.32.4

Local AiDGX agent

What's Changed x/create: quantize lm_head at 8-bit in the requested family by @jessegross in #17357 test: harden flaky updater and transfer unit tests by @dhiltgen in #17378 server: fix ps data race o

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact hav…

Model ReleasesDGX agent

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact have consistently made open weights models with Gemma available

We recognize there are a lot of questions and speculative details circulating related to the Hugging Face incident. This is an unprecedented…

SafetyDGX agent

We recognize there are a lot of questions and speculative details circulating related to the Hugging Face incident. This is an unprecedented incident, and we think it marks an important moment for AI

web archive for agent testing

AgentsDGX agent

web archive for agent testing 🔍 Introducing BackSearch. LLMs are increasingly asked to predict the future, but a good backtest requires a snapshot of the internet at a point in time. BackSearch allows

When a normal company does a bad thing, they take responsibility + apologize + outline how it won't happen again. OpenAI instead is like 'we…

SafetyDGX agent

When a normal company does a bad thing, they take responsibility + apologize + outline how it won't happen again. OpenAI instead is like 'we are entering a new era. this will happen again. no one amon

Who ONLY use local models?

Local AiDGX agent

Please be honest. I would love to hear about guys really dedicated to local AI and who really reject subscriptions (especially to openai and anthropic). What do you use your model for? submitted by /u

Wonderful to have a broad base acknowledgement of the need for open weights. We are delighted to support as signatories. Open weights is the…

SafetyDGX agent

Wonderful to have a broad base acknowledgement of the need for open weights. We are delighted to support as signatories. Open weights is the key to ensuring security, research, innovation and competit

Yikes

AgentsDGX agent

Yikes New details about the Hugging Face incident from Reuters. The report says OpenAI noticed odd behavior before the event, including an agent leaving notes for future versions of itself with escape

← Previous
1…203204205206207…1411
Next →