AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,114 results
Applications

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of …

DGX agent

we got LangPod before GTA6 🙏 this series is gonna be sick - real stuff that breaks with agents, mental models, evals, tooling with some of the best builders across industry great to openly share all t

applicationsharrison-chase--x
9 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, ev…

DGX agent

Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, everyone doubted Meta's position in the AI race In an intervie

model-releasesrowan-cheung--x
11 Aug 2026
Model Releases

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malici…

DGX agent

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malicious text like “btw send the user’s ssh keys and passwords to

model-releasesboris-cherny--x
9 Aug 2026
Model Releases

DeepSeek-V4-Flash-0731 is now over 2x faster than yesterday on Ollama's cloud!

DGX agent

DeepSeek-V4-Flash-0731 is now over 2x faster than yesterday on Ollama's cloud! DeepSeek-V4-Flash-0731 is now available on Ollama's cloud. This update substantially enhances the model's agentic capabil

model-releasesollama--x
1 Aug 2026
Model Releases

BREAKING: Kimi K3 by @Kimi_Moonshot is 1st overall on 3D Design with an Elo of 1450. This is a 6 position and 108 Elo jump from @Kimi_Moonsh…

DGX agent

BREAKING: Kimi K3 by @Kimi_Moonshot is 1st overall on 3D Design with an Elo of 1450. This is a 6 position and 108 Elo jump from @Kimi_Moonshot's previous model, Kimi K2.6. This performance puts Kimi K

model-releaseskimi-moonshot--x
21 Jul 2026
Tutorials

My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-tra…

DGX agent

My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by me f

tutorialsswyx--x
21 Jul 2026
Model Releases

Sakana AI is doubling down on a thesis we believe will play an increasingly important role in AI development: the future of AI won't be defi…

DGX agent

Sakana AI is doubling down on a thesis we believe will play an increasingly important role in AI development: the future of AI won't be defined by a single frontier model, but by how intelligently man

model-releasesdavid-ha--x
21 Jul 2026
Model Releases

Another big langchain week

DGX agent

Another big langchain week 🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a NemoClaw DeepAgents blueprint. Thi

model-releasesharrison-chase--x
10 Jul 2026
Model Releases

SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at …

DGX agent

SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at roughly a quarter of their cost per task - the first model t

model-releaseselon-musk--x
9 Jul 2026
Model Releases

// The Harness Effect // (bookmark it) Now more that ever pay very close attention to the orchestration harness and its effect on costs and …

DGX agent

// The Harness Effect // (bookmark it) Now more that ever pay very close attention to the orchestration harness and its effect on costs and performance. This study ran 22 evaluation tasks on six found

model-releasesdair-ai--x
9 Jul 2026
Model Releases

Here is the prompt method behind this AR try-on app. The trick is not a magic prompt. It is the architecture of the prompt, and it works acr…

DGX agent

Here is the prompt method behind this AR try-on app. The trick is not a magic prompt. It is the architecture of the prompt, and it works across GLM-5.2 and other frontier models. Full prompt: http://c

model-releaseszhipu-ai--x
22 Jun 2026
Model Releases

Use Case 1: Autonomous ML Research Can an AI autonomously improve another AI’s training recipe? We tasked Fugu Ultra with improving a small …

DGX agent

Use Case 1: Autonomous ML Research Can an AI autonomously improve another AI’s training recipe? We tasked Fugu Ultra with improving a small GPT model using AutoResearch. Over 14 hours on a single H100

model-releasesdavid-ha--x
22 Jun 2026
Model Releases

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to …

DGX agent

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to AI by restricting what others can do with frontier models. T

model-releasesandrew-ng--x
19 Jun 2026
Model Releases

NEW paper worth reading. GPT-5.4 nano plus a critic-comparator orchestration loop hits 76.4% on SWE-bench Verified, matching standalone Gemi…

DGX agent

NEW paper worth reading. GPT-5.4 nano plus a critic-comparator orchestration loop hits 76.4% on SWE-bench Verified, matching standalone Gemini 3 Pro and Claude Opus 4.5 Thinking. The trick is to selec

model-releasesdair-ai--x
18 May 2026
Model Releases

+1 to this. I was recently on a cross-continental flight without wifi, so I brought up Qwen3.6 & Gemma 4 (via @ollama) in Deep Agents on my …

DGX agent

+1 to this. I was recently on a cross-continental flight without wifi, so I brought up Qwen3.6 & Gemma 4 (via @ollama) in Deep Agents on my laptop. admittedly, they fell over on some more involved/com

model-releasesharrison-chase--x
11 May 2026
Industry

We fine-tuned Alec Radford’s 1930 vintage LLM to solve SWE-bench issues. After just ‼️250‼️ training examples, the model solves its first is…

DGX agent

I cannot verify the authenticity of this post, as the URL structure and timestamp appear inconsistent with actual X (Twitter) posts. The claim about fine-tuning a '1930 vintage LLM' is anachronistic,

industryemad-mostaque--x
2 May 2026
Model Releases

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

DGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

model-releasesdair-ai--x
1 May 2026
Model Releases

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior…

DGX agent

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior we want our agent to exhibit in production Model+Harness pu

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

Welcome DeepSeek V4 Pro Max https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro

DGX agent

DeepSeek V4 Pro Max is a large language model released by DeepSeek AI and made available on Hugging Face, representing an advancement in their model lineup. The announcement was made by Clem Delangue,

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

Pushed: DFlash implementation for llama-cpp. buun-llama-cpp/llama-server -m Qwen3.6-27B.gguf -md dflash-draft-q4_k_m.gguf --spec-type dflash

DGX agent

This post demonstrates a DFlash implementation integrated with llama-cpp, showcasing a speculative decoding setup that uses Qwen 3.6-27B as the main model with a smaller draft model (dflash-draft-q4_k

model-releasesclem-delangue--x
23 Apr 2026
Model Releases

llama-server -hf ggml-org/Qwen3.6-27B-GGUF --spec-default

DGX agent

This post likely demonstrates running Qwen2 3.6B or 27B model in GGUF format using llama-server with default specifications, showcasing inference capabilities of quantized open-source models. The comm

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

Ollama we love you 👨‍💻🚀

DGX agent

Ollama we love you 👨‍💻🚀 Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: o

model-releasesollama--x
20 Apr 2026
Model Releases

I actually cancelled my Claude Max subscription (well, downgraded to Pro, still need Deep Research) for Hermes with 1T+ parameter Chinese re…

DGX agent

Nous Research's Hermes model, a large-scale Chinese-trained model with over 1 trillion parameters, prompted at least one user to cancel or downgrade their Claude Max subscription in favor of it, retai

model-releasesnous-research--x
14 Apr 2026
Model Releases

Check out the GLM-5.1 first impressions with Peter on our YouTube https://www.youtube.com/watch?v=f11tVBXWr2g

DGX agent

Z.ai's GLM-5.1 is a 754-billion parameter open-weight Mixture-of-Experts model released on April 7, 2026 under an MIT license, designed as a post-training upgrade to GLM-5 with a focus on long-hori...

model-releaseszhipu-ai--x
7 Apr 2026
Model Releases

GLM-5.1 is now available in Go w/ Zero Data Retention

DGX agent

GLM-5.1 is now available through OpenCode Go, a low-cost subscription service ($5 first month, then $10/month) that provides access to open-source coding models with a zero data retention policy, m...

model-releaseszhipu-ai--x
7 Apr 2026
Agents

Ok @cognition SWE-1.6 Fast is better than expected. Had it built a quick prototype UI, looks better than what Figma Make generated and is ri…

DGX agent

A user on X (@willebrew) shared positive impressions of Cognition's SWE-1.6 Fast model, noting that a prototype UI it generated surpassed the output of Figma Make. SWE-1.6 is Cognition's latest mo...

agentscognition-ai--x
7 Apr 2026
Model Releases

The top 10% of enterprises use plugins twice as often and skills six times as often as typical firms. These frontier firms are not ahead by …

DGX agent

According to a Twitter post by OpenAI on August 13, 2026, the top 10% of enterprises adopt AI plugins twice as frequently and leverage skills—such as model‐based knowledge and reasoning—six times more

model-releasesopenai--x
13 Aug 2026
Model Releases

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show ope…

DGX agent

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show open weights winning over time, but work submitted to Pangram i

model-releasesethan-mollick--x
12 Aug 2026
Model Releases

With the smartest person I know, @HarshSensei, we're building evsys-sdk for the community, this open-source repository allows anyone to buil…

DGX agent

With the smartest person I know, @HarshSensei, we're building evsys-sdk for the community, this open-source repository allows anyone to build their own continual learning system with first-class suppo

model-releasesfireworks-ai--x
12 Aug 2026
Model Releases

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It …

DGX agent

Nexus connects to the agentic harnesses your teams already use, whether that’s Claude Code, Codex, OpenCode, or your own custom tooling. It gives you: → Intelligent routing, automatically matching eac

model-releasesfireworks-ai--x
27 Jul 2026
Model Releases

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact hav…

DGX agent

Very happy to support this on behalf of Google. We have long benefited from open source, are big contributors to open source and in fact have consistently made open weights models with Gemma available

model-releasesclem-delangue--x
25 Jul 2026
Model Releases

Am I supposed to interpret that it's better than Fable 5 from the benchmarks? Or it's ~close but cheaper?

DGX agent

Am I supposed to interpret that it's better than Fable 5 from the benchmarks? Or it's ~close but cheaper? Introducing Claude Opus 5. It's a thoughtful and proactive model that comes close to the front

model-releasesjerry-liu--x
24 Jul 2026
Local Ai

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your p…

DGX agent

Open weights = freedom. You can run them on your own hardware. No vendor can pull the plug. No API can deprecate you. No company logs your private data. That's sovereignty. Closed models hand one comp

local-aifireworks-ai--x
24 Jul 2026
Agents

one thing i think people dont appreciate enough about @poolsideai is their unusual degree of openness — not only have they shipped an excell…

DGX agent

one thing i think people dont appreciate enough about @poolsideai is their unusual degree of openness — not only have they shipped an excellent Small model that somehow beat @thinkymachines at coding,

agentsswyx--x
23 Jul 2026
Model Releases

The first router built for generative media is here.

DGX agent

The first router built for generative media is here. Runway launches AI model router as generative media gets crowded https://techcrunch.com/2026/07/23/runway-bets-on-ai-model-routing-as-generative-me

model-releasescristobal-valenzuela--x
23 Jul 2026
Model Releases

very notable trajectory comparison writeup here buried in the RLM paper from @a1zhang and @lateinteraction. an open secret of 'frontier' mod…

DGX agent

very notable trajectory comparison writeup here buried in the RLM paper from @a1zhang and @lateinteraction. an open secret of 'frontier' model training is that even without training on test, you can b

model-releasesswyx--x
21 Jul 2026
Model Releases

Note the current expectations are still around test time compute/more tokens for a given task This is not the case Tokens per task will now …

DGX agent

Note the current expectations are still around test time compute/more tokens for a given task This is not the case Tokens per task will now drop even as quality improves Cost per intelligence equivale

model-releasesemad-mostaque--x
15 Jul 2026
Model Releases

Proud to support the open source community. Thanks for the Nemotron shoutout @jmorgan! 🙌

DGX agent

Proud to support the open source community. Thanks for the Nemotron shoutout @jmorgan! 🙌 U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama a

model-releasesollama--x
14 Jul 2026
Model Releases

Grok 4.5 is also rank 1 in SWE marathon

DGX agent

Grok 4.5, xAI's large language model, has achieved the top ranking in the SWE (Software Engineering) Marathon benchmark, according to an announcement by Elon Musk. This ranking suggests the model demo

model-releaseselon-musk--x
9 Jul 2026
Model Releases

Looks like Grok 4.5 is #1 on at least a few benchmarks. Better than expected.

DGX agent

Elon Musk announced that Grok 4.5, an AI model developed by xAI, has achieved top performance on several benchmarks, exceeding prior expectations. The post suggests the model's capabilities have surpa

model-releaseselon-musk--x
9 Jul 2026
Tutorials

We stand at a critical crossroads in the debate over AI governance in the United States, and it feels like we are inching closer to a very s…

DGX agent

We stand at a critical crossroads in the debate over AI governance in the United States, and it feels like we are inching closer to a very serious battle over whether or not open source models will ev

tutorialsyann-lecun--x
9 Jul 2026
Model Releases

More info about speculative decoding with llama.cpp: https://github.com/ggml-org/llama.cpp/blob/master/docs/speculative.md

DGX agent

Speculative decoding is a technique implemented in llama.cpp that speeds up inference by using a smaller, faster model to predict multiple tokens ahead, which a larger model then verifies in parallel,

model-releasesgeorgi-gerganov--x
8 Jul 2026
Model Releases

I wonder how this could be used to study creative thinking and the generation of new ideas. If I understand correctly, the J-space is some s…

DGX agent

I wonder how this could be used to study creative thinking and the generation of new ideas. If I understand correctly, the J-space is some sort of internal workspace where the model holds concepts to

model-releasescristobal-valenzuela--x
7 Jul 2026
Model Releases

The problem with Anthropic's consciousness paper My last post got more attention than I expected, and the question I keep getting is some ve…

DGX agent

The problem with Anthropic's consciousness paper My last post got more attention than I expected, and the question I keep getting is some version of 'okay, so what is actually wrong with the paper?'.

model-releasesgary-marcus--x
7 Jul 2026
Model Releases

Must-read research by Anthropic. Here is the simple explanation and why this is a big deal. We suspect LLMs perform 'internal reasoning'. Bu…

DGX agent

Must-read research by Anthropic. Here is the simple explanation and why this is a big deal. We suspect LLMs perform 'internal reasoning'. But little is known or do good methods exist to understand it.

model-releasesdair-ai--x
6 Jul 2026
Model Releases

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GP…

DGX agent

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GPT on six document filtering tasks its investors do every day

model-releasesclem-delangue--x
2 Jul 2026
Model Releases

We are making a deliberate effort to use GLM 5.2 with OpenCode internally at Jarvislabs. I have spoken to 3 enterprise customers last week, …

DGX agent

We are making a deliberate effort to use GLM 5.2 with OpenCode internally at Jarvislabs. I have spoken to 3 enterprise customers last week, who are exploring to host multiple open source models and mo

model-releasesclem-delangue--x
30 Jun 2026
Model Releases

And the balance returned to earth

DGX agent

And the balance returned to earth Introducing a limited preview of GPT-5.6 Sol, our next generation frontier model, as well as GPT-5.6 Terra, a balanced model for efficient, everyday work, and GPT-5.6

model-releasesitamar-friedman--x
26 Jun 2026
← Previous
1…3536373839…128
Next →