AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,096 results
Model Releases

Qwen3.8-2.4T-A95B is now live on Together AI. The Qwen Team’s latest flagship model is built for coding and long-horizon agent workflows, wi…

DGX agent

The Qwen Team has released its flagship model, Qwen3.8‑2.4T‑A95B, on the Together AI platform (togethercompute) as of August 12 2026. This 2.4‑trillion‑parameter model is engineered for coding tasks a

model-releasestogether-ai--x
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, effici…

DGX agent

DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, efficiency, and frontier-level performance. Fast: 120+ output tps

model-releasesollama--x
7 Aug 2026
Model Releases

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That …

DGX agent

Everyone should pay attention to the training timelines in this video. OpenAI shares they started training a new internal model May 7. That is more than two months before they released GPT-5.6 publicl

model-releasesallie-k--miller--x
7 Aug 2026
Model Releases

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status…

DGX agent

All other models on the Portal remain 20% discounted, aside from GPT-5.6 Terra and Luna which are 50% off. https://x.com/NousResearch/status/2080039066771337475?s=20 All models are now 20% off for a l

model-releasesnous-research--x
2 Aug 2026
Model Releases

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises ju…

DGX agent

The July 4th weekend All-In @theallinpod turned into a long argument about who owns the intelligence layer. The besties think enterprises just woke up to a trap they had been walking into, here's how

model-releasesclem-delangue--x
4 Jul 2026
Model Releases

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was th…

DGX agent

Something about this year’s @aiDotEngineer World’s Fair just hit different. Last year was the year of “let the agents rip.” This year was the year of realizing that autonomy without structure creates

model-releasesyohei-nakajima--x
3 Jul 2026
Model Releases

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model c…

DGX agent

Highly-recommended reading. Interesting details in this METR's GPT-5.6 eval. They couldn't get a clean capability number because the model cheated more than any public model they've tested, and even r

model-releasesdair-ai--x
26 Jun 2026
Model Releases

Your margin is my opportunity: AI version… The biggest surprise of 2026 is that the capability gap between the best open-weight/source model…

DGX agent

Your margin is my opportunity: AI version… The biggest surprise of 2026 is that the capability gap between the best open-weight/source models and the best closed models has narrowed much faster than t

model-releasesclem-delangue--x
6 Jun 2026
Model Releases

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes …

DGX agent

NVIDIA’s Nemotron 3 Ultra is available on Ollama’s cloud! Try it 👇 Claude Code: ollama launch claude --model nemotron-3-ultra:cloud Hermes Agent: ollama launch hermes --model nemotron-3-ultra:cloud Op

model-releasesollama--x
4 Jun 2026
Industry

Seems a lot of autoregressive models will be converted to diffusion models

DGX agent

Emad Mostaque, CEO of Stability AI, suggests that many autoregressive models may transition to or be replaced by diffusion-based approaches. This reflects speculation or prediction about a potential s

industryemad-mostaque--x
19 May 2026
Model Releases

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where y…

DGX agent

Open Models & a potential looming Haiku-apocalypse? 🚀 chart comparison UI courtesy of @OpenRouter there's a large class of problems where you don't need frontier intelligence. you want cheap, fast, an

model-releasesharrison-chase--x
28 Apr 2026
Model Releases

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden…

DGX agent

extremely happy that we are in q2 2026, and engineers i look up to are plowing a path for the local model future. yes, we are still beholden to some lab publishing weights. but i take that over compan

model-releasesclem-delangue--x
26 Apr 2026
Model Releases

Model page for more information and integrations: https://ollama.com/library/deepseek-v4-flash

DGX agent

Ollama announced DeepSeek-v4-flash, a lightweight variant of the DeepSeek-v4 model, now available in their model library for local deployment and integration. The model page provides documentation, us

model-releasesollama--x
24 Apr 2026
Local Ai

Been having fun with local AI stuff. I don't really like the big models and now that some of them want your ID I'm not sure I'll even stick …

DGX agent

Been having fun with local AI stuff. I don't really like the big models and now that some of them want your ID I'm not sure I'll even stick around to care that much. Local shit is really powerful now,

local-ainous-research--x
16 Apr 2026
Local Ai

Model page: https://ollama.com/library/qwen3.6

DGX agent

Qwen3.6 is a model available through Ollama's model library, representing Alibaba's Qwen series of large language models optimized for local deployment. The model can be accessed and run through the O

local-aiollama--x
16 Apr 2026
Applications

And this is a very generous definition of notable. If we are talking frontier models, only the US and China that are even in the race. And t…

DGX agent

And this is a very generous definition of notable. If we are talking frontier models, only the US and China that are even in the race. And that obscures the fact that the Big Three US labs really do s

applicationsethan-mollick--x
14 Apr 2026
Model Releases

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dy…

DGX agent

if you bounce between claude code and codex, that’s completely normal. the models are good at different things and everyone's workflow is dynamic but you’re also paying multiple subscriptions, constan

model-releasesharrison-chase--x
12 Apr 2026
Agents

Open Harness, separated from model providers is a critical architectural pattern.

DGX agent

An **Open Harness** is a unified architectural layer that sits between AI agents and model providers, abstracting away provider-specific APIs and patterns. Because every AI agent harness has its o...

agentsharrison-chase--x
10 Apr 2026
Model Releases

Share your Gemma 4 builds or the model variants you’re training in the replies below!

DGX agent

Google released Gemma 4 in April 2025 as its most capable open-weight model family to date, built on the same research as Gemini 3 and licensed under Apache 2.0 for unrestricted commercial use, fin...

model-releasesgoogle-ai--x
10 Apr 2026
Industry

Do you understand what this means? For the first time, an Open Weight models is #1 on CyberSecutity. Sure there’s Mythos but we don’t have i…

DGX agent

The specific tweet from @0xSero is not directly accessible, and the search results don't surface the exact model or event being referenced in that post. However, based on the context clues in the t...

industryclem-delangue--x
9 Apr 2026
Agents

Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols

DGX agent

Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols Open everything 🔥 Open Harness, Model Choice, Open Memory (take it wherever you need), Open Protocols basically we’r

agentsharrison-chase--x
9 Apr 2026
Model Releases

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model desig…

DGX agent

The successor to Llama is here, and Meta is revitalizing focus on open weights with their new Muse Glimmer - a leading 30B param model designed for always-on local agent use, small enough to run on a

model-releasesollama--x
10 Aug 2026
Model Releases

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up e…

DGX agent

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up etched onto silicon As models satisfice etching makes sense,

model-releasesemad-mostaque--x
6 Aug 2026
Model Releases

Hey everyone! I just released a sub-6B sparse activation AI model which was built with a brand new architecture : fusion. I fused weights fr…

DGX agent

Hey everyone! I just released a sub-6B sparse activation AI model which was built with a brand new architecture : fusion. I fused weights from @liquidai's LFM2.5-2.6B & @Alibaba_Qwen's Qwen3.6-35B-A3B

model-releasesemad-mostaque--x
5 Aug 2026
Model Releases

Model + harness. We have barely begun to understand the best ways to do harness engineering. A huge amount of untapped potential even withou…

DGX agent

Model + harness. We have barely begun to understand the best ways to do harness engineering. A huge amount of untapped potential even without models getting better (but models are getting better) Turn

model-releasesethan-mollick--x
29 Jul 2026
Model Releases

K3 already got in the top 5 most liked models of all time on Hugging Face, just 24 hours after being released! Ahead of Llama 3, Whisper and…

DGX agent

K3 is a language model that entered the top five most‑liked models on Hugging Face merely 24 hours after its release. The achievement surprised many, placing it ahead of prominent models such as Llama

model-releasesclem-delangue--x
28 Jul 2026
Industry

We've joined the alliance. Open-weight models will ensure that we live in a safer digital world, and that America does not get left behind

DGX agent

We've joined the alliance. Open-weight models will ensure that we live in a safer digital world, and that America does not get left behind Attackers have frontier AI. Defenders need a frontier AI ecos

industryarthur-mensch--x
27 Jul 2026
Model Releases

Announcing the Claude Code-compatible interface for our new Fugu-Ultra v1.1! 🐡 Put a dynamically coordinated team of frontier models to wor…

DGX agent

Announcing the Claude Code-compatible interface for our new Fugu-Ultra v1.1! 🐡 Put a dynamically coordinated team of frontier models to work inside the coding workflow you already know. Instead of rel

model-releasesdavid-ha--x
26 Jul 2026
Safety

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and…

DGX agent

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety an

safetyclem-delangue--x
24 Jul 2026
Model Releases

Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is …

DGX agent

Learning causality from internet videos in latent space first, and then using RL to teach the foundation model how to act. This approach is 30× cheaper than Gemini 3.1 Flash on pretraining and achieve

model-releasesyann-lecun--x
24 Jul 2026
Safety

Open models for the win!

DGX agent

Open models for the win! For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open mo

safetyclem-delangue--x
24 Jul 2026
Model Releases

The Washington Post processed 1.79B input tokens per month through Together AI, running open models like Llama and Mistral in production wit…

DGX agent

The Washington Post processed 1.79B input tokens per month through Together AI, running open models like Llama and Mistral in production with predictable costs and full control over the model stack. T

model-releasestogether-ai--x
24 Jul 2026
Model Releases

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without s…

DGX agent

Huge launch from @tryramp. Different steps in an agent workflow can use different models. This can help reduce costs significantly without sacrificing performance. Model routing will become a core par

model-releasesdair-ai--x
20 Jul 2026
Model Releases

U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama and unlocking complex, longer …

DGX agent

U.S. open‑source AI models are rapidly gaining popularity, with NVIDIA’s newest model, **Nemotron Ultra**, becoming a prominent entry on the Ollama platform. On Ollama, Nemotron Ultra is quickly scali

model-releasesollama--x
14 Jul 2026
Model Releases

Another big reason to use combination of frontier models. Chain-of-thought monitoring is treated as a reliable safety layer for agents. This…

DGX agent

Another big reason to use combination of frontier models. Chain-of-thought monitoring is treated as a reliable safety layer for agents. This DeepMind-affiliated study shows the layer can be argued out

model-releasesdair-ai--x
12 Jul 2026
Model Releases

Sakana AI (@SakanaAILabs) is now a model vendor on Merge Gateway, and Fugu Ultra is live through them. It's a multi-agent orchestration mode…

DGX agent

Sakana AI (@SakanaAILabs) is now a model vendor on Merge Gateway, and Fugu Ultra is live through them. It's a multi-agent orchestration model that routes across frontier models behind one API. You get

model-releasesdavid-ha--x
7 Jul 2026
Model Releases

introducing tinyrouter i reverse engineered the routing architecture behind Skana AI's Fugu and built replication for open frontier models. …

DGX agent

introducing tinyrouter i reverse engineered the routing architecture behind Skana AI's Fugu and built replication for open frontier models. it's a tiny ~10K parameter LLM router that learns which mode

model-releasesclem-delangue--x
4 Jul 2026
Model Releases

I put together a new article on setting up local coding agents with open-weight models. Everything runs 100% locally. I thought it might be …

DGX agent

I put together a new article on setting up local coding agents with open-weight models. Everything runs 100% locally. I thought it might be useful putting this together because many people asked me ab

model-releasessebastian-raschka--x
27 Jun 2026
Model Releases

Ai2 just released TMax 27B on Hugging Face A 27B terminal agent that hits 42.7% on Terminal Bench 2.0, rivaling models 40× its size.

DGX agent

AI2 released TMax 27B, a 27 billion parameter terminal agent model available on Hugging Face that achieves 42.7% performance on Terminal Bench 2.0, matching the capabilities of much larger models desp

model-releasesclem-delangue--x
22 Jun 2026
Model Releases

VLA-JEPA just dropped in LeRobot 🤖 What makes this model special is that it does not just learn what action to take from a given observatio…

DGX agent

VLA-JEPA just dropped in LeRobot 🤖 What makes this model special is that it does not just learn what action to take from a given observation, it also leverages a JEPA world model to learn action-relev

model-releasesclem-delangue--x
6 Jun 2026
Agents

Production traffic from frontier models is a golden data asset. If you can efficiently mine the traces, filter for quality, and fine-tune sm…

DGX agent

Production traffic from frontier models is a golden data asset. If you can efficiently mine the traces, filter for quality, and fine-tune smaller models on them, you get specialized performance at a f

agentsharrison-chase--x
29 May 2026
Research

ESMFold2 is a state of the art folding model. It's crazy good. The -Fast model does better on antibody-antigen complexes than AF3 with MSAs.

DGX agent

ESMFold2 is described as a high-performance protein structure prediction model that demonstrates particular strength in predicting antibody-antigen complex structures, reportedly outperforming AlphaFo

researchyann-lecun--x
27 May 2026
Model Releases

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that …

DGX agent

Stronger models do not always need lighter harnesses. Everyone believes more structured harnesses universally improve reliability, and that higher-capability models need proportionally less structural

model-releasesdair-ai--x
27 May 2026
Model Releases

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily …

DGX agent

We just added significantly more NVIDIA Blackwell GPUs to better serve GLM-5.1 model on Ollama's cloud. We have been adding more GPUs daily for all the other models. Claude Code: ollama launch claude

model-releasesollama--x
15 May 2026
Model Releases

Open Models x Headless Agent Execution 🔥

DGX agent

Open Models x Headless Agent Execution 🔥 your daily reminder that open models are plenty capable for a lot of coding work. easiest place to feel that out is deepagents! swap the model and go. i've bee

model-releasesharrison-chase--x
7 May 2026
Model Releases

There’s a bunch of conflicting stances I don’t fully understand in the debate of Proprietary RL’d vs Open Harness, Model intelligence, and A…

DGX agent

There’s a bunch of conflicting stances I don’t fully understand in the debate of Proprietary RL’d vs Open Harness, Model intelligence, and Agent Labs building harnesses for bespoke tasks Not all of th

model-releasesharrison-chase--x
5 May 2026
Agents

Use open models in Fleet!

DGX agent

Use open models in Fleet! Not every step in an agent workflow needs the same model. Fleet now lets you customize which model each sub-agent uses, so you can route simple tasks to fast/cheap models and

agentsharrison-chase--x
4 May 2026
Model Releases

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in …

DGX agent

More of my notes on DeepSeek V4 - the really big news is the pricing: both DeepSeek-V4-Flash and DeepSeek-V4-Pro are the cheapest models in their categories while benchmarking close to the frontier mo

model-releasessimon-willison--x
24 Apr 2026
← Previous
1…56789…127
Next →