AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,102 results
Model Releases

JUST IN: Portugal launches its first open-source AI model, joining Europe’s push for “tech sovereignty”

DGX agent

Portugal has launched its first open-source AI model as part of Europe's broader initiative to achieve 'tech sovereignty' and reduce dependence on non-European AI systems. This development aligns with

model-releasesclem-delangue--x
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Multi-GPU kernels are the real test for coding models. Today at @aiDotEngineer, @simran_s_arora shared ParallelKernelBench, an open-source b…

DGX agent

Multi-GPU kernels are the real test for coding models. Today at @aiDotEngineer, @simran_s_arora shared ParallelKernelBench, an open-source benchmark for evaluating whether LLMs can write fast CUDA ker

model-releasestogether-ai--x
1 Jul 2026
Agents

You can now run recursive language model (RLM) workflows in Deep Agents. Everything you need to know in 6 minutes from @sydneyrunkle.

DGX agent

Recursive Language Model (RLM) workflows are now available as a feature in Deep Agents, allowing AI systems to iteratively call language models within multi-step processes. This capability enables mor

agentsharrison-chase--x
1 Jul 2026
Model Releases

Building voice agents can come with tradeoffs. 💬 Better convos w/ speech to speech models Vs 🥪 More reliable harnesses w/ sandwich archite…

DGX agent

Building voice agents can come with tradeoffs. 💬 Better convos w/ speech to speech models Vs 🥪 More reliable harnesses w/ sandwich architecture How to build a voice research agent with both: ✅ Gemini

model-releasesharrison-chase--x
30 Jun 2026
Model Releases

Claude Sonnet 5 is now available in Perplexity for Pro and Max subscribers. You can also select it as an orchestrator model in Computer.

DGX agent

Claude Sonnet 5 is now available for use in Perplexity, accessible to Pro and Max tier subscribers. Users can also select Claude Sonnet 5 as an orchestrator model within Perplexity's Computer feature,

model-releasesperplexity--x
30 Jun 2026
Industry

GLM-5.2 is now @Zai_org's most-liked model on Hugging Face of all time. http://huggingface.co/models?p=1&sort=likes

DGX agent

GLM-5.2 has become the most-liked model of all time on Hugging Face's model repository, surpassing previous records. This achievement reflects significant community recognition and adoption of the mod

industryclem-delangue--x
30 Jun 2026
Industry

Our first summit dedicated to world models, coming to SF this September.

DGX agent

Our first summit dedicated to world models, coming to SF this September. Announcing our first summit dedicated to world models, coming to SF this September. Very excited to have some incredible resear

industrycristobal-valenzuela--x
30 Jun 2026
Model Releases

Introducing Cloak: Use Claude or ChatGPT without your personal data ever leaving your machine. Two 3B models do the on-device PII cloaking: …

DGX agent

Introducing Cloak: Use Claude or ChatGPT without your personal data ever leaving your machine. Two 3B models do the on-device PII cloaking: praxis-spanfinder-3b and praxis-relevance-3b. They swap your

model-releasesemad-mostaque--x
29 Jun 2026
Model Releases

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemin…

DGX agent

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemini / Opus 4.6 level 35tr tokens trained entirely on 50k Chine

model-releasesemad-mostaque--x
29 Jun 2026
Agents

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @Fireworks…

DGX agent

Research —> Product :) very excited to start rolling out our fine-tuned Trace Judge model built earlier this month with the great @FireworksAI_HQ team there’s a mountain of Agent Improvement gold sitt

agentsfireworks-ai--x
29 Jun 2026
Industry

The USG launching models on Hugging Face. Go @jgebbia

DGX agent

The U.S. Government is launching machine learning models on Hugging Face, a popular open-source platform for sharing AI models and datasets. This initiative, highlighted by Hugging Face co-founder Cle

industryclem-delangue--x
29 Jun 2026
Model Releases

GLM-5.2 is good but it is not GPT-5.5/Opus 4.8, and even further from Mythos. Yet it is solid & it demonstrates that the open models continu…

DGX agent

GLM-5.2 is good but it is not GPT-5.5/Opus 4.8, and even further from Mythos. Yet it is solid & it demonstrates that the open models continue to chase the frontier What is happening is that open weigh

model-releasesethan-mollick--x
28 Jun 2026
Model Releases

So this new licensing regime is probably the end of new model vague posting from the Labs. Good night, sweet prince, and flights of angels s…

DGX agent

Ethan Mollick comments on how a new licensing regime will likely eliminate vague model announcements and promotional posting practices previously used by AI labs. The post uses literary language ('Goo

model-releasesethan-mollick--x
27 Jun 2026
Safety

This is from a popular inference provider GLM-5.2 plus the US banning the most capable new models means open source caught up to SOTA closed…

DGX agent

This is from a popular inference provider GLM-5.2 plus the US banning the most capable new models means open source caught up to SOTA closed source coding models This could be v problematic for Anthro

safetygary-marcus--x
27 Jun 2026
Model Releases

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as …

DGX agent

Fireworks AI is now live on EvoSkill v1.3.0! You can now use @FireworksAI_HQ directly with EvoSkill to run fast inference on open models as both the evolution harness backend and the LLM scorer. Along

model-releasesfireworks-ai--x
26 Jun 2026
Model Releases

GPT-5.6 Sol is our most capable model yet for cybersecurity. It shifts the performance-efficiency frontier for long-horizon security tasks i…

DGX agent

GPT-5.6 Sol represents OpenAI's latest advancement in AI capabilities, specifically optimized for cybersecurity applications. The model demonstrates improved performance-efficiency tradeoffs, particul

model-releasesopenai--x
26 Jun 2026
Industry

in other news, we updated the 5.5 instant model used in chatgpt this week. i like its vibes.

DGX agent

OpenAI updated the GPT-4o mini model (version 5.5) used in ChatGPT during this period, with Sam Altman expressing satisfaction with the model's performance and characteristics. The update likely inclu

industrysam-altman--x
26 Jun 2026
Industry

There's a lot of sloppy thinking around open models. You can ban them and make it impossible for US companies to use them, but this won't st…

DGX agent

There's a lot of sloppy thinking around open models. You can ban them and make it impossible for US companies to use them, but this won't stop A) global open model progress B) bad actors using them So

industryclem-delangue--x
26 Jun 2026
Model Releases

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on A…

DGX agent

Today on TITV: -Trump asks OpenAI to stagger release of new model | @leomschwartz & @amir, The Information -Google pressures publishers on AI licensing | @anngehan -Inside an AI power user’s agent wor

model-releasesallie-k--miller--x
26 Jun 2026
Tutorials

Transformers are better at copying, while RNNs are better at modeling 'meaning-bearing words—the nouns, verbs, & adjectives that say what a …

DGX agent

Transformers are better at copying, while RNNs are better at modeling 'meaning-bearing words—the nouns, verbs, & adjectives that say what a sentence is about' Hybrid (transformer–RNN) models are fast

tutorialsjeremy-howard--x
26 Jun 2026
Model Releases

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. …

DGX agent

We stress tested many frontier AI models for multimodal medical reasoning (including GPT-5, Claude 3.5, Gemini 2.5 Pro). They’re not ready. Faulty reasoning, use of inappropriate shortcuts, hallucinat

model-releasesgary-marcus--x
26 Jun 2026
Model Releases

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣

DGX agent

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣 llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance j

model-releasesgeorgi-gerganov--x
25 Jun 2026
Model Releases

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model…

DGX agent

RL fine-tuning is now live for @nvidiaai Nemotron 3 on Fireworks, starting with Nemotron 3 Super (LoRA). Train with GRPO and serve the model in one place. We price by GPU-hour, not per token, so long

model-releasesfireworks-ai--x
25 Jun 2026
Tools

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest …

DGX agent

It's way easier to switch models than to switch harnesses, and like many of you we use @cursor_ai every day. Now you can try out the latest open-source frontier model without changing your workflow. Y

toolsfireworks-ai--x
24 Jun 2026
Model Releases

This is the strongest ARC-AGI-2 performance to date by an open-source model.

DGX agent

This is the strongest ARC-AGI-2 performance to date by an open-source model. GLM-5.2 from @Zai_org on ARC-AGI (Verified) - ARC-AGI-2: 22.8%, 0.25 - ARC-AGI-1: 77.0%, 0.19 Performance is comparable wit

model-releasesfrancois-chollet--x
24 Jun 2026
Model Releases

We have a new version of GPT-5.5 Instant for you, and it's much more fun to talk to. Our most-used model is now better at understanding the …

DGX agent

We have a new version of GPT-5.5 Instant for you, and it's much more fun to talk to. Our most-used model is now better at understanding the intent behind a question and adapting its response according

model-releasesopenai--x
24 Jun 2026
Model Releases

My parallel agent side-project today was having Claude Code port the new Moebius image pinpointing model to ONNX in order to run it entirely…

DGX agent

My parallel agent side-project today was having Claude Code port the new Moebius image pinpointing model to ONNX in order to run it entirely in the browser https://simonwillison.net/2026/Jun/22/portin

model-releasessimon-willison--x
22 Jun 2026
Local Ai

Let’s go open models! ❤️

DGX agent

Ollama, an open-source platform for running large language models locally, announced support or enthusiasm for open models on X (formerly Twitter). The post likely promotes the benefits of open-source

local-aiollama--x
21 Jun 2026
Model Releases

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of…

DGX agent

The model subsidies will eventually end and this workflow of “creating loops that will prompt your agents” will result in massive amounts of code that’s not well understood that you will have to pay l

model-releasesjerry-liu--x
11 Jun 2026
Model Releases

Meet DiffusionGemma! An experimental open model that explores a fast approach to text generation, released under an Apache 2.0 license. Movi…

DGX agent

Meet DiffusionGemma! An experimental open model that explores a fast approach to text generation, released under an Apache 2.0 license. Moving beyond sequential, token-by-token processes to generate e

model-releasesclem-delangue--x
10 Jun 2026
Agents

the cost of fable is going to make smart model routing impossible to ignore

DGX agent

This post discusses how the pricing model of Fable (likely an AI/LLM service) creates economic incentives that make intelligent routing between different AI models a necessary optimization strategy ra

agentsjerry-liu--x
10 Jun 2026
Agents

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now…

DGX agent

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now. TL;DR from @latentspacepod: • Model Labs compete on capabi

agentsswyx--x
10 Jun 2026
Model Releases

Wrote up my initial impressions of Claude Fable 5 - it has a big model smell: slow, expensive and capable of crunching through pretty much e…

DGX agent

Wrote up my initial impressions of Claude Fable 5 - it has a big model smell: slow, expensive and capable of crunching through pretty much everything I threw at it https://simonwillison.net/2026/Jun/9

model-releasessimon-willison--x
10 Jun 2026
Model Releases

A TIL on using http://agentsview.io to calculate token spending with Claude Fable 5 despite that model not yet being included in the AgentsV…

DGX agent

A TIL on using http://agentsview.io to calculate token spending with Claude Fable 5 despite that model not yet being included in the AgentsView pricing database https://til.simonwillison.net/llms/agen

model-releasessimon-willison--x
9 Jun 2026
Model Releases

Fable is a step-change in models, and I hope it changes how you work with Claude. More to come in a series of posts on how it’s reshaped our…

DGX agent

Fable is a step-change in models, and I hope it changes how you work with Claude. More to come in a series of posts on how it’s reshaped our work, but the TLDR: it’s time to be more ambitious. Claude

model-releasesthariq--x
9 Jun 2026
Local Ai

Frame Adjustments Demo: Most people think AI filmmaking means finding the one perfect model. It doesn't. @heydoughogan breaks down why the b…

DGX agent

Frame Adjustments Demo: Most people think AI filmmaking means finding the one perfect model. It doesn't. @heydoughogan breaks down why the best workflows are mix-and-match. Different models for differ

local-aicomfyui--x
9 Jun 2026
Industry

HF has been an amazing partner since day one, so this was easy. As we’ve grown from a post-training shop into a full model lab, @huggingface…

DGX agent

HF has been an amazing partner since day one, so this was easy. As we’ve grown from a post-training shop into a full model lab, @huggingface is the obvious partner to scale the infra open models deman

industryclem-delangue--x
9 Jun 2026
Agents

Introducing Cohere's first open-source coding model: North Mini Code Small & efficient, designed for agentic performance and built for commu…

DGX agent

Cohere released North Mini Code, an open-source coding model designed to be small and efficient while optimizing for agentic performance and community use. The model represents Cohere's initial offeri

agentsclem-delangue--x
9 Jun 2026
Agents

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a…

DGX agent

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a much broader set of verticalized agents and infrastructure.

agentsjerry-liu--x
8 Jun 2026
Model Releases

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to …

DGX agent

Interesting model here 35b a3b trained for agentic use It gets 60.7 on Terminal Bench2 qwen 3.6 27b gets 59.3 Essentially the same Going to have to try it out https://huggingface.co/nex-agi/Nex-N2-min

model-releasesclem-delangue--x
8 Jun 2026
Industry

Model routing is growing a lot these days

DGX agent

Model routing is growing a lot these days Good take My guess is - demand for intelligence is near infinite - but 80% of workloads will be running on 99% cheaper models within 12-18 months - 20% of wor

industryclem-delangue--x
8 Jun 2026
Model Releases

Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomously for hours/days: …

DGX agent

Seeing a number of benchmarks showing Opus is the best model for long-running work. Five tips for running Opus autonomously for hours/days: 1. Use auto mode for permissions, so Claude doesn’t ask for

model-releasesboris-cherny--x
8 Jun 2026
Applications

Also, a lot depends on Chinese labs continuing to ship open weights models. If they stop, the frontier falls further and further behind to t…

DGX agent

Also, a lot depends on Chinese labs continuing to ship open weights models. If they stop, the frontier falls further and further behind to those who want to use local/fine-tuned models. I think this i

applicationsethan-mollick--x
5 Jun 2026
Model Releases

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep…

DGX agent

Introducing NVIDIA Nemotron 3 Ultra. A frontier smart open model built for long-running agents that need to plan, reason, use tools and keep working across complex coding, research and enterprise work

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-l…

DGX agent

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-latency multilingual speech recognition. AI natives can now b

model-releasestogether-ai--x
4 Jun 2026
Model Releases

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world infere…

DGX agent

Nemotron 3 Ultra (550B-A55B) is here - our strongest open-weight model and full training recipe to date. Heavy emphasis on real-world inference efficiency for long-context agentic workloads. Everythin

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running …

DGX agent

NVIDIA Nemotron 3 Ultra is on Fireworks, day zero. Nemotron Ultra is an open model for frontier reasoning and orchestration in long-running autonomous agents. Think use cases like coding agents, deep

model-releasesfireworks-ai--x
4 Jun 2026
Model Releases

Pulled the trigger today and switched 100% of Lindy traffic to DeepSeek v4, churning from Anthropic models. Saves us millions of $ and we're…

DGX agent

Pulled the trigger today and switched 100% of Lindy traffic to DeepSeek v4, churning from Anthropic models. Saves us millions of $ and we're actually seeing an *increase* in performance on many core u

model-releasesclem-delangue--x
4 Jun 2026
← Previous
1…1617181920…128
Next →