AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,102 results
Model Releases

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year?

DGX agent

The latest crop of models remains below 1% on ARC-AGI-3 -- for now. Where will the scores be by the end of the year? GPT-5.5 & Opus 4.7 on ARC-AGI-3 - GPT-5.5: 0.43% - Opus 4.7: 0.18% We found 3 failu

model-releasesfrancois-chollet--x
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please…

DGX agent

The new Grok comes in below the latest Chinese open weights models, Grok 4 was at the frontier when released. (& Artificial Analysis: please stop using GDPval-AA which is not a useful test of anything

model-releasesethan-mollick--x
1 May 2026
Model Releases

Claude Code is tuned for Claude. Codex is tuned for OpenAI models. Until now, 𝚍𝚎𝚎𝚙𝚊𝚐𝚎𝚗𝚝𝚜 had fixed harness defaults, which meant i…

DGX agent

Claude Code is tuned for Claude. Codex is tuned for OpenAI models. Until now, 𝚍𝚎𝚎𝚙𝚊𝚐𝚎𝚗𝚝𝚜 had fixed harness defaults, which meant it couldn't take advantage of the provider-specific optimizations that

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability require…

DGX agent

As models, contexts, and workloads grow, hidden assumptions in inference infrastructure can surface as output anomalies. Reliability requires more than throughput, latency, and availability. It also r

model-releaseszhipu-ai--x
29 Apr 2026
Agents

Today we’re shipping Laguna M.1 and Laguna XS.2 – our first public models. We’re also shipping our agent harness and a preview product exper…

DGX agent

Today we’re shipping Laguna M.1 and Laguna XS.2 – our first public models. We’re also shipping our agent harness and a preview product experience. Both models were trained from scratch on our own stac

agentsclem-delangue--x
28 Apr 2026
Model Releases

going to buy 2 rtx 6000 just because of the capabilities of local models becoming great! no more outsourcing of research to the api

DGX agent

going to buy 2 rtx 6000 just because of the capabilities of local models becoming great! no more outsourcing of research to the api Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, suppor

model-releasesclem-delangue--x
27 Apr 2026
Industry

Top 3 trending models of the week on HF: @deepseek_ai @OpenAI & @Alibaba_Qwen!

DGX agent

This post highlights the three most popular models on Hugging Face during a given week, featuring DeepSeek AI, OpenAI, and Alibaba's Qwen models. The post was shared by Clem Delangue, CEO of Hugging F

industryclem-delangue--x
27 Apr 2026
Industry

New Dflash drafting model for the 27b Lets gooooo https://huggingface.co/z-lab/Qwen3.6-27B-DFlash

DGX agent

A new Dflash drafting model based on Qwen 3.6 with 27 billion parameters has been released on Hugging Face, available at z-lab/Qwen3.6-27B-DFlash. This model likely implements speculative decoding or

industryclem-delangue--x
25 Apr 2026
Model Releases

500+ likes in 28 mins. On their way to be the fastest model ever to get to #1 trending on HF! https://huggingface.co/deepseek-ai/DeepSeek-V4…

DGX agent

DeepSeek-V4 rapidly gained over 500 likes within 28 minutes on Hugging Face, demonstrating exceptional user engagement and positioning it as a strong contender to become the fastest model to reach #1

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

🎉 Day-0 support for @deepseek_ai V4 Pro and Flash on vLLM — a new generation of DeepSeek model, purpose-built for tasks up to 1M tokens. Al…

DGX agent

🎉 Day-0 support for @deepseek_ai V4 Pro and Flash on vLLM — a new generation of DeepSeek model, purpose-built for tasks up to 1M tokens. Alongside the release, we're publishing a first-principles walk

model-releasesdylan-patel--x
24 Apr 2026
Model Releases

GPT-5.5 is now available on Perplexity for Max subscribers. GPT-5.5 is also rolling out as the default orchestration model in Computer for b…

DGX agent

Perplexity has made GPT-5.5 available to Max subscribers and is rolling it out as the default orchestration model in Perplexity Computer. The announcement indicates expanded access to OpenAI's GPT-5.5

model-releasesperplexity--x
24 Apr 2026
Model Releases

🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, and Qwen3.6-27B punches way above i…

DGX agent

🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, and Qwen3.6-27B punches way above its weight. 👇 What's new: 🧠 Outstanding agentic coding — surpa

model-releasesjeremy-howard--x
22 Apr 2026
Model Releases

Not the first time either - they shut down a bunch of of their original proprietary hosted embedding models in this announcement back in Apr…

DGX agent

Not the first time either - they shut down a bunch of of their original proprietary hosted embedding models in this announcement back in April 2024 https://openai.com/index/gpt-4-api-general-availabil

model-releasessimon-willison--x
22 Apr 2026
Industry

Qwen3.6-35B-A3B is trending at #1 on Hugging Face! 🥇🤗 Thank you for making us the top trending model on @huggingface this week. Let's keep…

DGX agent

Qwen3.6-35B-A3B has achieved #1 trending status on Hugging Face, indicating significant community interest and adoption of this large language model. The announcement highlights the model's popularity

industryclem-delangue--x
22 Apr 2026
Model Releases

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17

DGX agent

We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17 People are misreading the SpaceX/Cursor deal as an M&A story. It’s actually a b

model-releasesclem-delangue--x
22 Apr 2026
Industry

Anthropic's Mythos has been accessed by a small group of unauthorized users, raising questions about control of the AI model https://www.blo…

DGX agent

Anthropic's Mythos has been accessed by a small group of unauthorized users, raising questions about control of the AI model https://www.bloomberg.com/news/articles/2026-04-21/anthropic-s-mythos-model

industryclem-delangue--x
21 Apr 2026
Model Releases

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3…

DGX agent

🚀 Introducing Qwen3.6-Max-Preview, an early preview of our next flagship model Highlights: ⚡️ Improved agentic coding capability over Qwen3.6-Plus 📖 Stronger world knowledge and instruction following

model-releasesqwen--x
20 Apr 2026
Model Releases

RIP Anthropic API fees. Someone just made Claude Code run 100% locally on a MacBook for $0/month. 122B parameter model. 65 tokens per second…

DGX agent

RIP Anthropic API fees. Someone just made Claude Code run 100% locally on a MacBook for $0/month. 122B parameter model. 65 tokens per second. Nothing touches the cloud. The trick everyone else missed:

model-releasesclem-delangue--x
18 Apr 2026
Tutorials

Wow I can already say after just 5 hours using @AnthropicAI Opus 4.7 that this is the first model that 'gets' what I'm doing when I'm workin…

DGX agent

Wow I can already say after just 5 hours using @AnthropicAI Opus 4.7 that this is the first model that 'gets' what I'm doing when I'm working. It feels aligned with me in a way no previous model did.

tutorialsjeremy-howard--x
17 Apr 2026
Model Releases

180 tok/s generation on a 4090 with qwen 3.6. if you're on a 4090 and not running this model yet you're leaving performance on the table. 3B…

DGX agent

180 tok/s generation on a 4090 with qwen 3.6. if you're on a 4090 and not running this model yet you're leaving performance on the table. 3B active params at that speed is insane for agentic coding. t

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

Opus 4.7 is a model I’ve loved working with in Claude Code. It’s more agentic and instruction following but also incredibly smart and creati…

DGX agent

Opus 4.7 is a model I’ve loved working with in Claude Code. It’s more agentic and instruction following but also incredibly smart and creative. I think it takes a slight adjustment to get used to, but

model-releasesthariq--x
16 Apr 2026
Model Releases

The example prompt for Google's new Gemini Flash TTS text-to-speed model is a lot https://simonwillison.net/2026/Apr/15/gemini-31-flash-tts/

DGX agent

Google's Gemini 3.1 Flash TTS (text-to-speech) model includes a notably elaborate or extensive example prompt, which Simon Willison highlighted as noteworthy. The post likely comments on the complexit

model-releasessimon-willison--x
15 Apr 2026
Model Releases

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. This launch [excitement] includes aud…

DGX agent

Today we launched Gemini 3.1 Flash TTS, our most expressive and controllable text-to-speech model yet. This launch [excitement] includes audio tags! 🗣🏷 Audio tags [explanatory] are a seamless way to g

model-releasesgoogle-ai--x
15 Apr 2026
Tools

What if you could get 1.3B Transformer quality from a 770M model? That's not a compression result. It's a different architecture. Parcae, fr…

DGX agent

What if you could get 1.3B Transformer quality from a 770M model? That's not a compression result. It's a different architecture. Parcae, from @realDanFu (Together AI's VP of Kernels) and his lab at U

toolstogether-ai--x
15 Apr 2026
Industry

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call na…

DGX agent

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call nano-analyzer Models down to 3.6B active params (including ope

industryclem-delangue--x
14 Apr 2026
Model Releases

We conducted cyber evaluations of Claude Mythos Preview and found that it is the first model to complete an AISI cyber range end-to-end. 🧵

DGX agent

Boris Cherny and colleagues conducted cybersecurity evaluations of Claude Mythos Preview, finding it to be the first AI model to successfully complete an AISI (AI Safety Institute) cyber range end-to-

model-releasesboris-cherny--x
13 Apr 2026
Local Ai

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduc…

DGX agent

@hwchase17 I think harness/managed agents is a way for Anthropic to keep its Moat. As models get mature, the need for cloud LLMs might reduce and local LLM models might increase (save cost) and thus t

local-aiharrison-chase--x
12 Apr 2026
Agents

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexib…

DGX agent

For anyone building agentic workflows: the real bottleneck isn't the model, it's the harness. Open standards are a must, not just for flexibility, but to ensure we actually own the long-term memory th

agentsharrison-chase--x
10 Apr 2026
Tools

If you ask ChatGPT voice mode for its knowledge cutoff date it tells you April 2024 - it's a GPT-4o era model

DGX agent

ChatGPT's voice mode, when directly queried about its knowledge cutoff date, reports April 2024, indicating it is powered by a GPT-4o era model rather than a more recent one. GPT-4o initially had ...

toolssimon-willison--x
10 Apr 2026
Industry

Tesla is down to the last few hundred Model S & X cars in inventory. Poignant end of an era.

DGX agent

Tesla is down to the last few hundred Model S & X cars in inventory. Poignant end of an era. Traded in my 2020 Model S for a brand new plaid X before they discontinue it. Car is amazing, but the FSD h

industryelon-musk--x
10 Apr 2026
Industry

Meta's new AI can predict your brain better than a brain scan. TRIBE v2 is a foundation model trained on 1,000+ hours of brain imaging data …

DGX agent

Meta's new AI can predict your brain better than a brain scan. TRIBE v2 is a foundation model trained on 1,000+ hours of brain imaging data from 720 people. You feed it a video, sound clip, or text, a

industryrowan-cheung--x
9 Apr 2026
Model Releases

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Cla…

DGX agent

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Claude Opus 5 Max Agentic AI is about more than answering quest

model-releaseselon-musk--x
12 Aug 2026
Model Releases

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and…

DGX agent

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and delivers particularly strong results on document understand

model-releasescohere--x
12 Aug 2026
Model Releases

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models a…

DGX agent

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing the frontier of coding and knowledge work, but su

model-releasesjerry-liu--x
11 Aug 2026
Agents

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months a…

DGX agent

the term “RLM” (recursive language model) got a lot of buzz this week, but this idea is not new! @a1zhang wrote the og RLM paper 10 months ago! thats like 5 agent-years! would highly recommend followi

agentsharrison-chase--x
9 Aug 2026
Applications

It is past time to take AI & security seriously at the individual level as well. If its not the current OpenAI and Anthropic models doing it…

DGX agent

It is past time to take AI & security seriously at the individual level as well. If its not the current OpenAI and Anthropic models doing it, then the coming open weights models will when they catch u

applicationsethan-mollick--x
6 Aug 2026
Model Releases

All models are currently 20% discounted in Portal, other than GPT-5.6 Terra and Luna which 50% off and DeepSeek V4 Flash 0731 which is 90% o…

DGX agent

All models in the Portal are discounted by 20 %, except GPT‑5.6 Terra and Luna (50 % off) and DeepSeek V4 Flash 0731 (90 % off). The latest Alibaba Qwen release, Qwen3.8‑Max, is now available for Herm

model-releasesnous-research--x
4 Aug 2026
Model Releases

🛡️Introducing Shieldstral, Mistral’s 3B open-weights model for content safety that can be deployed on-device 🧵 http://mistral.ai/news/shie…

DGX agent

Mistral AI introduced Shieldstral, a 3‑billion‑parameter, open‑weights model designed for content‑safety tasks and capable of on‑device deployment. The announcement was shared via a tweet from @Mistra

model-releasesmistral-ai--x
4 Aug 2026
Safety

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and …

DGX agent

Today, we’re launching Alpamayo 2 Super, our frontier open reasoning model for autonomous vehicles. Beyond seeing, Alpamayo understands and reasons through the complex world - thinks before it acts. I

safetyelon-musk--x
4 Aug 2026
Model Releases

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorp…

DGX agent

AFAIK the most significant breakthrough since 2017 besides scaling old ideas was broadening from base models into larger systems that incorporate symbol/manipulating entities like harnesses, tools, an

model-releasesgary-marcus--x
3 Aug 2026
Model Releases

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer s…

DGX agent

An internal version of our next major model produced 10 new results on long-standing open problems in mathematics and theoretical computer science, using roughly $2,000 worth of tokens at GPT-5.6 Sol

model-releasesopenai--x
3 Aug 2026
Agents

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the i…

DGX agent

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the interaction they originate in. Each mode gets assigned to an

agentsdair-ai--x
3 Aug 2026
Model Releases

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within …

DGX agent

.@ssankar says Palantir was able to make Nvidia's Nemotron Ultra model 'better than frontier': 'I literally almost felt gaslit when, within 24 hours of getting Nemotron up with no post-training, this

model-releasesclem-delangue--x
3 Aug 2026
Applications

This is a wild result. Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tune…

DGX agent

This is a wild result. Locus, the automated research system from @intology, post-trained Qwen3 base models that beat the official human-tuned Qwen3 1.7B Instruct release. SoTA on PostTrainBench! The m

applicationsdair-ai--x
3 Aug 2026
Model Releases

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than …

DGX agent

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than K3) & works on a MacBook / Spark This is Q1 flagship (Opus 4

model-releasesemad-mostaque--x
31 Jul 2026
Model Releases

Introducing Qwen-Audio-3.0-ASR-Flash: More context-aware. Stronger domain-term recognition. 🚀Our latest ASR model upgrades: • Context consi…

DGX agent

Introducing Qwen-Audio-3.0-ASR-Flash: More context-aware. Stronger domain-term recognition. 🚀Our latest ASR model upgrades: • Context consistency • Domain-term recognition • Custom hotwords • Speech p

model-releasesqwen--x
31 Jul 2026
Model Releases

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets

DGX agent

OpenAI’s entire growth loop: new models, price cuts, and Tibo’s token resets major price cuts today: *80% drop for GPT-5.6 Luna, now 0.20 per million input tokens and 1.20 per million output *20% drop

model-releasesjerry-liu--x
31 Jul 2026
Model Releases

BREAKING: SpaceXAI's newly released Grok Voice Think Fast 2.0 beats voice models from OpenAI, Google, Alibaba, and DeepSlate in the Artifici…

DGX agent

SpaceXAI has released its new Grok Voice Think Fast 2.0, which on the Artificial Analysis Speech‑to‑Speech benchmark outperformed leading models from OpenAI, Google, Alibaba and DeepSlate. The claim w

model-releaseselon-musk--x
29 Jul 2026
← Previous
1…1415161718…128
Next →