Model Releases
so on-brand
so on-brand >be openai >release gpt 5.6 sol >max thinking budgets, full psycho >benchmark results take few days to roll in >“best model ever” >reviews locked in >4 days later >silently nerf it >nobody
so on-brand >be openai >release gpt 5.6 sol >max thinking budgets, full psycho >benchmark results take few days to roll in >“best model ever” >reviews locked in >4 days later >silently nerf it >nobody re-tests lol >users still paying full price >not realizing the model they benchmarked isn’t…
Related
- Pop quiz, which of these is no longer true (or at least directionally true), two years later?
- Big Breaking News: White House asks OpenAI to delay GPT- 5.6.
- GPT-5.6 Sol sets a new SOTA on ARC-AGI-3: 7.8% Sol is the first verified frontier model to ever beat an ARC-AGI-3 game It is the best model …
Source: Gary Marcus (X) | 2026-07-13