Model Releases
Given the messy naming scheme used by all the AI companies, I caused a chart to be made showing the gain in GPQA per 0.1 version in model na…
Given the messy naming scheme used by all the AI companies, I caused a chart to be made showing the gain in GPQA per 0.1 version in model names (estimated, since model names skip version numbers). The
Given the messy naming scheme used by all the AI companies, I caused a chart to be made showing the gain in GPQA per 0.1 version in model names (estimated, since model names skip version numbers). There has never been a more misnamed model that Claude 3.7, should have been 4.4.
Related
- So what's the deal with Amazon Nova? They released Nova 2 in December, and even then, the top flight Nova 2 model trailed Sonnet 4.5. And it…
- So we now have a pretty good picture of the state of the frontier AI model makers. US closed source models continue to lead. Google, OpenAI,…
- OpenAI should probably bite the bullet and just name their next set of models something more human sounding. Everyone anthropomorphizes thei…
- After playing with it a bit, Meta's Muse Spark Thinking is fine so far, but really doesn't match the current Big Three models. It also is a …
Source: Ethan Mollick (X) | 2026-04-14