Model Releases
Writing fiction seems to be a genuine weak spot for LLMs that is not improving as rapidly as almost every other area. There may be a lot of …
Writing fiction seems to be a genuine weak spot for LLMs that is not improving as rapidly as almost every other area. There may be a lot of reasons why this is happening. It would be a really interest
Writing fiction seems to be a genuine weak spot for LLMs that is not improving as rapidly as almost every other area. There may be a lot of reasons why this is happening. It would be a really interesting benchmark (but you would need human judges, AI judges love AI fiction).
Related
- So we now have a pretty good picture of the state of the frontier AI model makers. US closed source models continue to lead. Google, OpenAI,…
- The US frontier labs have all walked away from open weights. They continue to occasionally release excellent open models (Gemma 4, etc), but…
- After playing with it a bit, Meta's Muse Spark Thinking is fine so far, but really doesn't match the current Big Three models. It also is a …
- SuperClaude (Mythos) still seems irreducibly Claude-y given the transcripts in the system card. Here two versions of Mythos are forced to ta…
Source: Ethan Mollick (X) | 2026-04-08