Model Releases
@OpenAI released GPT 5.6 today and we ran a day 0 benchmark in ParseBench to test improvements in document understanding. The new family of …
@OpenAI released GPT 5.6 today and we ran a day 0 benchmark in ParseBench to test improvements in document understanding. The new family of models continues to excel at reading text and tables, but co
@OpenAI released GPT 5.6 today and we ran a day 0 benchmark in ParseBench to test improvements in document understanding. The new family of models continues to excel at reading text and tables, but continues to struggle with charts and layout. What's most interesting is that Luna is about 6 x cheaper than Sol and only results in minor degradations across all ParseBench metrics, showing that an increase in reasoning tokens doesn't always result in a commensurate improvement in visual understanding.
Source: Jerry Liu (X) | 2026-07-09