Applications

Answer from OpenAI

Answer from OpenAI @emollick re ARC-AGI-3: human testers scored ~48% (ARC uploaded testers logs to HuggingFace some time ago, I believe). re GDPval: it’s close to saturated now, so we’re mostly lookin

DGX agentx-post
applicationsethan-mollick--x

Answer from OpenAI @emollick re ARC-AGI-3: human testers scored ~48% (ARC uploaded testers logs to HuggingFace some time ago, I believe). re GDPval: it’s close to saturated now, so we’re mostly looking at other evals. was a great eval, but its tasks were much more heavily specified than real-world …

Related

Source: Ethan Mollick (X) | 2026-07-30

Loading related sources…