GPT-5.6 Sol Scores Similarly to Claude Fable 5 on GDPval-AA v2 Benchmark
OpenAI's GPT-5.6 Sol (max) achieved comparable performance to Anthropic's Claude Fable 5 (max) on the GDPval-AA v2 benchmark. This result indicates a similar capability between the two frontier models for completing economically valuable tasks.