The new GPT-5.6-Sol model achieved a 7.78% score on the ARC-AGI-3 benchmark, representing a substantial increase over Opus 4.8's previous score of 1.5%. This benchmark result indicates a notable advancement in the model's capabilities.
OpenAI's new model demonstrates a significant leap in AGI benchmark performance, intensifying the competitive landscape for frontier AI models.