The new GPT-5.6 Sol (max) model, integrated into Codex, now leads all evaluations in the Artificial Analysis Coding Agent Index. It specifically ties Grok 4.5 in Grok Build for the SWE-Atlas-QnA benchmark. Furthermore, GPT-5.6 Sol (max) demonstrates a lower Cost per Task compared to both Claude Fable 5 (max) and Claude Opus 4.8 (max).
OpenAI's latest model offers a more cost-effective, top-performing option for coding agent workflows, shifting demand for inference compute.