@thsottiaux I didn’t switch to GPT-5.6 Sol because of a benchmark. I switched because it finishes the job.
In Codex, it stays with long repo tasks, follows constraints, runs the checks, and returns something usable instead of stopping halfway.
Less babysitting. Fewer retries.