On DeepSWE v1.1, GPT-5.6 Sol at maximum effort scored 72.7% at a reported $6.47 per task. Claude Fable 5 scored 69.7% at $21.63 per task on the same 113-task, long-horizon engineering benchmark.
GPT-5.6 Sol Max was already the better deal on DeepSWE v1.1 - The recent price cut widens the gap even further. 🔥 Sol scores 72.7% at $6.47/task, compared with Fable 5 Max at 69.7% and $21.63/task. DeepSWE tests coding agents on 113 original, long-horizon engineering tasks.
That makes the earlier pricing story concrete: in this benchmark, Sol paired a slightly higher score with roughly 70% lower task cost. It does not prove Sol is cheaper for every workload, but it shows how the temporary token discount can widen an existing cost advantage in output-heavy coding-agent runs.