Artificial Analysis’s independent results back the core promise of GPT-6.1 Sol: performance close to Astra at a much lower bill. At maximum reasoning effort, Sol scores 52 on its Intelligence Index versus Astra’s 53, with a weighted average cost of $0.72 per task versus $3.26, about 78% less. The index combines tests of coding, knowledge, reasoning and professional work; its score is not a percentage of jobs completed.
That is also a meaningful upgrade over GPT-6 Sol, which scores 48 and costs $1.06 per task at maximum effort. The new model scores higher while costing about a third less on this test suite.
This strengthens the case for trying Sol on work previously reserved for Astra: the savings now show up in an outside evaluator’s actual token usage, not just OpenAI’s launch charts or token prices. It does not make Sol the overall leader. Claude Opus 5.5 scores 58 and Sonnet 5.5 scores 56 at maximum effort with their default fallback models enabled.


