GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence

Hacker News Top Models

Summary

OpenAI's GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, scoring near GPT-6 Astra on the Intelligence Index while costing less than a quarter per task and pushing the cost-efficiency Pareto frontier.

No content available
Original Article
View Cached Full Text

Cached at: 09/30/26, 11:01 AM

# GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence Source: [https://artificialanalysis.ai/articles/gpt-6-1-sol-replaces-gpt-6-sol-after-just-7-days-with-near-astra-intelligence](https://artificialanalysis.ai/articles/gpt-6-1-sol-replaces-gpt-6-sol-after-just-7-days-with-near-astra-intelligence) [See model page](https://artificialanalysis.ai/models/gpt-6-1-sol)**GPT\-6\.1 Sol replaces GPT\-6 Sol after just 7 days\. It scores 1 point below GPT\-6 Astra in the Intelligence Index at less than one quarter of the Cost per Task\.** ![](https://cdn.sanity.io/images/6vfeftx9/articles/3ff27ab0c04cb2db338f97b27b3688dfbde25dcf-2770x2120.png?w=1200&auto=format) Pricing matches GPT\-6 Sol at $2/$10 per million input/output tokens, except that the cache read discount rises from 90% to 95%\. GPT\-6\.1 Sol’s overall blended price for agentic workloads is therefore slightly lower than GPT\-6 Sol\. This represents an additional price cut, following GPT\-6 Sol’s original 50% discount from GPT\-5\.6 Sol\. **Key takeaways:** **➤ Achieves near\-Astra Intelligence:**GPT\-6\.1 Sol gains 4 points in the Intelligence Index vs GPT\-6 Sol, and 5 points vs GPT\-5\.6 Sol \- landing 1 point below GPT\-6 Astra\. It makes significant gains in agentic knowledge work, improving 4 points and 5 points in AA\-Briefcase v1\.1 and GDPval\-AA v2\.1 respectively\. Other notable gains include a 12 point jump in Terminal\-Bench 4\.0, a 5 point jump in Humanity’s Last Exam, a 6 point jump in GDP\.pdf, and an 8 point jump in AA\-Omniscience Accuracy coupled with hallucination rate falling from 60% to 54%\. **➤ Pushes cost efficiency frontier:**At max effort, GPT\-6\.1 Sol costs less than a quarter of GPT\-6 Astra per Intelligence Index task \($0\.72 vs $3\.26\)\. It also costs 31% less per task than GPT\-6 Sol \($1\.05\) and 64% less than GPT\-5\.6 Sol \($1\.99\)\. All effort levels of GPT\-6\.1 Sol push out the cost efficiency Pareto frontier: for a given level of intelligence, there is no cheaper model\. **➤ Pushes token efficiency frontier, but uses slightly more output tokens than GPT\-6 Sol:**GPT\-6\.1 Sol uses ~10\-30% more output tokens than GPT\-6 Sol across effort levels\. However, due to the increase in Intelligence Index score, its low and medium effort levels are Pareto optimal for token efficiency\. **➤ Gains in Coding Agent Index:**GPT\-6\.1 Sol gains 3 points on GPT\-6 Sol at max effort in the Artificial Analysis Coding Agent Index, and sits 2 points below GPT\-6 Astra\. ## Coding Agent Index GPT\-6\.1 Sol dominates the lower\-price range of the Pareto frontier for Artificial Analysis Coding Agent Index vs Cost per Task\. GPT\-6\.1 Sol \(xhigh\) scores 1 point above GPT\-6 Astra for less than 15% of the Cost per Task\. This represents a 6 point gain from GPT\-6 Sol \(max\)\. We observed the xhigh effort setting to outperform the max effort setting by 3 points\. ![](https://cdn.sanity.io/images/6vfeftx9/articles/9e15d07ee44019baf9b0133da9ca3aa0ac7b6acd-3308x2062.png?w=1200&auto=format) ## Token efficiency GPT\-6\.1 Sol uses 10\-30% more output tokens than GPT\-6 Sol across effort settings in the Intelligence Index\. However, due to increases in intelligence, its low and medium effort settings are Pareto optimal for token efficiency\. ![](https://cdn.sanity.io/images/6vfeftx9/articles/9e4d7d0678aa29f9bbad8fea64e47b1fe1aa0bc7-2724x2152.png?w=1200&auto=format) ## AA\-Omniscience At max effort, GPT\-6\.1 Sol jumps 8 points in AA\-Omniscience Accuracy coupled with a 6 point reduction in hallucination rate\. ![](https://cdn.sanity.io/images/6vfeftx9/articles/d9d4a500a9bdd48cf05c5222decaf080018f4305-2560x2896.png?w=1200&auto=format) ## AA\-Briefcase GPT\-6\.1 Sol improves by ~80 Elo in AA\-Briefcase\. This is driven by increases in its rubric score and Analytical Quality Elo, while Presentation Elo falls slightly\. ![](https://cdn.sanity.io/images/6vfeftx9/articles/12e7144cb6c3126037ac2f6d2f637d3b7b378abb-2560x2720.png?w=1200&auto=format) ## Results by evaluation Breakdown of the individual evaluations in the Artificial Analysis Intelligence Index v4\.3\.2\. ![](https://cdn.sanity.io/images/6vfeftx9/articles/71ffd1bd5547cebaba4c4879a40dd4a147fcbe7b-2560x4072.png?w=1200&auto=format) Compare GPT\-6\.1 Sol with other leading models at:[artificialanalysis\.ai/models/releases/gpt\-6\-1\-sol](https://artificialanalysis.ai/models/releases/gpt-6-1-sol)

Similar Articles