返回 AI 情报
模型2026-09-29T20:03:03.000ZX:Artificial Analysis (@ArtificialAnlys)

Artificial Analysis 评测 GPT-6.1 Sol:智能指数仅比 GPT-6 Astra 低 1 分,单任务成本不到其四分之一

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter ...

AI 摘要

Artificial Analysis 发布对 OpenAI 新模型 GPT-6.1 Sol 的评测,其 Intelligence Index 得分仅比 GPT-6 Astra 低 1 分,max effort 下每任务成本 $0.72,不到 GPT-6 Astra($3.26)的四分之一。

正文 · AI 翻译

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter of the Cost per Task

Pricing matches GPT-6 Sol at $2/$10 per million input/output tokens, except that the cache read discount rises from 90% to 95%. GPT-6.1 Sol’s overall blended price for agentic workloads is therefore slightly lower than GPT-6 Sol. This represents an additional price cut, following GPT-6 Sol’s original 50% discount from GPT-5.6 Sol.

Key takeaways:

➤ Achieves near-Astra Intelligence: GPT-6.1 Sol gains 4 points in the Intelligence Index vs GPT-6 Sol, and 5 points vs GPT-5.6 Sol - landing 1 point below GPT-6 Astra. It makes significant gains in agentic knowledge work, improving 4 points and 5 points in AA-Briefcase v1.1 and GDPval-AA v2.1 respectively. Other notable gains include a 12 point jump in Terminal-Bench 4.0, a 5 point jump in Humanity’s Last Exam, a 6 point jump in GDP.pdf, and an 8 point jump in AA-Omniscience Accuracy coupled with hallucination rate falling from 60% to 54%.

➤ Pushes cost efficiency frontier: At max effort, GPT-6.1 Sol costs less than a quarter of GPT-6 Astra per Intelligence Index task ($0.72 vs $3.26). It also costs 31% less per task than GPT-6 Sol ($1.05) and 64% less than GPT-5.6 Sol ($1.99). All effort levels of GPT-6.1 Sol push out the cost efficiency Pareto frontier: for a given level of intelligence, there is no cheaper model.

➤ Pushes token efficiency frontier, but uses slightly more output tokens than GPT-6 Sol: GPT-6.1 Sol uses ~10-30% more output tokens than GPT-6 Sol across effort levels. However, due to the increase in Intelligence Index score, its low and medium effort levels are Pareto optimal for token efficiency.

➤ Gains in Coding Agent Index: GPT-6.1 Sol gains 3 points on GPT-6 Sol at max effort in the Artificial Analysis Coding Agent Index, and sits 2 points below GPT-6 Astra.

Congratulations @OpenAI and @sama on the launch!

原文

Original Title

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter ...

Source

X:Artificial Analysis (@ArtificialAnlys)

Site

x.com

Published

2026-09-29T20:03:03.000Z

阅读原文· x.com

继续阅读

Artificial Analysis 评测 GPT-6.1 Sol:智能指数仅比 GPT-6 Astra 低 1 分,单任务成本不到其四分之一 - AI 情报频道 - 小黑丸