OpenAI发布GPT-6.1 Sol:中端模型逼近Astra性能
OpenAI Releases GPT-6.1 Sol: Near-Astra Coding and Computer Use at One-Fifth of Astra’s Token Price
GPT-6系列新迭代,在中端价位实现了接近旗舰的性能,对Agent开发者的成本控制有直接影响,值得关注。
This week, OpenAI released GPT-6.1 Sol. It upgrades GPT-6 Sol, the mid-tier model in the GPT-6 family. OpenAI’s claim is specific: near-Astra results on agentic coding, computer use, and professional work. The price is one-fifth of GPT-6 Astra’s standard input and output rates. Cached input drops to $0.10 per million tokens, 50% below GPT-6 Sol.
本周,OpenAI 发布了 GPT-6.1 Sol。它升级了 GPT-6 系列中的中端模型 GPT-6 Sol。OpenAI 的具体宣称是:在智能体编码、计算机使用和职业工作方面达到接近 Astra 的效果。其价格仅为 GPT-6 Astra 标准输入和输出费率的五分之一。缓存输入的费率降至每百万 token 0.10 美元,比 GPT-6 Sol 低 50%。
Is it deployable? Yes, as a hosted model. It is live today in the OpenAI API as gpt-6.1-sol, and in ChatGPT Work and Codex.
是否可部署?是的,作为托管模型。它今天已在 OpenAI API 中以 gpt-6.1-sol 的形式上线,并在 ChatGPT Work 和 Codex 中可用。
Where Sol Sits in the GPT-6 Lineup
Sol 在 GPT-6 产品线中的定位
OpenAI now ships 3 GPT-6 tiers. GPT-6 Astra costs $10 input, $50 output, and $1 cached input per million tokens. GPT-6.1 Sol costs $2, $10, and $0.10. GPT-6 Luna costs $0.10, $0.50, and $0.01.
OpenAI 目前提供 3 个 GPT-6 层级。GPT-6 Astra 的费率为输入 $10、输出 $50、每百万 token 缓存输入 $1。GPT-6.1 Sol 的费率为 $2、$10 和 $0.10。GPT-6 Luna 的费率为 $0.10、$0.50 和 $0.01。
The cached price is quite important most for agents. Agents resend the same system prompt, tool schemas, and history on every step. Cached reads now cost 5% of the uncached input rate, down from 10% on GPT-6 Sol.
缓存价格对智能体而言非常重要。智能体在每个步骤都会重新发送相同的系统提示词、工具模式和历史记录。缓存读取现在占未缓存输入费率的 5%,低于 GPT-6 Sol 的 10%。
Benchmark Results
基准测试结果
All figures below are vendor-reported in OpenAI’s launch post. OpenAI states that competitor numbers came from public reports.
以下所有数据均为 OpenAI 发布帖中厂商报告的数据。OpenAI 指出,竞争对手的数据来自公开报告。
- Coding: On DeepSWE v1.1, GPT-6.1 Sol matches GPT-6 Astra at roughly one-fifth of the cost. It beats GPT-6 Sol’s best score by 6.4 percentage points, at a lower reasoning effort.
- Professional work: On GDP.pdf, which tests answers over complex professional PDFs, Sol beats Claude Opus 5.5 with fallbacks. It does so at less than half the cost per task. On AutomationBench 1.0.6, Sol scores 2.2 points above Opus 5.5 at medium effort. That result comes at roughly one-third the cost.
- Computer use: On the OSWorld 2.0 offline set, Sol gains 7 points over GPT-6 Sol at maximum effort. It lands within 2.1 points of Astra at roughly one-seventh the cost per task.
- Science: On Terminal-Bench Science 0.1, Sol more than doubles GPT-6 Sol’s score at max effort. Average cost per task is $5.47, versus $23.21 for Opus 5.5 and $23.80 for Astra. Astra still leads at 68.1%, and OpenAI recommends it for the hardest research.
- Factuality: At low effort, the share of responses with a factual error falls from 11.4% to 7.7%. That is a reduction of about 32%. The eval uses deliberately difficult conversations where users had flagged earlier model errors.
- 编码:在 DeepSWE v1.1 上,GPT-6.1 Sol 以大约五分之一的成本达到了与 GPT-6 Astra 相当的水平。它以较低的推理努力,比 GPT-6 Sol 的最佳成绩高出 6.4 个百分点。
- 职业工作:在 GDP.pdf(测试复杂职业 PDF 的回答)上,Sol 击败了带有回退机制的 Claude Opus 5.5。其每项任务的成本低于一半。在 AutomationBench 1.0.6 上,Sol 在中等努力下比 Opus 5.5 高出 2.2 分。该结果的成本约为三分之一。
- 计算机使用:在 OSWorld 2.0 离线集上,Sol 在最大努力下比 GPT-6 Sol 高出 7 分。它以大约七分之一的每项任务成本,与 Astra 的成绩相差 2.1 分。
- 科学:在 Terminal-Bench Science 0.1 上,Sol 在最大努力下将 GPT-6 Sol 的成绩翻了一倍多。每项任务的平均成本为 5.47 美元,而 Opus 5.5 为 23.21 美元,Astra 为 23.80 美元。Astra 仍以 68.1% 领先,OpenAI 推荐将其用于最困难的研究。
- 事实性:在低努力下,包含事实错误的回答比例从 11.4% 下降到 7.7%。这减少了约 32%。该评估使用了故意设计的困难对话,其中用户曾标记过早期模型的错误。
API Details Developers Need
开发者需要的 API 详情
From the GPT-6.1 Sol model page:
来自 GPT-6.1 Sol 模型页面:
- Context window of 1,050,000 tokens, 128,000 max output tokens, and an April 30, 2026 knowledge cutoff.
- Text and image input; text output.
- reasoning.effort accepts low, medium (default), high, xhigh, and max. The none and minimal settings are not supported.
- Use the Responses API for tool calling. Chat Completions works without tool calling.
- Prompts above 272K input tokens cost 2x input and cache rates, and 1.5x output, for the full request.
- Batch and Flex are 50% cheaper. Fast mode costs 2x standard.
- US and EU data residency are supported. Fast mode is unavailable with EU residency.
- Fine-tuning is not supported.
- 上下文窗口为 1,050,000 个 token,最大输出 128,000 个 token,知识截止日为 2026 年 4 月 30 日。
- 支持文本和图片输入;文本输出。
- reasoning.effort 接受 low、medium(默认)、high、xhigh 和 max。不支持 none 和 minimal 设置。
- 使用 Responses API 进行工具调用。Chat Completions 不支持工具调用。
- 输入 token 超过 272K 的提示词,其输入和缓存费率是标准的 2 倍,输出费率是标准的 1.5 倍,适用于整个请求。
- Batch 和 Flex 模式价格低 50%。Fast 模式价格是标准模式的 2 倍。
- 支持美国(US)和欧盟(EU)数据驻留。Fast 模式在欧盟数据驻留下不可用。
- 不支持微调。
OpenAI also plans a GPT-6.1 Sol Ultrafast option in Codex within days. It promises up to 8x faster token generation than standard speed.
OpenAI 还计划在几天内在 Codex 中推出 GPT-6.1 Sol Ultrafast 选项。它承诺比标准速度快多达 8 倍的 token 生成速度。
Interactive Explainer
交互式解释器
How GPT-6.1 Sol Compares
GPT-6.1 Sol 对比分析
| Feature | GPT-6.1 Sol | Claude Opus 5.5 | Claude Sonnet 5.5 | Gemini 3.1 Pro Preview |
|---|---|---|---|---|
| Developer / API ID | OpenAI / gpt-6.1-sol | Anthropic / claude-opus-5-5 | Anthropic / claude-sonnet-5-5 | Google / gemini-3.1-pro-preview |
| Status | Generally available | Generally available | Generally available | Preview |
| Input price (per 1M) | $2.00 | $4.00 | $2.00 | $2.00 (≤200K prompt) |
| Output price (per 1M) | $10.00 | $20.00 | $10.00 | $12.00 (≤200K prompt) |
| Cached input read (per 1M) | $0.10 | $0.20 | $0.20 | $0.20 + $4.50/1M tokens/hr storage |
| Long-prompt surcharge | >272K input: 2x input and cache, 1.5x output | None; 1M at standard rates | None; 1M at standard rates | >200K: $4 input, $18 output |
| Context window | 1,050,000 | 1M | 1M | 1,048,576 |
| Max output tokens | 128,000 | 128K | 128K | 65,536 |
| Input modalities | Text, image | Text, image | Text, image | Text, image, video, audio, PDF |
| Reasoning control | low to max (5 levels), default medium | Adaptive thinking, always on; default medium | Adaptive thinking; default high | Thinking supported |
| Batch pricing (in / out) | 50% off standard | $2 / $10 | $1 / $5 | $1 / $6 |
| Knowledge cutoff | Apr 30, 2026 | Jun 2026 | Jun 2026 | Not listed |
| Open weights | No | No | No | No |
| 功能 | GPT-6.1 Sol | Claude Opus 5.5 | Claude Sonnet 5.5 | Gemini 3.1 Pro Preview |
|---|---|---|---|---|
| 开发者 / API ID | OpenAI / gpt-6.1-sol | Anthropic / claude-opus-5-5 | Anthropic / claude-sonnet-5-5 | Google / gemini-3.1-pro-preview |
| 状态 | 正式发布 | 正式发布 | 正式发布 | 预览版 |
| 输入价格(每百万) | $2.00 | $4.00 | $2.00 | $2.00 (≤200K prompt) |
| 输出价格(每百万) | $10.00 | $20.00 | $10.00 | $12.00 (≤200K prompt) |
| 缓存输入读取(每百万) | $0.10 | $0.20 | $0.20 | $0.20 + $4.50/1M tokens/hr 存储费 |
| 长提示附加费 | >272K 输入:输入和缓存 2 倍,输出 1.5 倍 | 无;按标准费率收取 1M | 无;按标准费率收取 1M | >200K:$4 输入,$18 输出 |
| 上下文窗口 | 1,050,000 | 1M | 1M | 1,048,576 |
| 最大输出 token | 128,000 | 128K | 128K | 65,536 |
| 输入模态 | 文本、图像 | 文本、图像 | 文本、图像 | 文本、图像、视频、音频、PDF |
| 推理控制 | 低到最高(5 个级别),默认中等 | 自适应思维,始终开启;默认中等 | 自适应思维;默认高 | 支持思考 |
| 批量定价(输入/输出) | 标准价格打五折 | $2 / $10 | $1 / $5 | $1 / $6 |
| 知识截止日期 | 2026 年 4 月 30 日 | 2026 年 6 月 | 2026 年 6 月 | 未列出 |
| 开放权重 | 否 | 否 | 否 | 否 |
Standard first-party API list prices, verified September 30, 2026.
标准第一方 API 列表价格,经 2026 年 9 月 30 日核实。
Claude Sonnet 5.5 matches Sol’s $2 and $10 list price, but Sol’s cached input costs half as much. OpenAI’s benchmarks compare Sol against Opus 5.5, not Sonnet 5.5. Gemini 3.1 Pro matches Sol on input, charges $12 for output, and remains in preview. List prices are not direct cost comparisons. Anthropic notes its newer tokenizer produces roughly 30% more tokens for the same text.
Claude Sonnet 5.5 的列表价格与 Sol 的 $2 和 $10 持平,但 Sol 的缓存输入成本减半。OpenAI 的基准测试将 Sol 与 Opus 5.5 进行比较,而非 Sonnet 5.5。Gemini 3.1 Pro 在输入方面与 Sol 持平,输出收费 $12,且仍处于预览阶段。列表价格并非直接的成本比较。Anthropic 指出,其较新的分词器对于相同文本可产生约多 30% 的 token。
Key Takeaways
关键要点
- GPT-6.1 Sol matches GPT-6 Astra on DeepSWE v1.1 at about one-fifth the cost.
- Pricing is $2 input, $10 output, and $0.10 cached input per 1M tokens.
- It beats Claude Opus 5.5 on AutomationBench by 2.2 points at about one-third the cost.
- It offers a 1,050,000-token context, 128K output, and 5 reasoning effort levels.
- It is API-only and closed-weight; Astra still leads the hardest science tasks.
- GPT-6.1 Sol 在 DeepSWE v1.1 上与 GPT-6 Astra 表现相当,但成本仅为后者的五分之一。
- 价格为每百万 token 输入 $2、输出 $10、缓存输入 $0.10。
- 它在 AutomationBench 上以约三分之一的成本领先 Claude Opus 5.5 达 2.2 分。
- 它提供1,050,000个token的上下文窗口、128K的输出长度以及5级推理强度。
- 它仅通过API提供且权重封闭;Astra仍在最难的科学任务中保持领先。
FAQ
常见问题解答
- What is GPT-6.1 Sol? It is OpenAI’s mid-tier GPT-6 model, released September 29, 2026, for coding, computer use, and professional work.
- How much does GPT-6.1 Sol cost? Standard API pricing is $2 per million input tokens, $0.10 cached input, and $10 output.
- Can I self-host GPT-6.1 Sol? No. It is available only through the OpenAI API, ChatGPT Work, and Codex.
- 什么是GPT-6.1 Sol?它是OpenAI的中端GPT-6模型,于2026年9月29日发布,适用于编程、计算机操作和专业工作。
- GPT-6.1 Sol的价格是多少?标准API定价为每百万输入token 2美元,缓存输入0.10美元,输出10美元。
- 我可以自行托管GPT-6.1 Sol吗?不可以。它仅可通过OpenAI API、ChatGPT Work和Codex获取。
更进一步:量化金融体系
看懂新闻只是起点——沿量化金融路径,把它变成能交付的工程能力