刷新

2026 OpenAI 官方 Token 价表怎么读:GPT-5.6 系列 $ /M 字段详解

内容刷新 / GEO:补 English summary 与最新核对清单 — oa-2026-08-oa-token-price-table

返回指南列表

封面:2026 OpenAI 官方 Token 价表怎么读:GPT-5.6 系列 $ /M 字段详解

2026 OpenAI 官方 Token 价表怎么读:GPT-5.6 系列 $/M 字段详解

OpenAI 官方 API 的 Token 计费是开发者决定是否采用 GPT-5.6 系列的核心依据。$/M(每百万 tokens)字段直接决定了你的实际支出:算一个项目要多少美元,是否值得切换模型。

ChatGPT Plus 用户通常不直接看这个字段(其使用量按订阅计算),而 API 开发者(包括 Cursor、Claude Code、Grok 等第三方工具)每天都在用它对账。

如果你正在查账单、对比不同模型的性价比,或者想知道为什么同样的一段 prompt 换成 Luna 可能省 80%,那这篇文章就是你的参考清单。它帮你快速读懂官方定价表,避免误判。

现状与数据更新

2026 年 7 月 30 日,OpenAI 对 GPT-5.6 系列进行了价格优化:

  • Luna 降价 80%(最便宜,适合日常工作)
  • Terra 降价 20%(均衡选择)
  • Sol 继续保持原价($4/$20),但官方标注“至少有效至 2026 年 11 月 21 日”

这些调整后,官方定价表已更新到 2026 年 9 月,短上下文(Short context) 和 长上下文(Long context) 价格完全不同。缓存价格(Cached input)也单独列出,写 token(Cache writes)需额外注意。

以下是当前(2026-09-20)标准定价表(单位:美元,1M tokens):

模型 上下文 输入($/M) 缓存输入($/M) 缓存写入($/M) 输出($/M)
gpt-5.6-luna 短 0.20 0.02 0.25 1.20
gpt-5.6-luna 长 0.40 0.04 0.50 1.80
gpt-5.6-terra 短 2.00 0.20 2.50 12.00
gpt-5.6-terra 长 4.00 0.40 5.00 18.00
gpt-5.6-sol 短 4.00 0.40 5.00 20.00
gpt-5.6-sol 长 8.00 0.80 10.00 30.00

数据来源:OpenAI 官方 API 定价页面(developers.openai.com/api/docs/pricing)。实际使用请以平台显示为准(部分地区可能上调 10%)。

核对清单

在对账单时,按以下 5 步操作,最快能 90% 准确:

1. 模型别名确认

平台显示 gpt-5.6 或 GPT-5.6 Sol/Terra/Luna 就对。Cyber 变种单独计费(输入更高)。

2. 上下文长度选择

短上下文(<128k tokens)用上面短列;超过就用长列。平台会自动判断。

3. 缓存使用率

开启 prompt caching 后,输入价格大幅降低,但需要支持缓存的工具/平台才生效。

4. 缓存写入成本

即使用了缓存,首次写入仍按全价算。后续重复 token 才走缓存价。

5. 总成本计算示例

- 10 万 tokens(约 1 万字 prompt)+ 2 万 tokens 输出,无缓存:

短上下文 Luna:(100k + 20k)× 0.0002 + 20k × 0.0012 = $0.036

短上下文 Terra:(100k + 20k)× 0.002 + 20k × 0.012 = $0.36

- 换长上下文 Sol:约 $1.80(贵 50 倍!)

建议用站内 Token 成本工具(/official-prices)一次性算出你的 prompt 总价。

风险边界

仅供参考,非法律意见。实际计费可能因地区、特殊处理模式(Fast 模式等)或平台差异而有浮动。

不要用此价表直接打官司或维权——只用于决策对比。

升级后必挂风险:Sol 仍为旗舰,但 Luna/Terra 性价比更高,长期坚持 Sol 会让预算快速消耗;如果你的工作不需要 Sol 的顶级智能,切换 Terra 即可省 20-80% 成本。

ChatGPT Plus 与 API 价格体系完全不同,别混淆。

站内路径

想更深一步?

English summary

This guide shows developers exactly how to read OpenAI’s 2026 GPT-5.6 series token pricing table to match their billing statements and calculate real costs. The $/M fields (input/output, short vs long context, cached input/cached writes) determine your monthly spend when using the API with tools like Cursor or Claude Code.

Updated after the July 30, 2026 price reductions (Luna –80%, Terra –20%) and the Sol promotional pricing that runs through at least November 2026, the table gives short-context and long-context rates plus separate caching pricing.

Follow the 5-step checklist to verify your model alias, context length, caching usage, and compute exact costs. One example: 100k input + 20k output on Luna short-context costs only $0.036.

For deeper comparison with ChatGPT Plus usage or full official documentation, see the linked paths on openaicn.cn. Prices are subject to regional adjustments and are for reference only—always check the official API dashboard for your account.

This information helps users decide whether to switch models, avoid miscalculations, and optimize API costs without resorting to third-party tools.

---

延伸阅读