ITCOW牛新网 9月30日消息,OpenAI 2026开发者日推出GPT‑6 Astra Ultrafast极速服务层级。Codex环境最高可实现8倍token生成速度,API调用最高提速6倍,峰值输出可达每秒300个词元。

该能力面向Pro 500订阅用户以及企业客户开放,可在ChatGPT Work、Codex中使用,现已开放API调用,官方后续还计划推出GPT‑6.1 Sol Ultrafast版本。
API计费区分短上下文与长上下文,还区分普通输入、缓存写入、复用缓存输入三类价格,长上下文模式整体收费更高,缓存复用可显著降低开销。Ultrafast兼顾强模型能力与低延迟,适配实时客服、事件响应、快讯生成、交互式编码等对速度敏感的业务场景。
API 调用费用方面,售价如下:
| 模型 | 价格类型 | 短上下文(输入< 272K Token) | 长上下文(输入> 272K Token) |
|---|---|---|---|
| gpt-6-astra | 输入(Input) | $60.00 / 1M tokens | $120.00 / 1M tokens |
| gpt-6-astra | 缓存输入(Cached input) | $6.00 / 1M tokens | $12.00 / 1M tokens |
| gpt-6-astra | 缓存写入(Cache writes) | $75.00 / 1M tokens | $150.00 / 1M tokens |
| gpt-6-astra | 输出(Output) | $300.00 / 1M tokens | $450.00 / 1M tokens |







