Ultrafast is OpenAI's new speed tier for GPT-6.1 Sol. OpenAI is rolling out Ultrafast for GPT-6.1 Sol in the API, Codex and ChatGPT Work, the company said on X on 8 October 2026. The company describes it as near-Astra intelligence at up to 8x the speed of Sol Standard. The API price is $12 per million input tokens and $60 per million output tokens, six times Sol Standard's $2 and $10 rates.
Ultrafast API pricing
OpenAI's longer announcement on its developer community forum gives the Ultrafast price as $12 per million input tokens and $60 per million output tokens. The company says this is 1.2 times the cost of Astra.
The Standard rates in this article are $2 per million input tokens and $10 per million output tokens. Against those rates, Ultrafast costs six times as much for both input and output. As a hypothetical example, a workload of 1 million input tokens and 1 million output tokens would cost $12 at Standard rates ($2 + $10) and $72 on Ultrafast ($12 + $60).
Fast mode, which costs twice the Standard price, is a separate option. Ultrafast has its own rate and should not be read as a version of Fast mode.
Who can use it
In Codex and ChatGPT Work, Ultrafast is listed for Pro 500, eligible usage-based Enterprise and credit-based Edu plans. Enterprise administrators must enable access for their organisation. The announcement names those plans, and this article reports only those.
In the API, the company says Ultrafast for GPT-6.1 Sol is available in all supported regions, including US and EU data residency.
EU data residency for Sol Fast
This article previously stated that Fast mode is unavailable with EU data residency. OpenAI's announcement says EU data residency support has now been added for GPT-6.1 Sol Fast and GPT-6 Luna Fast. Developers who avoided Fast mode for that reason should check availability in their own account.
Does this change your choice?
The decision depends on whether faster responses are worth roughly six times the Standard token price. OpenAI positions Ultrafast for work where speed matters, such as debugging an outage, agents that navigate apps and live experiences where each second counts. The company's "up to 8x" figure is a ceiling rather than a typical result.
In our view, batch jobs or workloads where latency matters little have less reason to pay the premium. The announcement does not include independent speed or quality benchmarks, so anyone deciding whether to use Ultrafast should measure latency and cost on their own workload.





0 comments
No approved comments yet. You can start the conversation.
Leave a comment