ModelsPublished All news from that day
Ongoing case: OpenAI's models 18 storiesSdkb / Selena Deckelmann at Wikimania 2025 / CC BY-SA 4.0, croppedOpenAI added Ultrafast mode for GPT-6.1 Sol, priced at $12 per million input tokens.
OpenAI's changelog dates it 8 October. Ultrafast is the API's fastest service tier; for GPT-6.1 Sol it is priced at $12 per million input tokens and $60 output at short context.
Check our sources · 13 facts from 3 sourcesOpenAI's API changelog for 8 October 2026 says Ultrafast mode is added for GPT-6.1 Sol in the Responses API. A request sets the model to gpt-6.1-sol and the service tier to ultrafast, which OpenAI says reduces the time between generated output tokens.
OpenAI says it is available to all API users, subject to rate limits, with global processing and US and EU data residency. Its guide calls Ultrafast the fastest service tier in the API, says it is available for GPT-6 Astra and GPT-6.1 Sol, and says to use it when speed justifies the higher cost.
The pricing page lists gpt-6.1-sol on Ultrafast at $12.00 per million input tokens, $0.60 cached input and $60.00 output at short context, and $24.00, $1.20 and $90.00 at long context. The same page lists Standard at $2.00, $0.10 and $10.00 at short context. For GPT-6 Astra, Ultrafast is $60.00 input and $300.00 output at short context.
The guide says Ultrafast has separate rate limits from Standard and Fast modes; for GPT-6.1 Sol they are 1,000,000 tokens per minute at the Build tier, 4,000,000 at Launch and 40,000,000 at Grow. It gives no tokens-per-second figure. Astra's Ultrafast supports US data residency and global processing only. The guide strongly recommends WebSockets, especially for agents that make many tool calls in quick succession.
Underlined sentences link to the facts we checked. Dotted words open a short explainer.
Your reaction
We count reactions per story and day, never who reacted. The counts help us choose what goes in the monthly issue. If you are signed in, your own page shows yours too.
We count reactions per story and day, never who reacted. They help choose what goes in the monthly issue.
Check our sources
Every sentence above is checked against these 3 sources.
1 OpenAI API changelog
Open the source archived copy-
OpenAI API changelog, 8 October 2026: Ultrafast mode added for GPT-6.1 Sol in the Responses API.
Toned down to what the source says
citeAdded [Ultrafast mode](https://developers.openai.com/api/docs/guides/ultrafast-mode) for [GPT-6.1 Sol](https://developers.openai.com/api/docs/models/gpt-6.1-sol) in the Responses API.
-
The changelog says to use gpt-6.1-sol with service_tier: "ultrafast" to reduce the time between generated output tokens.
citeUse `gpt-6.1-sol` with `service_tier: "ultrafast"` to reduce the time between generated output tokens.
-
The changelog says Ultrafast for GPT-6.1 Sol is available to all API users, subject to rate limits, with global processing and US and EU data residency.
citeIt is available to all API users, subject to rate limits, with global processing and US and EU data residency.
2 Ultrafast mode | OpenAI API
Open the source archived copy-
OpenAI's guide says Ultrafast mode is the fastest service tier in the OpenAI API, broadly available for GPT-6 Astra and GPT-6.1 Sol, to use when speed justifies the higher cost.
citeUltrafast mode is the fastest service tier in the OpenAI API. It is broadly available for GPT-6 Astra and GPT-6.1 Sol. Use it when speed justifies the higher cost.
-
OpenAI's guide says Ultrafast has separate rate limits from Standard and Fast modes.
citeUltrafast has separate rate limits from Standard and Fast modes.
-
OpenAI's guide says it strongly recommends WebSockets, especially for agentic applications that make many tool calls in quick succession.
citeWe strongly recommend WebSockets, especially for agentic applications that make many tool calls in quick succession.
-
OpenAI's Ultrafast guide says: "GPT-6.1 Sol supports US and EU data residency and global processing. GPT-6 Astra Ultrafast supports US data residency and global processing only."
citeUltrafast mode for GPT-6.1 Sol supports US and EU data residency and global processing. GPT-6 Astra Ultrafast supports US data residency and global processing only.
-
The guide lists Ultrafast rate limits for gpt-6.1-sol of 1,000,000 tokens per minute at the Build tier, 4,000,000 at Launch and 40,000,000 at Grow.
citeBuild | 1,000,000
-
The guide gives no tokens-per-second figure for Ultrafast; it calls it the fastest service tier without stating a speed.
citeUltrafast mode is the fastest service tier in the OpenAI API.
3 Pricing | OpenAI API
Open the source archived copy-
OpenAI's pricing page lists gpt-6.1-sol under Ultrafast, short context, at $12.00 per million input tokens, $0.60 cached input and $60.00 output.
citegpt-6.1-sol | $12.00 | $0.60 | $60.00 | $24.00 | $1.20 | $90.00
-
OpenAI's pricing page lists gpt-6.1-sol under Ultrafast, long context, at $24.00 per million input tokens, $1.20 cached input and $90.00 output.
citegpt-6.1-sol | $12.00 | $0.60 | $60.00 | $24.00 | $1.20 | $90.00
-
For comparison the pricing page lists gpt-6.1-sol at Standard, short context, at $2.00 per million input tokens, $0.10 cached input and $10.00 output.
citegpt-6.1-sol | $2.00 | $0.10 | $10.00 | $4.00 | $0.20 | $15.00
-
The pricing page lists gpt-6-astra under Ultrafast, short context, at $60.00 per million input tokens, $6.00 cached input and $300.00 output.
citegpt-6-astra | $60.00 | $6.00 | $300.00 | $120.00 | $12.00 | $450.00
Every sentence above is checked against these 3 sources. Each fact has its own address you can cite.
Source 1 · OpenAI · 8 Oct 2026
OpenAI API changelog
OpenAI API changelog, 8 October 2026: Ultrafast mode added for GPT-6.1 Sol in the Responses API.
Toned down to what the source saysAdded [Ultrafast mode](https://developers.openai.com/api/docs/guides/ultrafast-mode) for [GPT-6.1 Sol](https://developers.openai.com/api/docs/models/gpt-6.1-sol) in the Responses API.
The changelog says to use gpt-6.1-sol with service_tier: "ultrafast" to reduce the time between generated output tokens.
Use `gpt-6.1-sol` with `service_tier: "ultrafast"` to reduce the time between generated output tokens.
The changelog says Ultrafast for GPT-6.1 Sol is available to all API users, subject to rate limits, with global processing and US and EU data residency.
It is available to all API users, subject to rate limits, with global processing and US and EU data residency.
Source 2 · OpenAI
Ultrafast mode | OpenAI API
OpenAI's guide says Ultrafast mode is the fastest service tier in the OpenAI API, broadly available for GPT-6 Astra and GPT-6.1 Sol, to use when speed justifies the higher cost.
Ultrafast mode is the fastest service tier in the OpenAI API. It is broadly available for GPT-6 Astra and GPT-6.1 Sol. Use it when speed justifies the higher cost.
OpenAI's guide says Ultrafast has separate rate limits from Standard and Fast modes.
Ultrafast has separate rate limits from Standard and Fast modes.
OpenAI's guide says it strongly recommends WebSockets, especially for agentic applications that make many tool calls in quick succession.
We strongly recommend WebSockets, especially for agentic applications that make many tool calls in quick succession.
Show all 6 factsShow fewer facts
OpenAI's Ultrafast guide says: "GPT-6.1 Sol supports US and EU data residency and global processing. GPT-6 Astra Ultrafast supports US data residency and global processing only."
Ultrafast mode for GPT-6.1 Sol supports US and EU data residency and global processing. GPT-6 Astra Ultrafast supports US data residency and global processing only.
The guide lists Ultrafast rate limits for gpt-6.1-sol of 1,000,000 tokens per minute at the Build tier, 4,000,000 at Launch and 40,000,000 at Grow.
Build | 1,000,000
The guide gives no tokens-per-second figure for Ultrafast; it calls it the fastest service tier without stating a speed.
Ultrafast mode is the fastest service tier in the OpenAI API.
Source 3 · OpenAI
Pricing | OpenAI API
OpenAI's pricing page lists gpt-6.1-sol under Ultrafast, short context, at $12.00 per million input tokens, $0.60 cached input and $60.00 output.
gpt-6.1-sol | $12.00 | $0.60 | $60.00 | $24.00 | $1.20 | $90.00
OpenAI's pricing page lists gpt-6.1-sol under Ultrafast, long context, at $24.00 per million input tokens, $1.20 cached input and $90.00 output.
gpt-6.1-sol | $12.00 | $0.60 | $60.00 | $24.00 | $1.20 | $90.00
For comparison the pricing page lists gpt-6.1-sol at Standard, short context, at $2.00 per million input tokens, $0.10 cached input and $10.00 output.
gpt-6.1-sol | $2.00 | $0.10 | $10.00 | $4.00 | $0.20 | $15.00
Show all 4 factsShow fewer facts
The pricing page lists gpt-6-astra under Ultrafast, short context, at $60.00 per million input tokens, $6.00 cached input and $300.00 output.
gpt-6-astra | $60.00 | $6.00 | $300.00 | $120.00 | $12.00 | $450.00
Topics
The morning email
On the mornings we publish: the three top stories and up to four short ones. Free.
We email you a link to confirm. An issue may include one sponsor, always labelled Sponsored · Advertisement. Our emails count opens and clicks, not who made them. Unsubscribe in one click. What we keep
