Skip to main content

News · AI

OpenAI adds Ultrafast mode for GPT-6.1 Sol in the API

Businesses using gpt-6.1-sol via the Responses API can enable service_tier: "ultrafast" from October 8, 2026 to cut latency between output tokens; open to all API users under rate limits.

1 min read Source: developers.openai.com

What changed: OpenAI added Ultrafast mode for GPT-6.1 Sol in the Responses API. Using gpt-6.1-sol with service_tier: "ultrafast" reduces the time between generated output tokens. The feature is available to all API users subject to rate limits, with global processing and US and EU data residency options.

Who is affected

Businesses using the OpenAI Responses API, specifically the gpt-6.1-sol model.

What it means for your business

For businesses using gpt-6.1-sol through the Responses API, this means they can set service_tier to "ultrafast" to reduce the delay between output tokens, letting their applications respond faster.

What to do

If your Responses API integration uses gpt-6.1-sol, set service_tier to "ultrafast" and test the speed gain; review rate limits and the US/EU data residency options.

When it takes effect

Rolled out as of October 8, 2026.

Source: https://developers.openai.com/api/docs/changelog#2026-10-08

Related reading

First conversation

Tell us what you want to do, and we will work out together where to start.

In the first call we talk through your business, where things stand and what matters most. We say plainly which parts make sense for us to take on and which you should run yourself.

Cookies and measurement

Apart from what the site needs to work, measurement or advertising tags only run if you allow them. No measurement tags are active on this site right now. Details