Fast Mode in the OpenAI API
A paid speed tier that runs GPT-5.6 Sol up to 2.5 times faster for double the price. It replaces the old Priority Processing option.

Details
- Fast mode serves GPT-5.6 Sol up to 2.5 times faster than standard processing and costs twice as much per token
- Model quality is unchanged, the only difference is serving speed
- It replaces Priority Processing, and requests already tagged priority route to Fast mode automatically with no code change
- Available in the OpenAI API
Best for
Anything where a person sits waiting on the answer, like a live chat surface or an internal tool used in the moment.
When to use it
Latency is what users complain about and you can absorb double the token cost on that specific call.
Who benefits
Developers running customer-facing features on GPT-5.6 Sol.
Staying informed is great. Staying ahead is better.
Pick a topic and leave with something you created.
Get new module drops and weekly AI strategy.

