OpenAI Ultrafast mode now available on AI Gateway
Key point
The service tier costs 6× the standard per-token rate and supports US and global processing.
Details
AI Gateway now supports OpenAI's Ultrafast service tier for GPT-6 Astra, providing faster output for interactive applications and rapid coding iterations.
Access and Usage
To use Ultrafast, request it for openai/gpt-6-astra through the AI SDK, the Chat Completions API, or the Responses API. For workflows with frequent tool calls, OpenAI recommends the Responses API over a persistent WebSocket connection to reduce overhead between turns.
Regional Availability and Billing
Ultrafast supports US and global processing. Requests pinned to unsupported regions, such as the EU, run at the standard (default) tier. Standard processing remains the default when no service tier is specified.
Requests served at Ultrafast are billed at 6× the standard per-token rate, while requests that fall back to another tier are billed at the rate for the tier actually served.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.