GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning to allocate computation dynamically, responding quickly to simple queries while spending more depth on complex tasks. The model produces clearer, more grounded explanations with reduced jargon, making it easier to follow even on technical or multi-step problems.
Built for broad task coverage, GPT-5.1 delivers consistent gains across math, coding, and structured analysis workloads, with more coherent long-form answers and improved tool-use reliability. It also features refined conversational alignment, enabling warmer, more intuitive responses without compromising precision. GPT-5.1 serves as the primary full-capability successor to GPT-5
| $1.25 | $10.00 | $0.13 | 1.12s | 55 tps | ||
| $1.25 | $10.00 | $0.125 | 1.33s | 38 tps | ||
Not used in Standard routing:Why these endpoints are not used | ||||||
Flex | $0.625 | $5.00 | $0.0625 | 0.99s | 43 tps | |
| $1.375 | $11.00 | $0.143 | 1.62s | 107 tps | ||
Fast | $2.50 | $20.00 | $0.25 | -- | -- | |
P50, best across providers
P50, best provider
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.