GPT-6.1 Sol Pro is the same underlying model as GPT-6.1 SolOpens in new tab, served with reasoning.mode set to pro for higher-quality responses on complex tasks.
Cost note: pro mode spends far more reasoning tokens per request, so a typical request costs several times more than the same request on GPT-6.1 Sol and takes much longer to complete. It is intended for hard, high-stakes problems where the extra accuracy justifies the cost. For everyday coding, agentic, and chat workloads, use GPT-6.1 SolOpens in new tab instead.
Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-modeOpens in new tab
| $2.00 | $10.00 | $0.10 | 10.54s | 57 tps | ||
| $2.00 | $10.00 | $0.10 | 8.51s | 43 tps | ||
Not used in Standard routing:Why these endpoints are not used | ||||||
Flex | $1.00 | $5.00 | $0.05 | 6.67s | 84 tps | |
| $2.20 | $11.00 | $0.11 | -- | -- | ||
| $2.20 | $11.00 | $0.11 | 90.47s | 65 tps | ||
Fast | $4.00 | $20.00 | $0.20 | -- | -- | |
P50, best across providers
P50, best provider
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.