GPT-6 Sol is built for complex coding and agentic workflows.
reasoning.effort supports none, low, medium (default), high, xhigh, and max.
Use the Responses API for built-in tools and function calling. Chat Completions supports function calling only with reasoning_effort set to none.
EU data residency is available only with Standard processing. See data residency eligibility.
Cached input tokens are priced at 10% of the uncached input token rate.
Cache writes are billed at 1.25x the uncached input token rate.
Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.
Regional processing adds a 10% premium where available. EU data residency is available only with Standard processing.
Batch and Flex are priced at 50% of Standard rates. Fast mode is priced at 2x the applicable rates.
Use gpt-6-sol in your API requests.
| Tier | RPM | TPM | Batch queue limit |
|---|---|---|---|
| Free | Not supported | ||
| Tier 1 | 500 | 500,000 | 1,500,000 |
| Tier 2 | 5,000 | 1,000,000 | 3,000,000 |
| Tier 3 | 5,000 | 2,000,000 | 100,000,000 |
| Tier 4 | 10,000 | 4,000,000 | 200,000,000 |
| Tier 5 | 15,000 | 40,000,000 | 15,000,000,000 |