GPT-6.1 Sol delivers near-Astra performance at a lower cost for complex coding, computer use, and professional work. Compare it with Astra on your tasks to assess the tradeoff between quality and cost.
reasoning.effort supports low, medium (default), high, xhigh, and
max. The none and minimal reasoning efforts are not supported.
Use the Responses API for tool calling. Chat Completions is supported without tool calling.
GPT-6.1 Sol supports US and EU data residency. Fast mode is unavailable with EU data residency. See data residency eligibility.
See GPT-6.1 Sol in the GPT-6 guide and model-selection guidance.
Cached input tokens are priced at 5% of the uncached input token rate.
Cache writes are billed at 1.25x the uncached input token rate.
Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.
Fast mode prices are 2x Standard. Batch and Flex prices are 50% lower than Standard.
Regional processing adds a 10% premium where available.
Use gpt-6.1-sol to select this model.
| Tier | RPM | TPM |
|---|---|---|
| Free | Not supported | |
| Tier 1 | 500 | 500,000 |
| Tier 2 | 5,000 | 1,000,000 |
| Tier 3 | 5,000 | 2,000,000 |
| Tier 4 | 10,000 | 4,000,000 |
| Tier 5 | 15,000 | 40,000,000 |