Flagship models
| Short context | Long context | |||||||
|---|---|---|---|---|---|---|---|---|
| Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output |
| gpt-6-astra | $10.00 | $1.00 | $12.50 | $50.00 | $20.00 | $2.00 | $25.00 | $75.00 |
| gpt-6.1-sol | $2.00 | $0.10 | $2.50 | $10.00 | $4.00 | $0.20 | $5.00 | $15.00 |
| gpt-6-luna | $0.10 | $0.01 | $0.125 | $0.50 | $0.20 | $0.02 | $0.25 | $0.75 |
| Short context | Long context | |||||||
|---|---|---|---|---|---|---|---|---|
| Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output |
| gpt-6-astra | $5.00 | $0.50 | $6.25 | $25.00 | $10.00 | $1.00 | $12.50 | $37.50 |
| gpt-6.1-sol | $1.00 | $0.05 | $1.25 | $5.00 | $2.00 | $0.10 | $2.50 | $7.50 |
| gpt-6-luna | $0.05 | $0.005 | $0.0625 | $0.25 | $0.10 | $0.01 | $0.125 | $0.375 |
| Short context | Long context | |||||||
|---|---|---|---|---|---|---|---|---|
| Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output |
| gpt-6-astra | $5.00 | $0.50 | $6.25 | $25.00 | $10.00 | $1.00 | $12.50 | $37.50 |
| gpt-6.1-sol | $1.00 | $0.05 | $1.25 | $5.00 | $2.00 | $0.10 | $2.50 | $7.50 |
| gpt-6-luna | $0.05 | $0.005 | $0.0625 | $0.25 | $0.10 | $0.01 | $0.125 | $0.375 |
| Short context | Long context | |||||||
|---|---|---|---|---|---|---|---|---|
| Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output |
| gpt-6-astra | $20.00 | $2.00 | $25.00 | $100.00 | $40.00 | $4.00 | $50.00 | $150.00 |
| gpt-6.1-sol | $4.00 | $0.20 | $5.00 | $20.00 | $8.00 | $0.40 | $10.00 | $30.00 |
| gpt-6-luna | $0.20 | $0.02 | $0.25 | $1.00 | $0.40 | $0.04 | $0.50 | $1.50 |
| Short context | Long context | |||||||
|---|---|---|---|---|---|---|---|---|
| Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output |
| gpt-6-astra | $60.00 | $6.00 | $75.00 | $300.00 | $120.00 | $12.00 | $150.00 | $450.00 |
Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026. FedRAMP endpoints are also charged a 10% uplift. Priority processing was renamed Fast mode on July 30, 2026. GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.
Cyber models
| Short context | Long context | |||||||
|---|---|---|---|---|---|---|---|---|
| Model | Input | Cached input | Cache writes | Output | Input | Cached input | Cache writes | Output |
| gpt-5.6-sol | $4.00 | $0.40 | $5.00 | $20.00 | $8.00 | $0.80 | $10.00 | $30.00 |
| gpt-5.6-cyber | $12.50 | $1.25 | $15.625 | $75.00 | - | - | - | - |
Multimodal models
GPT-Live sessions
GPT-Live 1 voice sessions are billed per second, without rounding up to a whole minute. Backend model and tool usage is charged separately.
| Model | Price per minute |
|---|---|
| gpt-live-1 | $0.05 |
Realtime and audio generation models
| Model | Modality | Input | Cached input | Output / cost |
|---|---|---|---|---|
| gpt-realtime-2.1 | Audio | $32.00 | $0.40 | $64.00 |
| Text | $4.00 | $0.40 | $24.00 | |
| Image | $5.00 | $0.50 | - | |
| gpt-realtime-2.1-mini | Audio | $10.00 | $0.30 | $20.00 |
| Text | $0.60 | $0.06 | $2.40 | |
| Image | $0.80 | $0.08 | - |
Image generation models
| Model | Modality | Input | Cached input | Output |
|---|---|---|---|---|
| gpt-image-2.5-sunburst | Image | $8.00 | $2.00 | $30.00 |
| Text | $5.00 | $1.25 | - | |
| gpt-image-2.5-flare | Image | $8.00 | $2.00 | $30.00 |
| Text | $5.00 | $1.25 | - |
| Model | Modality | Input | Cached input | Output |
|---|---|---|---|---|
| gpt-image-2 | Image | $4.00 | $1.00 | $15.00 |
| Text | $2.50 | $0.625 | - |
Transcription models
| Model | Use case | Input | Output | Estimated cost |
|---|---|---|---|---|
| gpt-realtime-translate | Live translation | - | - | $0.034 / minute |
| gpt-live-transcribe | Live transcription | - | - | $0.017 / minute |
| gpt-realtime-whisper | Live transcription | - | - | $0.017 / minute |
| gpt-transcribe | Transcription | - | - | $0.0045 / minute |
| gpt-4o-transcribe | Transcription | $2.50 | $10.00 | $0.006 / minute |
| gpt-4o-mini-transcribe | Transcription | $1.25 | $5.00 | $0.003 / minute |
Tools
| Tool | Details | Pricing |
|---|---|---|
| Web search | Web search (all models) | $10.00 / 1k calls + Search content tokens billed at model rates. |
| Image Web search (all models) | $10.00 / 1k calls + Search content tokens billed at model rates. | |
Web search preview (reasoning models, including gpt-5, o-series) | $10.00 / 1k calls + Search content tokens billed at model rates. | |
| Web search preview (non-reasoning models) | $25.00 / 1k calls + Search content tokens are free. | |
| Containers | Hosted Shell and Code Interpreter | 1 GB $0.03, 4 GB $0.12, 16 GB $0.48, 64 GB $1.92 per 20-minute session per container. |
| File search | Storage | $0.10 / GB per day (1 GB free) |
| Tool call | $2.50 / 1k calls | |
| Agent Kit | ChatKit file and image upload storage | $0.10 / GB-day after 1 GB free per account per month |
Specialized models
| Category | Model | Input | Cached input | Output |
|---|---|---|---|---|
| ChatGPT | chat-latest | $5.00 | $0.50 | $30.00 |
| Codex | gpt-5.3-codex | $1.75 | $0.175 | $14.00 |
| Life Sciences | gpt-rosalind-research | $5.00 | $0.50 | $25.00 |
gpt-rosalind-research begins on October 5, 2026. Cache-write pricing does not apply to this model. Access is limited to approved internal research through the trusted-access program. All eligible organizations will continue to get access to the latest GPT-Rosalind models as they’re released.Regional processing (data residency) endpoints are charged a 10% uplift for models released on or after March 5, 2026, that are eligible for data residency. See our Your data guide for supported regions and processing details.
| Category | Model | Input | Cached input | Output |
|---|---|---|---|---|
| Codex | gpt-5.3-codex | $3.50 | $0.35 | $28.00 |
Finetuning
OpenAI is winding down the fine-tuning platform. The platform is no longer accessible to new users, but existing users of the fine-tuning platform will be able to create training jobs for the coming months.
All fine-tuned models will remain available for inference until their base models are deprecated. The full timeline is here.
| Model | Training | Input | Cached input | Output |
|---|---|---|---|---|
| o4-mini-2025-04-16 | $100.00 / hour | $4.00 | $1.00 | $16.00 |
| o4-mini-2025-04-16 with data sharing | $100.00 / hour | $2.00 | $0.50 | $8.00 |
| Model | Training | Input | Cached input | Output |
|---|---|---|---|---|
| o4-mini-2025-04-16 | $100.00 / hour | $2.00 | $0.50 | $8.00 |
| o4-mini-2025-04-16 with data sharing | $100.00 / hour | $1.00 | $0.25 | $4.00 |
Cloud platforms
OpenAI models on Amazon Bedrock and Microsoft Azure are billed through those services.