For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending .md to the page URL.
Primary navigation
gpt-6-sol
GPT-6 Sol
Built to power complex coding and agentic workflows.
Reasoning
Speed
Price
$2•$10
Input
Output

GPT-6 Sol is built for complex coding and agentic workflows.

reasoning.effort supports none, low, medium (default), high, xhigh, and max. Use the Responses API for built-in tools and function calling. Chat Completions supports function calling only with reasoning_effort set to none.

EU data residency is available only with Standard processing. See data residency eligibility.

1,050,000 context window
128,000 max output tokens
Apr 20, 2026 knowledge cutoff
Reasoning token support
Pricing
Pricing is based on the number of tokens used, or other metrics based on the model type. For tool-specific models, like search and computer use, there’s a fee per tool call. See details in the pricing page.
Text tokens
Per 1M tokens
Input
$2.00
Cached input
$0.20
Cache writes
$2.50
Output
$10.00

Cached input tokens are priced at 10% of the uncached input token rate.

Cache writes are billed at 1.25x the uncached input token rate.

Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.

Regional processing adds a 10% premium where available. EU data residency is available only with Standard processing.

Batch and Flex are priced at 50% of Standard rates. Fast mode is priced at 2x the applicable rates.

Modalities
Text
Input and output
Image
Input only
Audio
Not supported
Video
Not supported
Endpoints
Live
v1/live/sessions
Chat Completions
v1/chat/completions
Responses
v1/responses
Realtime
v1/realtime
Realtime translation
v1/realtime/translations
Realtime transcription
v1/realtime/transcription_sessions
Assistants
v1/assistants
Batch
v1/batch
Fine-tuning
v1/fine-tuning
Embeddings
v1/embeddings
Image generation
v1/images/generations
Videos
v1/videos
Image edit
v1/images/edits
Speech generation
v1/audio/speech
Transcription
v1/audio/transcriptions
Translation
v1/audio/translations
Moderation
v1/moderations
Completions (legacy)
v1/completions
Features
Streaming
Supported
Function calling
Supported
Structured outputs
Supported
Fine-tuning
Not supported
Tools
Tools supported by this model when using the Responses API.
Web search
Supported
File search
Supported
Image generation
Supported
Code interpreter
Supported
Hosted shell
Supported
Apply patch
Supported
Skills
Supported
Computer use
Supported
MCP
Supported
Tool search
Supported
Snapshots

Use gpt-6-sol in your API requests.

gpt-6-sol
gpt-6-sol
gpt-6-sol
gpt-6-sol
Rate limits
Rate limits ensure fair and reliable access to the API by placing specific caps on requests, tokens, audio duration, or other usage within a given time period. Your usage tier determines how high these limits are set and automatically increases as you send more requests and spend more on the API.
TierRPMTPMBatch queue limit
FreeNot supported
Tier 1500500,0001,500,000
Tier 25,0001,000,0003,000,000
Tier 35,0002,000,000100,000,000
Tier 410,0004,000,000200,000,000
Tier 515,00040,000,00015,000,000,000