For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending .md to the page URL.
Primary navigation
Compare models
Near-Astra performance for complex work at a lower cost.
Reasoning
Speed
Input
Output
Reasoning tokens
Pricing
Text tokens
Input / 1M tokens
$2.00
Cached input / 1M tokens
$0.10
Cache writes / 1M tokens
$2.50
Output / 1M tokens
$10.00
Context
Window
1,050,000
Max Output Tokens
128,000
Knowledge Cutoff
Apr 30, 2026
Endpoints
v1/chat/completions
v1/responses
v1/batch
Supported Features
Streaming
Function calling
Structured outputs
Image input
Rate Limits
TPM
Free
-
Tier 1
500,000
Tier 2
1,000,000
Tier 3
2,000,000
Tier 4
4,000,000
Tier 5
40,000,000
Built to power complex coding and agentic workflows.
Reasoning
Speed
Input
Output
Reasoning tokens
Pricing
Text tokens
Input / 1M tokens
$2.00
Cached input / 1M tokens
$0.20
Cache writes / 1M tokens
$2.50
Output / 1M tokens
$10.00
Context
Window
1,050,000
Max Output Tokens
128,000
Knowledge Cutoff
Apr 20, 2026
Endpoints
v1/chat/completions
v1/responses
v1/batch
Supported Features
Streaming
Function calling
Structured outputs
Image input
Rate Limits
TPM
Free
-
Tier 1
500,000
Tier 2
1,000,000
Tier 3
2,000,000
Tier 4
4,000,000
Tier 5
40,000,000