For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending .md to the page URL.
Primary navigation
Compare models
Our most capable model for the most demanding work.
Reasoning
Speed
Input
Output
Reasoning tokens
Pricing
Text tokens
Input / 1M tokens
$10.00
Cached input / 1M tokens
$1.00
Cache writes / 1M tokens
$12.50
Output / 1M tokens
$50.00
Context
Window
1,050,000
Max Output Tokens
128,000
Knowledge Cutoff
Apr 30, 2026
Endpoints
v1/chat/completions
v1/responses
v1/batch
Supported Features
Streaming
Function calling
Structured outputs
Image input
Rate Limits
TPM
Free
-
Build
1,000,000
Launch
4,000,000
Grow
40,000,000
Near-Astra performance for complex work at a lower cost.
Reasoning
Speed
Input
Output
Reasoning tokens
Pricing
Text tokens
Input / 1M tokens
$2.00
Cached input / 1M tokens
$0.10
Cache writes / 1M tokens
$2.50
Output / 1M tokens
$10.00
Context
Window
1,050,000
Max Output Tokens
128,000
Knowledge Cutoff
Apr 30, 2026
Endpoints
v1/chat/completions
v1/responses
v1/batch
Supported Features
Streaming
Function calling
Structured outputs
Image input
Rate Limits
TPM
Free
-
Build
1,000,000
Launch
4,000,000
Grow
40,000,000