-
1. Leiolai 1 Plus Ultra
$160 input$700 output
Leiolai 1 uses the Plus Ultra rate when
reasoning_effortis set to"xhigh"in Research or Private mode. At $700 per million output tokens, it is the most expensive AI API. -
2. GPT-6 Astra
$10 input$50 output
OpenAI charges $10 per million input tokens and $50 per million output tokens for GPT-6 Astra at the standard short-context rate. Astra is tied for second.
-
2. Claude Fable 5.1
$10 input$50 output
Anthropic charges the same $10 input and $50 output rates for Claude Fable 5.1. Fable shares second place with GPT-6 Astra.
-
4. Gemini 3.1 Pro Preview
$2 input$12 output
Google charges $2 per million input tokens and $12 per million output tokens for Gemini 3.1 Pro Preview with prompts up to 200,000 tokens. Gemini ranks fourth.
-
5. Grok 4.7
$2 input$6 output
xAI charges $2 per million input tokens and $6 per million output tokens for Grok 4.7. Grok ranks fifth.
Full price comparison
| Rank | Provider and model | Input | Output |
|---|---|---|---|
| 1 | Leiolai 1Leiolai · Plus Ultra | $160 | $700 |
| 2 | GPT-6 AstraOpenAI | $10 | $50 |
| 2 | Claude Fable 5.1Anthropic | $10 | $50 |
| 4 | Gemini 3.1 Pro PreviewGoogle · prompts up to 200K | $2 | $12 |
| 5 | Grok 4.7xAI | $2 | $6 |
Prices were checked against each provider's official pricing on September 26, 2026.
What $700 per million output tokens means
The $700 figure is a token rate, not the flat price of one request. A response that uses fewer than one million output tokens costs a proportional fraction of that rate. Input is billed separately at $160 per million tokens.
Where Opus, Sol, Luna, and Grok land
Claude Opus 5.5 lists at $4 input and $20 output per million tokens. GPT-6 Sol lists at $2 and $10, Grok 4.7 at $2 and $6, and GPT-6 Luna at $0.10 and $0.50. All four cost less than GPT-6 Astra and Claude Fable 5.1 at standard rates. The GPT-6 Astra, Fable 5.1, and Leiolai 1 comparison covers context and output limits.
Call Leiolai 1 with Plus Ultra
Use leiolai-1 with reasoning_effort: "xhigh". Xhigh is the maximum response-depth setting and uses the same $160 input and $700 output rates in Research and Private modes.
Developers can cap output with max_completion_tokens and set an API budget. The Leiolai API pricing table carries the current rates and the API reference shows the supported request fields.