Explore the interactive model lab
AI model API pricing and workload costs
Current-price snapshot, not historical benchmark-run cost. Include billed reasoning tokens and check the linked provider page before deployment. Tools, media, retries, caching and regional charges are excluded.
| Model | Input | Billed output |
|---|---|---|
| GPT-5.2gpt-5.2 | $1.75 | $14.00 |
| Gemini 3.1 Pro Previewgemini-3.1-pro-preview | $2.00 | $12.00 |
| Claude Opus 4.8claude-opus-4-8 | $5.00 | $25.00 |
| Claude Sonnet 5.5claude-sonnet-5-5 | $2.00 | $10.00 |
| Gemini 3 Flash Previewgemini-3-flash-preview | $0.50 | $3.00 |
| Gemini 3.8 Flashgemini-3.8-flash | $0.75 | $3.75 |
- GPT-5.2. previous flagship. Output limit: 128,000 tokens; includes reasoning.
- Gemini 3.1 Pro Preview. preview. Shown tier applies to prompts up to 200,000 input tokens. Output limit: 65,536 tokens; includes reasoning.
- Claude Opus 4.8. active legacy. Output limit: 128,000 tokens; includes reasoning.
- Claude Sonnet 5.5. active latest. No comparable score selected; cost-only until exact compatible evaluation data is verified. Output limit: 128,000 tokens; includes reasoning.
- Gemini 3 Flash Preview. preview, legacy Flash. Output limit: 65,536 tokens; includes reasoning.
- Gemini 3.8 Flash. stable latest Flash. No comparable score selected; cost-only until exact compatible evaluation data is verified. Shown prices valid through 2026-12-31; announced Jan 1, 2027 rates: $1.50 input / $7.50 output per million tokens. Output limit: 65,536 tokens; includes reasoning.