Price per million tokens
| Model | Input /1M | Cached /1M | Output /1M | Context |
|---|---|---|---|---|
| Claude Fable 5.1 | $10.00 | $0.25 | $50.00 | 1,000,000 |
| Claude Opus 5.5 | $4.00 | $0.20 | $20.00 | 1,000,000 |
| Claude Sonnet 5.5 | $2.00 | $0.20 | $10.00 | 1,000,000 |
| Claude Haiku 4.5 | $1.00 | $0.10 | $5.00 | 200,000 |
| GPT-5.6 Sol | $5.00 | $0.50 | $30.00 | 1,050,000 |
| GPT-5.6 Terra | $2.00 | $0.20 | $12.00 | 1,050,000 |
| GPT-5.6 Luna | $0.20 | $0.0200 | $1.20 | 1,050,000 |
How to choose
Price per token isn't the whole story: a pricier model may solve in one call what a cheap one needs three for, and tokenizers differ. Claude models from 4.7 on use a tokenizer that produces about 30% more tokens than the previous one; the token counter shows the difference.
- High volume, simple tasks: GPT-5.6 Luna and Claude Haiku 4.5.
- General product use (chat, support, RAG): Claude Sonnet 5.5 and GPT-5.6 Terra.
- Code, agents and long reasoning: Claude Opus 5.5, Claude Fable 5.1 and GPT-5.6 Sol.
Long context
Claude models from 4.6 on include a 1M-token context at standard pricing. On GPT-5.6, requests above 272K input tokens pay 2× input and 1.5× output for the whole request.