AI Model Pricing
32 models across 9 providers
| Model | Input /1M ⇅ | Output /1M ⇅ |
|---|---|---|
| $10.00 | $50.00 | |
| $5.00 | $25.00 | |
| $3.00 | $15.00 | |
| $1.00 | $5.00 | |
| Free | Free | |
| $5.00 | $30.00 | |
| $2.50 | $15.00 | |
| $0.75 | $4.50 | |
| $0.20 | $1.25 | |
| $2.00 | $8.00 | |
| $1.10 | $4.40 | |
| $1.50 | $9.00 | |
| $2.00 | $12.00 | |
| $0.25 | $1.50 | |
| $1.25 | $10.00 | |
| $0.30 | $2.50 | |
| $0.20 | $0.80 | |
| $0.10 | $0.30 | |
| $0.13 | $0.40 | |
| $1.50 | $7.50 | |
| $0.50 | $1.50 | |
| $0.15 | $0.60 | |
| $0.43 | $0.87 | |
| $0.14 | $0.28 | |
| $0.50 | $2.15 | |
| $0.27 | $1.12 | |
| $1.48 | $4.42 | |
| $0.32 | $1.28 | |
| $1.25 | $2.50 | |
| $1.25 | $2.50 | |
| $2.50 | $10.00 | |
| $0.15 | $0.60 |
↓ = price drop since last update · Value score = context / combined cost (higher = more tokens per dollar)
32 models
Cost Reduction Strategies
Prompt Caching
Cache system prompts with Claude and cut re-use costs by up to 90%. Write once, pay ~10% per cached hit. Critical for any high-volume agent.
Read docs →Model Routing
Route simple tasks (classification, intent, formatting) to fast/cheap models. Reserve frontier inference for generation and reasoning. 60–80% cost reduction typical.
Batch API
Anthropic and OpenAI both offer 50% off for async batch requests. Offline eval runs, bulk tagging, document processing — always batch these.
Read docs →Output Tokens
Output tokens cost 3–5× more than input. Tell the model to be concise. Structured JSON output with strict schemas eliminates filler and preamble.
Context Window
Every token in context costs money on every call. Summarize conversation history. Use RAG to retrieve only relevant chunks instead of full documents.
Verify at official pricing pages