Anthropic Claude API
The Claude API gives you the best model on the market: Claude Opus 5.5, #1 on the Artificial Analysis Intelligence Index at 58. The lineup is simple. Haiku 4.5 costs $1 / $5 per million tokens, Sonnet 5 costs $2 / $10, Opus 5.5 costs $4 / $20 and Fable 5.1 costs $10 / $50. Every current model except Haiku has a 1M-token context and 128K output.
The developer features are strong: prompt caching (reads cost as little as 2.5% of the input price on Fable), a 50% Batch discount, an effort setting to trade quality for cost, and the same model IDs on AWS, Google Cloud and Microsoft Foundry. A US-only inference option costs 10% extra.
The weak spot is the budget end. Haiku 4.5 is almost a year old, so there is no cheap Claude to rival GPT-6 Luna or Gemini 2.5 Flash-Lite. Pick it if quality on coding and agent tasks matters most. Skip it if you need the lowest cost per token at huge volume.
Score breakdown
Key facts
- Pricing
- $1 / $5 per 1M tokens (Haiku 4.5) (Sonnet 5 $2 / $10; Opus 5.5 $4 / $20; Fable 5.1 $10 / $50. Batch 50% off. Cache reads 2.5–10% of input price. US-only inference at 1.1x.)
- Free option
- No
- Platforms
- API, AWS Bedrock, Google Cloud, Microsoft Foundry, Claude Platform on AWS
- Flagship
- Claude Opus 5.5 (AA index 58, #1)
- Context
- 1M tokens on Opus 5.5, Fable 5.1 and Sonnet 5
- Max output
- 128K sync; up to 300K on Batch (beta)
- Clouds
- Claude API, Bedrock, Google Cloud, Microsoft Foundry
What we like
- Home of the top-ranked model, Claude Opus 5.5
- Same models on AWS, Google Cloud and Microsoft Foundry
- Deep caching and Batch discounts
- Clear model lineup with long retirement notice
Watch out for
- No modern budget model; Haiku 4.5 dates from October 2025
- Text and image in, text out only (no native audio or image output)
- Opus models use many output tokens