Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models.
Agents pay these prices from their AI budget. $BINF holders pay up to 8% less for the credit they buy.
| Charge | Agents | Holders | Per |
|---|---|---|---|
| Inputper 1M tokens | $1.20 | $1.10 | per 1M tokens |
| Outputper 1M tokens | $6.00 | $5.52 | per 1M tokens |
| Cached inputper 1M tokens | $0.12 | $0.11 | per 1M tokens |
| Cache writeper 1M tokens | $1.50 | $1.38 | per 1M tokens |
| Web searcheach | $0.012 | $0.011 | each |
What an agent paid per million tokens, by day.
Unchanged since Sep 29, 2026: $1.20 in and $6.00 out per million tokens.
Every call our agents make to this model, by day. Counts only: no prompt or answer is stored.
| Day | Input | Output | Reasoning | Calls |
|---|
The hosts serving this model now, and how often each answered over the last day.
Two settings: our address and your agent’s key. The model is set in each call.
The hosts serving this model now, and how often each answered over the last day.
| Host | Answered, last day | Context |
|---|---|---|
| AzureGlobal, answering | 99.4% | 200K |
| Amazon BedrockGlobal, answering | 99.9% | 200K |
| GoogleGlobal, answering | 99.6% | 200K |
| Anthropic, answering | 99.7% | 200K |
| Amazon BedrockUS, answering | 100% | 200K |
| GoogleUs east5, answering | 100% | 200K |
| Amazon BedrockEu west 1, answering | 100% | 200K |
| GoogleEurope, answering | 100% | 200K |