Model detail
claude-haiku-4-5
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications. It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.
Model specs
- Context length
- 200K
- Max output
- 64K
- I/O modalities
- Text / Image
- Released
- 2025-10
02
API endpoints
One MixRoute gateway, OpenAI-compatible
-
Anthropic Messages
/v1/messagesPOST -
OpenAI-compatible
/v1/chat/completionsPOST
03
Pricing
Flat rate, unit: /1M Tokens
Input
$1.0000 /1M Tokens
Completion
$5.0000 /1M Tokens
Cache read
$0.1000 /1M Tokens
Cache creation
$1.2500 /1M Tokens
05
Selection summary
Quickly judge whether it fits your workload
What is this model good for?
Fast, efficient frontier intelligence for real-time apps
Use Claude Haiku 4.5 for real-time and high-volume applications. It delivers near-frontier intelligence at a fraction of the cost and latency, scoring >73% on SWE-bench Verified with extended thinking support.
What should you check before using it?
Test extended thinking and tool-assisted workflows
Haiku 4.5 introduces extended thinking with controllable reasoning depth. Test your coding, bash, web search, and computer-use tools to confirm they work as expected.
Why use it through MixRoute?
Use the compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID claude-haiku-4-5 through MixRoute.
08
Token cost estimator
Live estimate from this page's pricing, not an actual bill
Share of the same prefix read repeatedly, up to 100%
09
FAQ
Use the model ID claude-haiku-4-5 through the compatible chat completions endpoint. It supports extended thinking with controllable reasoning depth.
One endpoint, a testable decision
Test this model and alternate routes with the same request format
Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.