Skip to content

Model detail

claude-haiku-4-5

Provided by Anthropic
Pay-as-you-go

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications. It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.

Model specs

Context length
200K
Max output
64K
I/O modalities
Text / Image
Released
2025-10

02

API endpoints

One MixRoute gateway, OpenAI-compatible

  • Anthropic Messages /v1/messages POST
  • OpenAI-compatible /v1/chat/completions POST

03

Pricing

Flat rate, unit: /1M Tokens

Input

$1.0000 /1M Tokens

Completion

$5.0000 /1M Tokens

Cache read

$0.1000 /1M Tokens

Cache creation

$1.2500 /1M Tokens

05

Selection summary

Quickly judge whether it fits your workload

What is this model good for?

Fast, efficient frontier intelligence for real-time apps

Use Claude Haiku 4.5 for real-time and high-volume applications. It delivers near-frontier intelligence at a fraction of the cost and latency, scoring >73% on SWE-bench Verified with extended thinking support.

What should you check before using it?

Test extended thinking and tool-assisted workflows

Haiku 4.5 introduces extended thinking with controllable reasoning depth. Test your coding, bash, web search, and computer-use tools to confirm they work as expected.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID claude-haiku-4-5 through MixRoute.

08

Token cost estimator

Live estimate from this page's pricing, not an actual bill

Share of the same prefix read repeatedly, up to 100%

09

FAQ

Use the model ID claude-haiku-4-5 through the compatible chat completions endpoint. It supports extended thinking with controllable reasoning depth.

One endpoint, a testable decision

Test this model and alternate routes with the same request format

Start from a real workload, then let quality, total cost, and failure conditions decide whether to send production traffic.