コンテンツへスキップ

モデル詳細

claude-haiku-4-5-20251001

Anthropic 提供
従量課金

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications. It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.

モデル仕様

コンテキスト長
200K
最大出力
64K
入出力モダリティ
テキスト / 画像
リリース
2025-10

02

API エンドポイント

MixRoute 単一ゲートウェイ、OpenAI 互換

  • Anthropic Messages /v1/messages POST
  • OpenAI-compatible /v1/chat/completions POST

03

料金

固定料金、単位: /1M Tokens

入力

$1.0000 /1M Tokens

生成

$5.0000 /1M Tokens

キャッシュ読み取り

$0.1000 /1M Tokens

キャッシュ作成

$1.2500 /1M Tokens

05

選定サマリー

ワークロード適合性を素早く判断

What is this model good for?

Fast, efficient frontier intelligence for real-time apps

Use Claude Haiku 4.5 for real-time and high-volume applications. It delivers near-frontier intelligence at a fraction of the cost and latency, scoring >73% on SWE-bench Verified with extended thinking support.

What should you check before using it?

Test extended thinking and tool-assisted workflows

Haiku 4.5 introduces extended thinking with controllable reasoning depth. Test your coding, bash, web search, and computer-use tools to confirm they work as expected.

Why use it through MixRoute?

Use the compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID claude-haiku-4-5 through MixRoute.

08

トークンコスト見積もり

このページの料金に基づく見積もりで、実請求ではありません

同一プレフィックスが繰り返し読み取られる入力の割合(最大100%)

単一エンドポイントで検証可能な判断

同じリクエスト形式でこのモデルと代替ルートをテスト

実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。