模型详情
claude-sonnet-4-5
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with improvements across system design, code security, and specification adherence. The model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking. Sonnet 4.5 also introduces stronger agentic capabilities, including improved tool orchestration, speculative parallel execution, and more efficient context and memory management. With enhanced context tracking and awareness of token usage across tool calls, it is particularly well-suited for multi-context and long-running workflows. Use cases span software engineering, cybersecurity, financial analysis, research agents, and other domains requiring sustained reasoning and tool use.
模型规格
- 上下文长度
- 200K
- 最大输出
- 64K
- 输入/输出模态
- 文本 / 图像
- 发布日期
- 2025-09
02
API 端点
单一 MixRoute 网关,OpenAI 兼容
-
Anthropic Messages
/v1/messagesPOST -
OpenAI-compatible
/v1/chat/completionsPOST
03
价格
固定费率,单位:/1M Tokens
输入
$3.0000 /1M Tokens
补全
$15.0000 /1M Tokens
缓存读取
$0.3000 /1M Tokens
缓存创建
$3.7500 /1M Tokens
05
选型摘要
快速判断是否适合你的工作负载
What is this model good for?
Real-world agents and coding with state-of-the-art performance
Use Claude Sonnet 4.5 for software engineering, cybersecurity, financial analysis, and research agents. It delivers state-of-the-art coding performance with improved tool orchestration and context management.
What should you check before using it?
Test extended autonomous operation and tool orchestration
Sonnet 4.5 supports extended autonomous operation with fact-based progress tracking. Test speculative parallel execution and context management for long-running workflows.
Why use it through MixRoute?
Use the compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID claude-sonnet-4-5 through MixRoute.
08
Token 成本估算
基于本页价格的实时估算,非实际账单
同一前缀被重复读取的输入占比,最高 100%
单一端点,可验证的决策
使用相同的请求格式测试此模型和备用路由
从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。