コンテンツへスキップ

モデル詳細

gpt-4.1

OpenAI 提供
従量課金

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and GPT-4.5 across coding (54.6% SWE-bench Verified), instruction compliance (87.4% IFEval), and multimodal understanding benchmarks. It is tuned for precise code diffs, agent reliability, and high recall in large document contexts, making it ideal for agents, IDE tooling, and enterprise knowledge retrieval.

モデル仕様

コンテキスト長
1.047576M
最大出力
32.768K
入出力モダリティ
テキスト / 画像
リリース
2025-04

02

API エンドポイント

MixRoute 単一ゲートウェイ、OpenAI 互換

  • OpenAI-compatible /v1/chat/completions POST

03

料金

固定料金、単位: /1M Tokens

入力

$2.0000 /1M Tokens

生成

$4.0000 /1M Tokens

キャッシュ読み取り

$0.5000 /1M Tokens

05

選定サマリー

ワークロード適合性を素早く判断

What is this model good for?

Long-context reasoning and software engineering

Use GPT-4.1 for tasks requiring a 1M token context window, including long-document analysis, complex codebases, and enterprise knowledge retrieval. It excels at instruction following and agent reliability.

What should you check before using it?

Plan for the 1M context window and instruction-following needs

Confirm your workload benefits from the 1M token context window. Test instruction-following accuracy on your specific prompts and verify agent reliability for your most demanding workflows.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-4.1 through MixRoute, then run the same integration tests used for the current client.

08

トークンコスト見積もり

このページの料金に基づく見積もりで、実請求ではありません

同一プレフィックスが繰り返し読み取られる入力の割合(最大100%)

単一エンドポイントで検証可能な判断

同じリクエスト形式でこのモデルと代替ルートをテスト

実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。