コンテンツへスキップ

モデル詳細

gpt-4.1-nano-2025-04-14

OpenAI 提供
従量課金

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million token context window, and scores 80.1% on MMLU, 50.3% on GPQA, and 9.8% on Aider polyglot coding – even higher than GPT‑4o mini. It’s ideal for tasks like classification or autocompletion.

モデル仕様

コンテキスト長
1.047576M
最大出力
32.768K
入出力モダリティ
テキスト / 画像
リリース
2025-04

02

API エンドポイント

MixRoute 単一ゲートウェイ、OpenAI 互換

  • OpenAI-compatible /v1/chat/completions POST

03

料金

固定料金、単位: /1M Tokens

入力

$0.1000 /1M Tokens

生成

$0.2000 /1M Tokens

05

選定サマリー

ワークロード適合性を素早く判断

What is this model good for?

Ultra-fast classification and autocompletion at scale

Use GPT-4.1 nano for high-volume tasks like classification, autocompletion, and data extraction. It delivers the 1M token context window at the lowest cost and latency in the GPT-4.1 series.

What should you check before using it?

Confirm accuracy for your task class

GPT-4.1 nano prioritizes speed and cost over deep reasoning. Test your classification and autocompletion tasks to confirm the accuracy meets your requirements before production deployment.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-4.1-nano through MixRoute, then run the same integration tests used for the current client.

08

トークンコスト見積もり

このページの料金に基づく見積もりで、実請求ではありません

このモデルにはキャッシュ読み取り価格がないため、キャッシュ料金は計算されません

単一エンドポイントで検証可能な判断

同じリクエスト形式でこのモデルと代替ルートをテスト

実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。