モデル詳細
gpt-3.5-turbo-16k
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up to Sep 2021.
モデル仕様
- コンテキスト長
- 16.385K
- 最大出力
- 4.096K
- 入出力モダリティ
- テキスト
- リリース
- 2023-08
02
API エンドポイント
MixRoute 単一ゲートウェイ、OpenAI 互換
-
OpenAI-compatible
/v1/chat/completionsPOST
03
料金
固定料金、単位: /1M Tokens
入力
$3.0000 /1M Tokens
生成
$6.0000 /1M Tokens
05
選定サマリー
ワークロード適合性を素早く判断
What is this model good for?
GPT-3.5 Turbo with extended 16K context
Use gpt-3.5-turbo-16k for tasks requiring longer context, supporting approximately 20 pages of text in a single request. It offers four times the context length of standard GPT-3.5 Turbo.
What should you check before using it?
Consider newer models with larger context windows
Newer models offer significantly larger context windows (up to 1M tokens). Test your long-context tasks to determine if this 16K variant still meets your needs.
Why use it through MixRoute?
Use the OpenAI-compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-3.5-turbo-16k through MixRoute.
08
トークンコスト見積もり
このページの料金に基づく見積もりで、実請求ではありません
このモデルにはキャッシュ読み取り価格がないため、キャッシュ料金は計算されません
単一エンドポイントで検証可能な判断
同じリクエスト形式でこのモデルと代替ルートをテスト
実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。