模型详情
gpt-3.5-turbo-16k
This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up to Sep 2021.
模型规格
- 上下文长度
- 16.385K
- 最大输出
- 4.096K
- 输入/输出模态
- 文本
- 发布日期
- 2023-08
02
API 端点
单一 MixRoute 网关,OpenAI 兼容
-
OpenAI-compatible
/v1/chat/completionsPOST
03
价格
固定费率,单位:/1M Tokens
输入
$3.0000 /1M Tokens
补全
$6.0000 /1M Tokens
05
选型摘要
快速判断是否适合你的工作负载
What is this model good for?
GPT-3.5 Turbo with extended 16K context
Use gpt-3.5-turbo-16k for tasks requiring longer context, supporting approximately 20 pages of text in a single request. It offers four times the context length of standard GPT-3.5 Turbo.
What should you check before using it?
Consider newer models with larger context windows
Newer models offer significantly larger context windows (up to 1M tokens). Test your long-context tasks to determine if this 16K variant still meets your needs.
Why use it through MixRoute?
Use the OpenAI-compatible endpoint with stable model ID
Use the confirmed compatible endpoint with model ID gpt-3.5-turbo-16k through MixRoute.
08
Token 成本估算
基于本页价格的实时估算,非实际账单
此模型没有缓存读取价格,不计算缓存费用
单一端点,可验证的决策
使用相同的请求格式测试此模型和备用路由
从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。