跳至内容

模型详情

gpt-3.5-turbo-16k

由 OpenAI 提供
按量付费

This model offers four times the context length of gpt-3.5-turbo, allowing it to support approximately 20 pages of text in a single request at a higher cost. Training data: up to Sep 2021.

模型规格

上下文长度
16.385K
最大输出
4.096K
输入/输出模态
文本
发布日期
2023-08

02

API 端点

单一 MixRoute 网关,OpenAI 兼容

  • OpenAI-compatible /v1/chat/completions POST

03

价格

固定费率,单位:/1M Tokens

输入

$3.0000 /1M Tokens

补全

$6.0000 /1M Tokens

05

选型摘要

快速判断是否适合你的工作负载

What is this model good for?

GPT-3.5 Turbo with extended 16K context

Use gpt-3.5-turbo-16k for tasks requiring longer context, supporting approximately 20 pages of text in a single request. It offers four times the context length of standard GPT-3.5 Turbo.

What should you check before using it?

Consider newer models with larger context windows

Newer models offer significantly larger context windows (up to 1M tokens). Test your long-context tasks to determine if this 16K variant still meets your needs.

Why use it through MixRoute?

Use the OpenAI-compatible endpoint with stable model ID

Use the confirmed compatible endpoint with model ID gpt-3.5-turbo-16k through MixRoute.

08

Token 成本估算

基于本页价格的实时估算,非实际账单

此模型没有缓存读取价格,不计算缓存费用

单一端点,可验证的决策

使用相同的请求格式测试此模型和备用路由

从真实工作负载开始,再由质量、总成本和故障条件决定是否接入生产流量。