モデル詳細
qwen3.6-flash
従量課金
動的料金
2 ティア
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in above 256K tokens. Prompt caching is supported, with both explicit cache read and cache creation pricing.
モデル仕様
- コンテキスト長
- –
- 最大出力
- –
- 入出力モダリティ
- –
- リリース
- 2026-04
02
API エンドポイント
MixRoute 単一ゲートウェイ、OpenAI 互換
-
OpenAI-compatible
/v1/chat/completionsPOST
03
段階料金
単位: /1M Tokens
| ティア | 入力 /1M Tokens | 出力 /1M Tokens | キャッシュ読み取り /1M Tokens | キャッシュ書き込み /1M Tokens |
|---|---|---|---|---|
| short Length ≤ 256K | $0.2500 | $1.5000 | $0.0250 | $0.3125 |
| long Length > 256K | $1.0000 | $4.0000 | $0.1000 | $1.2500 |
08
トークンコスト見積もり
このページの料金に基づく見積もりで、実請求ではありません
同一プレフィックスが繰り返し読み取られる入力の割合(最大100%)
単一エンドポイントで検証可能な判断
同じリクエスト形式でこのモデルと代替ルートをテスト
実際のワークロードから始め、品質、総コスト、障害条件を基に本番トラフィックを送るか判断します。