| Input length | Input / 1M | Output / 1M | Cache hit / 1M |
|---|---|---|---|
| — - 128k | $0.03 | $0.29 | $0.01 |
| 128k - 256k | $0.12 | $1.15 | $0.02 |
| 256k - 1M | $0.18 | $1.72 | $0.02 |
Qwen3.5 Flash 在 NextModel 中是什麼?
台灣團隊可用的 NextModel 目錄中的 Qwen 模型,常用於 General chat via Qwen、API workloads 工作負載。當前展示價格為 Starting at $0.03 / 1M tokens、輸出 $0.29 / 1M tokens,上下文視窗為 — token。
適用場景
- General chat via Qwen
- API workloads
OpenAI 相容呼叫範例
保持 OpenAI SDK 呼叫方式不變,把 base_url 改為 NextModel,並使用模型目錄 ID qwen--qwen3.5-flash。
Python
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://api.nextmodel.app/v1"
)
resp = client.chat.completions.create(
model="qwen/qwen3.5-flash",
messages=[{"role": "user", "content": "Hello from NextModel"}]
)
print(resp.choices[0].message.content)相似替代項
Qwen目錄
Available through the NextModel gateway via Qwen.
適用場景General chat via Qwen, API workloads
路由已設定
串流輸出
平台整理NextModel gateway catalog (Go origin)
Qwen目錄
Available through the NextModel gateway via Qwen.
適用場景General chat via Qwen, API workloads
路由已設定
串流輸出
平台整理NextModel gateway catalog (Go origin)
Qwen目錄
Available through the NextModel gateway via Qwen.
適用場景General chat via Qwen, API workloads
路由已設定
串流輸出
平台整理NextModel gateway catalog (Go origin)
參考
Qwen3.5 Flash 相關文章
上線前如何估算 AI API 成本
Listed price only tells you the unit cost. This walkthrough turns that into a traffic estimate.
Lin Qiming
LLM Gateway: Compare Options and Alternatives
Same OpenAI-compatible base URL, with routing and billing sitting in front of the model.
Zhao Yu'an
常見問題