Processing...Please wait while we secure this action
Qwen model

Qwen3.5 Flash API

qwen/qwen3.5-flash has an input price of $0.03 per 1M tokens and an output price of $0.29 per 1M tokens. Latency band is 1000-3000ms. It has streaming and the use case "quality".

QwenPlatform curatedCatalog
Streaming
Input priceStarting at $0.03 / 1M tokens
Output price$0.29 / 1M tokens
Context length— tokens
Max output8.2k tokens
Input lengthInput / 1MOutput / 1MCache hit / 1M
— - 128k$0.03$0.29$0.01
128k - 256k$0.12$1.15$0.02
256k - 1M$0.18$1.72$0.02

What is Qwen3.5 Flash API in NextModel?

qwen/qwen3.5-flash has an input price of $0.03 per 1M tokens and an output price of $0.29 per 1M tokens. Latency band is 1000-3000ms. It has streaming and the use case "quality".

Best use cases

  • Call this id if you want a low-input-price Qwen model, need streaming, and can proceed without a published context window or first-party CacheSafety score.
  • You need the listed capabilities on this catalog entry.
  • The listed price fits the workload better than the siblings in the same family.

When not to use

  • Do not call it if your decision requires a published context window or a published first-party CacheSafety benchmark.
  • Both are missing for qwen/qwen3.5-flash.
  • A required number is unpublished in the catalog snapshot.

OpenAI-compatible code example

Keep the OpenAI SDK style, set base_url to NextModel, and use the catalog model ID qwen--qwen3.5-flash.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="qwen/qwen3.5-flash",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Similar alternatives

QwenCatalog

Available through the NextModel gateway via Qwen.

Starting at $0.12 / 1M tokensInput$0.69 / 1M tokensOutputContext
Best forGeneral chat via Qwen, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
QwenCatalog

Available through the NextModel gateway via Qwen.

Starting at $1.20 / 1M tokensInput$6 / 1M tokensOutputContext
Best forGeneral chat via Qwen, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
QwenCatalog

Available through the NextModel gateway via Qwen.

Starting at $0.05 / 1M tokensInput$0.4 / 1M tokensOutputContext
Best forGeneral chat via Qwen, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

Reading

Articles that explain Qwen3.5 Flash

FAQ

Qwen3.5 Flash API questions

What model id do I send for Qwen3.5 Flash?

Send qwen/qwen3.5-flash to https://api.nextmodel.app/v1 with your OpenAI SDK. If that id is missing from /v1/models, the page is stale.

What does the catalog list for price and context?

Qwen3.5 Flash is listed at $0.03 / 1M tokens input and $0.29 / 1M tokens output. Context window: not published in catalog.

What should I treat as unverified?

Do not call it if your decision requires a published context window or a published first-party CacheSafety benchmark. Both are missing for qwen/qwen3.5-flash.