Qwen modele

Qwen3 VL Flash GLB

Qwen3 VL Flash GLB Comparez les tarifs API, le fournisseur, la longueur de contexte, les capacites, les cas d'usage, la latence et les alternatives. General chat via Qwen, API workloads. $0.0072 / 1M tokens / $0.058 / 1M tokens. — tokens.

QwenPlatform curatedCatalog
Streaming
Prix entree$0.0072 / 1M tokens
Prix sortie$0.058 / 1M tokens
Longueur du contexte— tokens
Sortie max8.2k tokens
Input lengthInput / 1MOutput / 1MCache hit / 1M
— - 32k$0.0072$0.058$0.0014
32k - 128k$0.012$0.087$0.0029
128k - 256k$0.017$0.139$0.0043

Qu'est-ce que Qwen3 VL Flash GLB dans NextModel ?

Qwen3 VL Flash GLB Comparez les tarifs API, le fournisseur, la longueur de contexte, les capacites, les cas d'usage, la latence et les alternatives. General chat via Qwen, API workloads. $0.0072 / 1M tokens / $0.058 / 1M tokens. — tokens.

Meilleurs cas d'usage

  • General chat via Qwen
  • API workloads

Exemple compatible OpenAI

Conservez le style SDK OpenAI, pointez base_url vers NextModel et utilisez l'ID de catalogue qwen--qwen3-vl-flash-glb.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="qwen/qwen3-vl-flash-glb",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Alternatives proches

QwenCatalog

Available through the NextModel gateway via Qwen.

Starting at $0.174 / 1M tokensInputStarting at $0.868 / 1M tokensOutputContext
Best forGeneral chat via Qwen, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
QwenCatalog

Available through the NextModel gateway via Qwen.

Starting at $0.052 / 1M tokensInputStarting at $0.208 / 1M tokensOutputContext
Best forGeneral chat via Qwen, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
QwenCatalog

Available through the NextModel gateway via Qwen.

Starting at $0.0043 / 1M tokensInputStarting at $0.032 / 1M tokensOutputContext
Best forGeneral chat via Qwen, API workloads
RoutingConfigured
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Qwen3 VL Flash GLB FAQ