GPT-6 AstraNew
GPT-6 Astra via NextModel: 1.1M-token context; Starting at $10 / 1M tokens in / $50 / 1M tokens out per 1M tokens. Capabilities: streaming, tool, json.
Pick by workload first: every card states who supplies the model, what it costs in and out, and how much context you get. Then run your candidates through one OpenAI-compatible endpoint.
129 of 129 models
Three things decide most of your cost: who supplies the model, what it charges in and out, and whether the context window fits your task — every card states all three. Latency, routing status, and the rest are filters.
GPT-6 Astra via NextModel: 1.1M-token context; Starting at $10 / 1M tokens in / $50 / 1M tokens out per 1M tokens. Capabilities: streaming, tool, json.
Qwen3.8-Max via NextModel: 1M-token context; $1.77 / 1M tokens in / $5.30 / 1M tokens out per 1M tokens. Capabilities: streaming.
Claude Fable 5.1 via NextModel: 1M-token context; $10 / 1M tokens in / $50 / 1M tokens out per 1M tokens. Capabilities: streaming, tool, json.
Qwen3.8-Max-GLB via NextModel: 1M-token context; $2 / 1M tokens in / $6 / 1M tokens out per 1M tokens. Capabilities: streaming.
FLUX.2 Flex via NextModel: context unpublished; $0.2 / image in / — out per 1M tokens. Capabilities: streaming, vision.
FLUX.2 PRO via NextModel: context unpublished; $0.075 / image in / — out per 1M tokens. Capabilities: streaming, vision.
GPT Image 1 via NextModel: context unpublished; $0.26 / image in / — out per 1M tokens. Capabilities: streaming, vision.
GPT Image 1 Mini via NextModel: context unpublished; $0.0539 / image in / — out per 1M tokens. Capabilities: streaming, vision.
GPT Image 1.5 via NextModel: context unpublished; $0.218 / image in / — out per 1M tokens. Capabilities: streaming, vision.
GPT Image 2.0 via NextModel: context unpublished; $0.221 / image in / — out per 1M tokens. Capabilities: streaming, vision.
Text Embedding 3 Small via NextModel: context unpublished; $0.02 / 1M tokens in / $0 / 1M tokens out per 1M tokens. Capabilities: streaming.
Text Embedding 3 Large via NextModel: context unpublished; $0.13 / 1M tokens in / $0 / 1M tokens out per 1M tokens. Capabilities: streaming.
GLM 5.3 via NextModel: 1M-token context; $1.40 / 1M tokens in / $4.40 / 1M tokens out per 1M tokens. Capabilities: streaming, tool, json.
Deepseek-V4-Pro-0813 via NextModel: 1M-token context; $1.33 / 1M tokens in / $3.98 / 1M tokens out per 1M tokens. Capabilities: streaming, tool, json.
Doubao Seed 2.0 Lite via NextModel: 256k-token context; Starting at $0.09 / 1M tokens in / $0.53 / 1M tokens out per 1M tokens. Capabilities: streaming.
Doubao Seed 2.0 Pro via NextModel: 256k-token context; Starting at $0.48 / 1M tokens in / $2.36 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3-VL-Flash CN via NextModel: 256k-token context; Starting at $0.03 / 1M tokens in / $0.22 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3.5 Flash via NextModel: 1M-token context; Starting at $0.03 / 1M tokens in / $0.29 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3.5 Plus via NextModel: 1M-token context; Starting at $0.12 / 1M tokens in / $0.69 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3-VL-Plus CN via NextModel: 256k-token context; Starting at $0.15 / 1M tokens in / $1.44 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3.6 Flash GLB via NextModel: 1M-token context; Starting at $0.25 / 1M tokens in / $1.50 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3-Max CN via NextModel: 256k-token context; Starting at $0.36 / 1M tokens in / $1.44 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3-VL-Plus GLB via NextModel: 256k-token context; Starting at $0.2 / 1M tokens in / $1.60 / 1M tokens out per 1M tokens. Capabilities: streaming.
Qwen3.6 Plus GLB via NextModel: 1M-token context; Starting at $0.5 / 1M tokens in / $3 / 1M tokens out per 1M tokens. Capabilities: streaming.
18 video models
| Model | Provider | Input | Output | Context | Capabilities | Best for | Latency | Status | Source |
|---|---|---|---|---|---|---|---|---|---|
| GPT-6 Astraopenai/gpt-6-astra | OpenAI | Starting at $10 / 1M tokens | $50 / 1M tokens | 1.1M | StreamingTool callingJSON modeVision | image understanding, multimodal chat | 2766-6283ms | Catalog | Platform curated |
| Qwen3.8-Maxqwen/qwen3.8-max | Qwen | $1.77 / 1M tokens | $5.30 / 1M tokens | 1M | Streaming | General chat via Qwen, API workloads | 1672-5802ms | Catalog | Platform curated |
| Claude Fable 5.1anthropic/claude-fable-5.1 | Anthropic | $10 / 1M tokens | $50 / 1M tokens | 1M | StreamingTool callingJSON modeVision | image understanding, multimodal chat | 4209-4963ms | Catalog | Platform curated |
| Qwen3.8-Max-GLBqwen/qwen3.8-max-glb | Qwen | $2 / 1M tokens | $6 / 1M tokens | 1M | Streaming | General chat via Qwen, API workloads | 1957-5865ms | Catalog | Platform curated |
| FLUX.2 Flexblack-forest-labs/FLUX.2-flex | Black Forest Labs | $0.2 / image | — | — | StreamingVision | image understanding, multimodal chat | 0-0ms | Catalog | Platform curated |
| FLUX.2 PROblack-forest-labs/FLUX.2-pro | Black Forest Labs | $0.075 / image | — | — | StreamingVision | image understanding, multimodal chat | 0-0ms | Catalog | Platform curated |
| GPT Image 1openai/gpt-image-1 | OpenAI | $0.26 / image | — | — | StreamingVision | image understanding, multimodal chat | 0-0ms | Catalog | Platform curated |
| GPT Image 1 Miniopenai/gpt-image-1-mini | OpenAI | $0.0539 / image | — | — | StreamingVision | image understanding, multimodal chat | 0-0ms | Catalog | Platform curated |
| GPT Image 1.5openai/gpt-image-1.5 | OpenAI | $0.218 / image | — | — | StreamingVision | image understanding, multimodal chat | 0-0ms | Catalog | Platform curated |
| GPT Image 2.0openai/gpt-image-2 | OpenAI | $0.221 / image | — | — | StreamingVision | image understanding, multimodal chat | 0-0ms | Catalog | Platform curated |
| Text Embedding 3 Smallopenai/text-embedding-3-small | OpenAI | $0.02 / 1M tokens | $0 / 1M tokens | — | Streaming | General chat via OpenAI, API workloads | 0-0ms | Catalog | Platform curated |
| Text Embedding 3 Largeopenai/text-embedding-3-large | OpenAI | $0.13 / 1M tokens | $0 / 1M tokens | — | Streaming | General chat via OpenAI, API workloads | 0-0ms | Catalog | Platform curated |
| GLM 5.3z-ai/glm-5.3 | Z Ai | $1.40 / 1M tokens | $4.40 / 1M tokens | 1M | StreamingTool callingJSON modeLong context | General chat via Z Ai, API workloads | 5137-7205ms | Catalog | Platform curated |
| Deepseek-V4-Pro-0813deepseek/deepseek-v4-pro-0813 | DeepSeek | $1.33 / 1M tokens | $3.98 / 1M tokens | 1M | StreamingTool callingJSON modeLong context | Chinese Q&A, general chat | 2075-14930ms | Catalog | Platform curated |
| Doubao Seed 2.0 Litedoubao-seed-2-0-lite | Volcengine | Starting at $0.09 / 1M tokens | $0.53 / 1M tokens | 256k | Streaming | Chinese Q&A, general chat | 19256-25059ms | Catalog | Platform curated |
| Doubao Seed 2.0 Prodoubao-seed-2-0-pro | Volcengine | Starting at $0.48 / 1M tokens | $2.36 / 1M tokens | 256k | Streaming | Chinese Q&A, general chat | 5387-7239ms | Catalog | Platform curated |
| Qwen3-VL-Flash CNqwen/qwen3-vl-flash-cn | Qwen | Starting at $0.03 / 1M tokens | $0.22 / 1M tokens | 256k | Streaming | General chat via Qwen, API workloads | 835-1062ms | Catalog | Platform curated |
| Qwen3.5 Flashqwen/qwen3.5-flash | Qwen | Starting at $0.03 / 1M tokens | $0.29 / 1M tokens | 1M | Streaming | General chat via Qwen, API workloads | 1085-1339ms | Catalog | Platform curated |
| Qwen3.5 Plusqwen/qwen3.5-plus | Qwen | Starting at $0.12 / 1M tokens | $0.69 / 1M tokens | 1M | Streaming | General chat via Qwen, API workloads | 2698-2964ms | Catalog | Platform curated |
| Qwen3-VL-Plus CNqwen/qwen3-vl-plus-cn | Qwen | Starting at $0.15 / 1M tokens | $1.44 / 1M tokens | 256k | Streaming | General chat via Qwen, API workloads | 1076-1257ms | Catalog | Platform curated |
| Qwen3.6 Flash GLBqwen/qwen3.6-flash-glb | Qwen | Starting at $0.25 / 1M tokens | $1.50 / 1M tokens | 1M | Streaming | General chat via Qwen, API workloads | 3643-4402ms | Catalog | Platform curated |
| Qwen3-Max CNqwen/qwen3-max-cn | Qwen | Starting at $0.36 / 1M tokens | $1.44 / 1M tokens | 256k | Streaming | General chat via Qwen, API workloads | 1152-1427ms | Catalog | Platform curated |
| Qwen3-VL-Plus GLBqwen/qwen3-vl-plus-2025-12-19-glb | Qwen | Starting at $0.2 / 1M tokens | $1.60 / 1M tokens | 256k | Streaming | General chat via Qwen, API workloads | 606-670ms | Catalog | Platform curated |
| Qwen3.6 Plus GLBqwen/qwen3.6-plus-glb | Qwen | Starting at $0.5 / 1M tokens | $3 / 1M tokens | 1M | Streaming | General chat via Qwen, API workloads | 5518-6156ms | Catalog | Platform curated |