Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
125 / 125 个模型
每张卡片都标好来源,也能直接复制成 OpenAI 兼容调用。
这里能一次看清什么?
NextModel 模型市场会同时展示提供方、输入价格、输出价格、上下文长度、延迟估算、能力、使用场景、可用性、路由状态和来源标签,帮助团队在正式导流前收敛模型候选。
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Available through the NextModel gateway via DeepSeek.
Available through the NextModel gateway via DeepSeek.
Available through the NextModel gateway via DeepSeek.
Available through the NextModel gateway via Volcengine.
Available through the NextModel gateway via Volcengine.
Doubao Seed 2.0 Mini 是目前通过 NextModel 公共网关暴露的最低成本生产模型。它适合作为中文问答、分类、摘要和轻量多模态任务的默认选择。
Available through the NextModel gateway via Volcengine.
Available through the NextModel gateway via Volcengine.
Available through the NextModel gateway via Volcengine.
Available through the NextModel gateway via Doubao.
Available through the NextModel gateway via Doubao.
Available through the NextModel gateway via Doubao.
Available through the NextModel gateway via Doubao.
Available through the NextModel gateway via Doubao.
Available through the NextModel gateway via Doubao.
Available through the NextModel gateway via Dreamina.
Available through the NextModel gateway via Dreamina.
Available through the NextModel gateway via Dreamina.
Available through the NextModel gateway via Gemini.
Available through the NextModel gateway via Gemini.
Available through the NextModel gateway via Gemini.
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Available through the NextModel gateway via Kling.
Available through the NextModel gateway via Kling.
Available through the NextModel gateway via Kling.
Available through the NextModel gateway via Kling.
Available through the NextModel gateway via Kling.
Available through the NextModel gateway via Kling.
Available through the NextModel gateway via MiniMax.
Available through the NextModel gateway via Moonshotai.
Available through the NextModel gateway via Moonshotai.
Available through the NextModel gateway via Moonshotai.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Available through the NextModel gateway via OpenAI.
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic...
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
GPT Chat Latest points to OpenAI's stable API alias `chat-latest` that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates...
Available through the NextModel gateway via OpenAI.
Available through the NextModel gateway via OpenAI.
Available through the NextModel gateway via OpenAI.
Available through the NextModel gateway via OpenAI.
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Qwen.
Available through the NextModel gateway via Vidu.
Available through the NextModel gateway via Vidu.
Available through the NextModel gateway via X Ai.
Available through the NextModel gateway via X Ai.
Available through the NextModel gateway via X Ai.
Available through the NextModel gateway via X Ai.
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
GLM 5 CN
80Available through the NextModel gateway via Z Ai.
Available through the NextModel gateway via Z Ai.
Available through the NextModel gateway via Z Ai.
决策表
一屏看完价格、上下文、能力、状态与来源。
适合在进入生产测试、成本估算或提供方策略决策前快速缩小候选范围。
| 模型 | 提供方 | 输入 | 输出 | 上下文 | 能力 | 适用场景 | 延迟 | 状态 | 来源 |
|---|---|---|---|---|---|---|---|---|---|
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | Anthropic | $1.45 / 1M tokens | $7.23 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | Anthropic | $0.145 / 1M tokens | $0.723 / 1M tokens | 200k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Opus 4.5anthropic/claude-opus-4.5 | Anthropic | $0.723 / 1M tokens | $3.62 / 1M tokens | 200k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Opus 4.6anthropic/claude-opus-4.6 | Anthropic | $0.723 / 1M tokens | $3.62 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Opus 4.7anthropic/claude-opus-4.7 | Anthropic | $0.723 / 1M tokens | $3.62 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8 | Anthropic | $0.723 / 1M tokens | $3.62 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Claude Opus 5anthropic/claude-opus-5 | Anthropic | $0.723 / 1M tokens | $3.62 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | Anthropic | $0.434 / 1M tokens | $2.17 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Sonnet 4.6anthropic/claude-sonnet-4.6 | Anthropic | $0.434 / 1M tokens | $2.17 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | Anthropic | $0.289 / 1M tokens | $1.45 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Deepseek V3.2 CNdeepseek/deepseek-v3.2-cn | DeepSeek | $0.042 / 1M tokens | $0.064 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Deepseek V4 Flash CNdeepseek/deepseek-v4-flash-cn | DeepSeek | $0.02 / 1M tokens | $0.041 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Deepseek V4 PRO CNdeepseek/deepseek-v4-pro-cn | DeepSeek | $0.239 / 1M tokens | $0.479 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seed 2 0 Codedoubao-seed-2-0-code | Volcengine | $0.069 / 1M tokens | $0.341 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seed 2 0 Litedoubao-seed-2-0-lite | Volcengine | $0.013 / 1M tokens | $0.077 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seed 2 0 Minidoubao-seed-2-0-mini | Volcengine | $0.0043 / 1M tokens | $0.043 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seed 2 0 PROdoubao-seed-2-0-pro | Volcengine | $0.069 / 1M tokens | $0.341 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seed 2 1 PROdoubao-seed-2-1-pro | Volcengine | $0.129 / 1M tokens | $0.639 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seed 2 1 Turbodoubao-seed-2-1-turbo | Volcengine | $0.065 / 1M tokens | $0.32 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seedance 2 0 260128doubao/doubao-seedance-2-0-260128 | Doubao | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seedance 2 0 Fast 260128doubao/doubao-seedance-2-0-fast-260128 | Doubao | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seedance 2 0 Mini 260615doubao/doubao-seedance-2-0-mini-260615 | Doubao | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seedream 4 0 250828doubao/doubao-seedream-4-0-250828 | Doubao | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seedream 4 5 251128doubao/doubao-seedream-4-5-251128 | Doubao | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Doubao Seedream 5 0 260128doubao/doubao-seedream-5-0-260128 | Doubao | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Dreamina Seedance 2 0 260128dreamina/dreamina-seedance-2-0-260128 | Dreamina | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Dreamina Seedance 2 0 Fast 260128dreamina/dreamina-seedance-2-0-fast-260128 | Dreamina | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Dreamina Seedance 2 0 Mini 260615dreamina/dreamina-seedance-2-0-mini-260615 | Dreamina | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Gemini 3 PRO Imagegemini/gemini-3-pro-image | Gemini | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Gemini 3.1 Flash Imagegemini/gemini-3.1-flash-image | Gemini | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Gemini 3.1 Flash Lite Imagegemini/gemini-3.1-flash-lite-image | Gemini | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Google: Gemini 2.5 Flashgoogle/gemini-2.5-flash | $0.043 / 1M tokens | $0.362 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 | |
| Google: Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | $0.072 / 1M tokens | $0.434 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 | |
| Google: Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | $0.036 / 1M tokens | $0.217 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 | |
| Google: Gemini 3.5 Flashgoogle/gemini-3.5-flash | $0.217 / 1M tokens | $1.30 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 | |
| Google: Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | $0.043 / 1M tokens | $0.362 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 | |
| Kling V1 5 CNkling/kling-v1-5-cn | Kling | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Kling V1 6 CNkling/kling-v1-6-cn | Kling | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Kling V2 1 CNkling/kling-v2-1-cn | Kling | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Kling V2 5 Turbo CNkling/kling-v2-5-turbo-cn | Kling | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Kling V2 6 CNkling/kling-v2-6-cn | Kling | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Kling V3 CNkling/kling-v3-cn | Kling | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Minimax M2.5 CNminimax/minimax-m2.5-cn | MiniMax | $0.045 / 1M tokens | $0.177 / 1M tokens | — | 流式输出 | 中文问答, general chat | 1000-3000ms | 目录 | 平台整理 |
| Kimi K2.5 CNmoonshotai/kimi-k2.5-cn | Moonshotai | $0.084 / 1M tokens | $0.437 / 1M tokens | — | 流式输出 | General chat via Moonshotai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Kimi K2.6 CNmoonshotai/kimi-k2.6-cn | Moonshotai | $0.13 / 1M tokens | $0.538 / 1M tokens | — | 流式输出 | General chat via Moonshotai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Kimi K2.7 Code CNmoonshotai/kimi-k2.7-code-cn | Moonshotai | $0.13 / 1M tokens | $0.538 / 1M tokens | — | 流式输出 | General chat via Moonshotai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| MoonshotAI: Kimi K3moonshotai/kimi-k3 | Moonshotai | $0.427 / 1M tokens | $2.13 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-4.1openai/gpt-4.1 | OpenAI | $0.289 / 1M tokens | $1.16 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-4.1 Miniopenai/gpt-4.1-mini | OpenAI | $0.058 / 1M tokens | $0.231 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-4.1 Nanoopenai/gpt-4.1-nano | OpenAI | $0.014 / 1M tokens | $0.058 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-4oopenai/gpt-4o | OpenAI | $0.362 / 1M tokens | $1.45 / 1M tokens | 128k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-4o-miniopenai/gpt-4o-mini | OpenAI | $0.022 / 1M tokens | $0.087 / 1M tokens | 128k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5openai/gpt-5 | OpenAI | $0.181 / 1M tokens | $1.45 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| GPT 5 Codexopenai/gpt-5-codex | OpenAI | $0.181 / 1M tokens | $1.45 / 1M tokens | — | 流式输出 | General chat via OpenAI, API workloads | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5 Miniopenai/gpt-5-mini | OpenAI | $0.036 / 1M tokens | $0.289 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5 Nanoopenai/gpt-5-nano | OpenAI | $0.0072 / 1M tokens | $0.058 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5 Proopenai/gpt-5-pro | OpenAI | $2.17 / 1M tokens | $17.36 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.1openai/gpt-5.1 | OpenAI | $0.181 / 1M tokens | $1.45 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.1-Codexopenai/gpt-5.1-codex | OpenAI | $0.181 / 1M tokens | $1.45 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.1-Codex-Maxopenai/gpt-5.1-codex-max | OpenAI | $0.181 / 1M tokens | $1.45 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.1-Codex-Miniopenai/gpt-5.1-codex-mini | OpenAI | $0.036 / 1M tokens | $0.289 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.2openai/gpt-5.2 | OpenAI | $0.253 / 1M tokens | $2.03 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.2-Codexopenai/gpt-5.2-codex | OpenAI | $0.253 / 1M tokens | $2.03 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.3-Codexopenai/gpt-5.3-codex | OpenAI | $0.253 / 1M tokens | $2.03 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.4openai/gpt-5.4 | OpenAI | $0.362 / 1M tokens | $2.17 / 1M tokens | 1.1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.4 Miniopenai/gpt-5.4-mini | OpenAI | $0.109 / 1M tokens | $0.651 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.4 Nanoopenai/gpt-5.4-nano | OpenAI | $0.029 / 1M tokens | $0.181 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.4 Proopenai/gpt-5.4-pro | OpenAI | $4.34 / 1M tokens | $26.04 / 1M tokens | 1.1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.5openai/gpt-5.5 | OpenAI | $0.723 / 1M tokens | $4.34 / 1M tokens | 1.1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.6 Lunaopenai/gpt-5.6-luna | OpenAI | $0.029 / 1M tokens | $0.174 / 1M tokens | 1.1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.6 Solopenai/gpt-5.6-sol | OpenAI | $0.723 / 1M tokens | $4.34 / 1M tokens | 1.1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT-5.6 Terraopenai/gpt-5.6-terra | OpenAI | $0.289 / 1M tokens | $1.74 / 1M tokens | 1.1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: GPT Chat Latestopenai/gpt-chat-latest | OpenAI | $0.723 / 1M tokens | $4.34 / 1M tokens | 400k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| GPT Image 1openai/gpt-image-1 | OpenAI | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| GPT Image 1 Miniopenai/gpt-image-1-mini | OpenAI | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| GPT Image 1.5openai/gpt-image-1.5 | OpenAI | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| GPT Image 2openai/gpt-image-2 | OpenAI | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: o3openai/o3 | OpenAI | $0.289 / 1M tokens | $1.16 / 1M tokens | 200k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| OpenAI: o4 Miniopenai/o4-mini | OpenAI | $0.159 / 1M tokens | $0.637 / 1M tokens | 200k | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Qwen Image 2.0 CNqwen/qwen-image-2.0-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Qwen Image 2.0 GLBqwen/qwen-image-2.0-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Qwen Image 2.0 PRO CNqwen/qwen-image-2.0-pro-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Qwen Image 2.0 PRO GLBqwen/qwen-image-2.0-pro-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Qwen3 MAX 2026 01 23 GLBqwen/qwen3-max-2026-01-23-glb | Qwen | $0.174 / 1M tokens | $0.868 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3 MAX CNqwen/qwen3-max-cn | Qwen | $0.052 / 1M tokens | $0.208 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3 VL Flash CNqwen/qwen3-vl-flash-cn | Qwen | $0.0043 / 1M tokens | $0.032 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3 VL Flash GLBqwen/qwen3-vl-flash-glb | Qwen | $0.0072 / 1M tokens | $0.058 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3 VL Plus 2025 12 19 GLBqwen/qwen3-vl-plus-2025-12-19-glb | Qwen | $0.029 / 1M tokens | $0.231 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3 VL Plus CNqwen/qwen3-vl-plus-cn | Qwen | $0.022 / 1M tokens | $0.208 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.5 Flash CNqwen/qwen3.5-flash-cn | Qwen | $0.0043 / 1M tokens | $0.042 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.5 Flash GLBqwen/qwen3.5-flash-glb | Qwen | $0.014 / 1M tokens | $0.058 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.5 Plus CNqwen/qwen3.5-plus-cn | Qwen | $0.017 / 1M tokens | $0.1 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.5 Plus GLBqwen/qwen3.5-plus-glb | Qwen | $0.058 / 1M tokens | $0.347 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.6 Flash CNqwen/qwen3.6-flash-cn | Qwen | $0.025 / 1M tokens | $0.143 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.6 Flash GLBqwen/qwen3.6-flash-glb | Qwen | $0.036 / 1M tokens | $0.217 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.6 MAX Preview CNqwen/qwen3.6-max-preview-cn | Qwen | $0.179 / 1M tokens | $1.07 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.6 MAX Preview GLBqwen/qwen3.6-max-preview-glb | Qwen | $0.188 / 1M tokens | $1.13 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.6 Plus CNqwen/qwen3.6-plus-cn | Qwen | $0.041 / 1M tokens | $0.24 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.6 Plus GLBqwen/qwen3.6-plus-glb | Qwen | $0.072 / 1M tokens | $0.434 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.7 MAX CNqwen/qwen3.7-max-cn | Qwen | $0.239 / 1M tokens | $0.718 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Qwen3.7 MAX GLBqwen/qwen3.7-max-glb | Qwen | $0.362 / 1M tokens | $1.09 / 1M tokens | — | 流式输出 | General chat via Qwen, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Wan2.6 T2I CNqwen/wan2.6-t2i-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.6 T2I GLBqwen/wan2.6-t2i-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.6 T2V CNqwen/wan2.6-t2v-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.6 T2V GLBqwen/wan2.6-t2v-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 I2V CNqwen/wan2.7-i2v-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 I2V GLBqwen/wan2.7-i2v-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 Image CNqwen/wan2.7-image-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 Image GLBqwen/wan2.7-image-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 Image PRO CNqwen/wan2.7-image-pro-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 Image PRO GLBqwen/wan2.7-image-pro-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 R2V CNqwen/wan2.7-r2v-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 R2V GLBqwen/wan2.7-r2v-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 T2V CNqwen/wan2.7-t2v-cn | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Wan2.7 T2V GLBqwen/wan2.7-t2v-glb | Qwen | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Viduq3 PRO CNvidu/viduq3-pro-cn | Vidu | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Viduq3 Turbo CNvidu/viduq3-turbo-cn | Vidu | $0 / 1M tokens | $0 / 1M tokens | — | 流式输出视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| Grok 4.1 Fast NON Reasoningx-ai/grok-4.1-fast-non-reasoning | X Ai | $0.029 / 1M tokens | $0.072 / 1M tokens | — | 流式输出 | General chat via X Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Grok 4.1 Fast Reasoningx-ai/grok-4.1-fast-reasoning | X Ai | $0.029 / 1M tokens | $0.072 / 1M tokens | — | 流式输出 | General chat via X Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Grok 4.20 NON Reasoningx-ai/grok-4.20-non-reasoning | X Ai | $0.181 / 1M tokens | $0.362 / 1M tokens | — | 流式输出 | General chat via X Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| Grok 4.20 Reasoningx-ai/grok-4.20-reasoning | X Ai | $0.181 / 1M tokens | $0.362 / 1M tokens | — | 流式输出 | General chat via X Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | X Ai | $0.181 / 1M tokens | $0.362 / 1M tokens | 1M | 流式输出工具调用JSON 模式视觉 | 图像理解, multimodal chat | 1000-3000ms | 目录 | 平台整理 |
| GLM 5 CNz-ai/glm-5-cn | Z Ai | $0.084 / 1M tokens | $0.373 / 1M tokens | — | 流式输出 | General chat via Z Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| GLM 5.1 CNz-ai/glm-5.1-cn | Z Ai | $0.12 / 1M tokens | $0.479 / 1M tokens | — | 流式输出 | General chat via Z Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |
| GLM 5.2 CNz-ai/glm-5.2-cn | Z Ai | $0.171 / 1M tokens | $0.596 / 1M tokens | — | 流式输出 | General chat via Z Ai, API workloads | 1000-3000ms | 目录 | 平台整理 |