Comparaison de modeles

Kimi K2 Thinking vs Google Gemini Flash Latest

Compare MoonshotAI: Kimi K2 Thinking (OpenRouter) and Google Gemini Flash Latest (OpenRouter) by price, context, capabilities, and latency.

Lequel choisir ?

  • Price: MoonshotAI: Kimi K2 Thinking is cheaper ($0.114 / 1M tokens input / $0.47 / 1M tokens output) vs Google Gemini Flash Latest ($0.282 / 1M tokens input / $1.69 / 1M tokens output).
  • Context: Google Gemini Flash Latest has the larger context window (1M tokens).
  • Google Gemini Flash Latest adds: Vision.

Cote a cote

MoonshotAI: Kimi K2 Thinking vs Google Gemini Flash Latest — comparaison complete

Comparez le prix, le fournisseur, le contexte, les capacites, la latence et la base de source.

ModelProviderInputOutputContextCapabilitiesBest forLatencyStatusSource
MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinkingOpenRouter$0.114 / 1M tokens$0.47 / 1M tokens262.1k
StreamingTool callingJSON modeLong context
General chat via OpenRouter, API workloads1000-3000msCatalogPlatform curated
Google Gemini Flash Latest~google/gemini-flash-latestOpenRouter$0.282 / 1M tokens$1.69 / 1M tokens1M
StreamingTool callingJSON modeVision
image understanding, multimodal chat1000-3000msCatalogPlatform curated

FAQ

MoonshotAI: Kimi K2 Thinking vs Google Gemini Flash Latest FAQ

Is MoonshotAI: Kimi K2 Thinking or Google Gemini Flash Latest cheaper?

MoonshotAI: Kimi K2 Thinking is cheaper ($0.114 / 1M tokens input / $0.47 / 1M tokens output) vs Google Gemini Flash Latest ($0.282 / 1M tokens input / $1.69 / 1M tokens output). Actual cost depends on your input/output token mix — estimate it with the pricing calculator.

Which has a larger context window, MoonshotAI: Kimi K2 Thinking or Google Gemini Flash Latest?

Google Gemini Flash Latest is larger (1M tokens) vs 262.1k tokens.

MoonshotAI: Kimi K2 Thinking vs Google Gemini Flash Latest for Long context: which should I pick?

Both target Long context. Pick MoonshotAI: Kimi K2 Thinking to optimize cost, or Google Gemini Flash Latest for the longer context window. Test both on real prompts before committing production traffic.