Model comparison

Magnum v4 72B vs Gemma 3n 4B

Compare Magnum v4 72B (OpenRouter) and Google: Gemma 3n 4B (Google) by price, context, capabilities, and latency.

Which should you pick?

  • Price: Google: Gemma 3n 4B is cheaper ($0.012 / 1M tokens input / $0.023 / 1M tokens output) vs Magnum v4 72B ($0.564 / 1M tokens input / $0.94 / 1M tokens output).
  • Context: Google: Gemma 3n 4B has the larger context window (32.8k tokens).

Side by side

Magnum v4 72B vs Google: Gemma 3n 4B — full comparison

Compare price, provider, context, capabilities, latency, and source basis.

ModelProviderInputOutputContextCapabilitiesBest forLatencyStatusSource
Magnum v4 72Banthracite-org/magnum-v4-72bOpenRouter$0.564 / 1M tokens$0.94 / 1M tokens16.4k
StreamingJSON mode
General chat via OpenRouter, API workloads1000-3000msCatalogPlatform curated
Google: Gemma 3n 4Bgoogle/gemma-3n-e4b-itGoogle$0.012 / 1M tokens$0.023 / 1M tokens32.8k
StreamingJSON mode
General chat via Google, API workloads1000-3000msCatalogPlatform curated

FAQ

Magnum v4 72B vs Google: Gemma 3n 4B FAQ

Is Magnum v4 72B or Google: Gemma 3n 4B cheaper?

Google: Gemma 3n 4B is cheaper ($0.012 / 1M tokens input / $0.023 / 1M tokens output) vs Magnum v4 72B ($0.564 / 1M tokens input / $0.94 / 1M tokens output). Actual cost depends on your input/output token mix — estimate it with the pricing calculator.

Which has a larger context window, Magnum v4 72B or Google: Gemma 3n 4B?

Google: Gemma 3n 4B is larger (32.8k tokens) vs 16.4k tokens.

Magnum v4 72B vs Google: Gemma 3n 4B for High quality: which should I pick?

Both target High quality. Pick Google: Gemma 3n 4B to optimize cost, or Google: Gemma 3n 4B for the longer context window. Test both on real prompts before committing production traffic.