Processing...Please wait while we secure this action
Google model

Gemma 3n 4B

Google: Gemma 3n 4B is a Google model listed in the NextModel catalog for General chat via Google, API workloads workloads. Its listed price is $0.012 / 1M tokens input and $0.023 / 1M tokens output per 1M tokens, with a 32.8k token context window.

GooglePlatform curatedCatalog
StreamingJSON mode
Input price$0.012 / 1M tokens
Output price$0.023 / 1M tokens
Context length32.8k tokens
Max output8.2k tokens

What is Gemma 3n 4B in NextModel?

Google: Gemma 3n 4B is a Google model listed in the NextModel catalog for General chat via Google, API workloads workloads. Its listed price is $0.012 / 1M tokens input and $0.023 / 1M tokens output per 1M tokens, with a 32.8k token context window.

Best use cases

  • General chat via Google
  • API workloads

OpenAI-compatible code example

Keep the OpenAI SDK style, set base_url to NextModel, and use the catalog model ID google--gemma-3n-e4b-it.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="google/gemma-3n-e4b-it",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

Similar alternatives

GoogleCatalog

Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini). Gemma models are well-suited for a variety of...

$0.123 / 1M tokensInput$0.123 / 1M tokensOutput8.2kContext
Best forGeneral chat via Google, API workloads
RoutingConfigured
StreamingJSON mode
Platform curatedNextModel gateway catalog (Go origin)
View details
GoogleCatalog

Available through the NextModel gateway via Google.

$0.02 / 1M tokensInput$0.075 / 1M tokensOutputContext
Best forGeneral chat via Google, API workloads
RoutingDegraded
Streaming
Platform curatedNextModel gateway catalog (Go origin)
View details
OpenRouterCatalog

Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and...

$0.029 / 1M tokensInput$0.094 / 1M tokensOutput65.5kContext
Best forGeneral chat via OpenRouter, API workloads
RoutingDegraded
StreamingJSON mode
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Google: Gemma 3n 4B API questions