Gemma 3n 4B 在 NextModel 中是什麼?
香港團隊可用的 NextModel 目錄中的 Google 模型,常用於 General chat via Google、API workloads 工作負載。當前展示價格為輸入 $0.012 / 1M tokens、輸出 $0.023 / 1M tokens 每 100 萬 token,上下文視窗為 32.8k token。
適用場景
- General chat via Google
- API workloads
OpenAI 相容呼叫範例
保持 OpenAI SDK 呼叫方式不變,把 base_url 改為 NextModel,並使用模型目錄 ID google--gemma-3n-e4b-it。
Python
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://api.nextmodel.app/v1"
)
resp = client.chat.completions.create(
model="google/gemma-3n-e4b-it",
messages=[{"role": "user", "content": "Hello from NextModel"}]
)
print(resp.choices[0].message.content)相似替代項
Google目錄
Gemma 2 27B by Google is an open model built from the same research and technology used to create the [Gemini models](/models?q=gemini). Gemma models are well-suited for a variety of...
適用場景General chat via Google, API workloads
路由已設定
串流輸出JSON 模式
平台整理NextModel gateway catalog (Go origin)
Google目錄
Available through the NextModel gateway via Google.
適用場景General chat via Google, API workloads
路由降級
串流輸出
平台整理NextModel gateway catalog (Go origin)
OpenRouter目錄
Olmo 3 32B Think is a large-scale, 32-billion-parameter model purpose-built for deep reasoning, complex logic chains and advanced instruction-following scenarios. Its capacity enables strong performance on demanding evaluation tasks and...
適用場景General chat via OpenRouter, API workloads
路由降級
串流輸出JSON 模式
平台整理NextModel gateway catalog (Go origin)
常見問題