Gemini 3.5 Flash Lite
Google: Gemini 3.5 Flash Lite is a Google model listed in the NextModel catalogue for image understanding, multimodal chat workloads. Its listed price is $0.043 / 1M tokens input and $0.362 / 1M tokens output per 1M tokens, with a 1M token context window.
What is Gemini 3.5 Flash Lite in NextModel?
Google: Gemini 3.5 Flash Lite is a Google model listed in the NextModel catalogue for image understanding, multimodal chat workloads. Its listed price is $0.043 / 1M tokens input and $0.362 / 1M tokens output per 1M tokens, with a 1M token context window.
Best use cases
- image understanding
- multimodal chat
OpenAI-compatible code example
Keep the OpenAI SDK style, set base_url to NextModel, and use the catalogue model ID google--gemini-3.5-flash-lite.
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://api.nextmodel.app/v1"
)
resp = client.chat.completions.create(
model="google/gemini-3.5-flash-lite",
messages=[{"role": "user", "content": "Hello from NextModel"}]
)
print(resp.choices[0].message.content)Similar alternatives
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
Compare Google: Gemini 3.5 Flash Lite
FAQ