प्रकाशित 2026-08-16 · 赵予安

सीधा उत्तर

Google Gemini 3.5 Flash Lite is listed at 0.043 / 0.362 per million tokens, with a 1,048,576 token context window. First-party cache scores are unpublished. यह गाइड उन product और platform टीमों के लिए है जो model quality, cost, routing policy, और launch risk की तुलना कर रही हैं।

What shipped

Google: Gemini 3.5 Flash Lite appears in the NextModel live catalog as google/gemini-3.5-flash-lite, with input at $0.043 and output at $0.362 per million tokens. The public slug is google--gemini-3.5-flash-lite. The provider is Google.

Pricing and context

Input price is $0.043 per 1 million tokens. Output price is $0.362 per 1 million tokens. The context window is 1,048,576 tokens. The latency band is 1000 to 3000 milliseconds. Related catalog entries include Google: Gemini 2.5 Flash at the same input and output prices. Google: Gemini 3.1 Flash Lite is $0.036 in and $0.217 out. Google: Gemini 3 Flash Preview is $0.072 in and $0.434 out.

Who this is for

The catalog lists streaming, tool, json, vision, and long as capabilities. The assigned use cases are vision, long-context, and agent work. This model is listed for long-context and agent applications with vision support.

Still unverified

The release date is unpublished in this catalog snapshot. Independent third-party quality scores are not in these notes. No first-party CacheSafety Bench row exists for this exact model, so Safe Hit, Bad Hit, and trap rates are not reported.