Diterbitkan pada 2026-08-16 · 赵予安

Jawaban langsung

Google Gemini 3.5 Flash Lite is listed at 0.043 / 0.362 per million tokens, with a 1,048,576 token context window. First-party cache scores are unpublished. Panduan ini ditulis untuk tim produk dan platform yang membandingkan kualitas model, biaya, kebijakan routing, dan risiko peluncuran.

What shipped

Google: Gemini 3.5 Flash Lite appears in the NextModel live catalog as google/gemini-3.5-flash-lite, with input at $0.043 and output at $0.362 per million tokens. The public slug is google--gemini-3.5-flash-lite. The provider is Google.

Pricing and context

Input price is $0.043 per 1 million tokens. Output price is $0.362 per 1 million tokens. The context window is 1,048,576 tokens. The latency band is 1000 to 3000 milliseconds. Related catalog entries include Google: Gemini 2.5 Flash at the same input and output prices. Google: Gemini 3.1 Flash Lite is $0.036 in and $0.217 out. Google: Gemini 3 Flash Preview is $0.072 in and $0.434 out.

Who this is for

The catalog lists streaming, tool, json, vision, and long as capabilities. The assigned use cases are vision, long-context, and agent work. This model is listed for long-context and agent applications with vision support.

Still unverified

The release date is unpublished in this catalog snapshot. Independent third-party quality scores are not in these notes. No first-party CacheSafety Bench row exists for this exact model, so Safe Hit, Bad Hit, and trap rates are not reported.