Published 2026-08-16 · 赵予安

Direct answer

Google Gemini 3.1 Flash Lite is listed at 0.036 / 0.217 per million tokens, with a 1,048,576 token context window. First-party cache scores are unpublished. This guide is written for Australian product and platform teams comparing model quality, spend, routing policy, and production rollout risk.

What shipped

The NextModel catalog includes Google: Gemini 3.1 Flash Lite as the display name. The public slug is google--gemini-3.1-flash-lite. The provider is Google. The catalog lists streaming, tool, JSON, vision, and long-context capabilities. The latency band is 1000 to 3000 ms.

Model id and pricing

The model id is google/gemini-3.1-flash-lite. Input price is $0.036 per 1M tokens. Output price is $0.217 per 1M tokens. Context window is 1,048,576 tokens. The related Google: Gemini 2.5 Flash entry in the same catalog is $0.043 per 1M input tokens and $0.362 per 1M output tokens.

Who this is for

The catalog labels vision, long-context, and agent as the use cases for this model. Both listed prices are lower than the related Gemini 2.5 Flash entry. The catalog pairs a 1,048,576 token context window with those prices.

Still unverified

The release date is unpublished in this catalog snapshot. Independent third-party quality scores are not in this catalog snapshot. No first-party CacheSafety Bench row exists for this exact model. Safe Hit, Bad Hit, and trap rates are not reported.