Google 모델

Gemini 3.5 Flash Lite

Google: Gemini 3.5 Flash Lite API 가격, 공급자, 컨텍스트 길이, 기능, 활용 사례, 지연 시간, 대안을 비교합니다. image understanding, multimodal chat. $0.043 / 1M tokens / $0.362 / 1M tokens. 1M tokens.

GooglePlatform curatedCatalog
StreamingTool callingJSON modeVisionLong context
입력 가격$0.043 / 1M tokens
출력 가격$0.362 / 1M tokens
컨텍스트 길이1M tokens
최대 출력8.2k tokens

NextModel에서 Gemini 3.5 Flash Lite은 무엇인가요?

Google: Gemini 3.5 Flash Lite API 가격, 공급자, 컨텍스트 길이, 기능, 활용 사례, 지연 시간, 대안을 비교합니다. image understanding, multimodal chat. $0.043 / 1M tokens / $0.362 / 1M tokens. 1M tokens.

적합한 활용 사례

  • image understanding
  • multimodal chat

OpenAI 호환 코드 예시

OpenAI SDK 스타일을 유지하고 base_url 을 NextModel 로 향하게 하며 카탈로그 ID를 사용합니다 google--gemini-3.5-flash-lite.

Python
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://api.nextmodel.app/v1"
)

resp = client.chat.completions.create(
    model="google/gemini-3.5-flash-lite",
    messages=[{"role": "user", "content": "Hello from NextModel"}]
)

print(resp.choices[0].message.content)

유사 대안

GoogleCatalog

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

$0.043 / 1M tokensInput$0.362 / 1M tokensOutput1MContext
Best forimage understanding, multimodal chat
RoutingConfigured
StreamingTool callingJSON modeVision
Platform curatedNextModel gateway catalog (Go origin)
View details
GoogleCatalog

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

$0.072 / 1M tokensInput$0.434 / 1M tokensOutput1MContext
Best forimage understanding, multimodal chat
RoutingConfigured
StreamingTool callingJSON modeVision
Platform curatedNextModel gateway catalog (Go origin)
View details
GoogleCatalog

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

$0.036 / 1M tokensInput$0.217 / 1M tokensOutput1MContext
Best forimage understanding, multimodal chat
RoutingConfigured
StreamingTool callingJSON modeVision
Platform curatedNextModel gateway catalog (Go origin)
View details

FAQ

Google: Gemini 3.5 Flash Lite FAQ