G

Google: gemini-3.5-flash-lite

gemini-3.5-flash-lite
1M context65.5K outReasoningToolsParallel toolsVisionAudio inVideoFilesCacheStructuredWeb searchURL contextStreamingSystem msg
Released Jul 21, 2026Knowledge cutoff Jan 2025Updated Oct 3, 2026

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Mode chatTokenizer GeminiDeprecation Jul 2027

Pricing

Input price
$0.14/ 1M tokens52% off
Output price
$1.19/ 1M tokens52% off
Compatible endpoints gemini, openaiVendor Google

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltySome providersNot sent by default
include_reasoningAll providers-
max_tokensAll providers-
presence_penaltySome providersNot sent by default
reasoningAll providers-
reasoning_effortAll providers-
repetition_penaltySome providersNot sent by default
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureSome providersNot sent by default
tool_choiceAll providers-
toolsAll providers-
top_kSome providersNot sent by default
top_pSome providersNot sent by default

Frequently asked questions

Similar models