G

Google: gemini-3.5-flash-lite-free:free

gemini-3.5-flash-lite-free:free
1M context65.5K outReasoningToolsParallel toolsVisionAudio inVideoFilesCacheStructuredWeb searchURL contextStreamingSystem msg
Released Jul 21, 2026Knowledge cutoff Dec 2024Updated Sep 4, 2026

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Mode chatTokenizer GeminiDeprecation Jul 2027

All providers for this model are busy right now

Every upstream provider has hit its rate limit. The model comes back automatically once limits lift, usually within hours. Try again in a little while or switch to another model.

Request this model on Discord

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Compatible endpoints -Vendor Google

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltySome providersNot sent by default
include_reasoningAll providers-
max_tokensAll providers-
presence_penaltySome providersNot sent by default
reasoningAll providers-
reasoning_effortAll providers-
repetition_penaltySome providersNot sent by default
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureSome providersNot sent by default
tool_choiceAll providers-
toolsAll providers-
top_kSome providersNot sent by default
top_pSome providersNot sent by default

Frequently asked questions

Similar models