G

Google: gemini-2.5-flash-lite:free

gemini-2.5-flash-lite:free
1M context65.5K outReasoningToolsParallel toolsVisionAudio inVideoFilesCacheStructuredWeb searchURL contextSystem msg
Released Jul 22, 2025Knowledge cutoff Jan 2025Updated Jun 20, 2026

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater accuracy and nuanced context handling. Additionally, Gemini 2.5 Flash is configurable through the "max tokens for reasoning" parameter, as described in the documentation (https://openrouter.ai/docs/use-cases/reasoning-tokens#max-tokens-for-reasoning).

Mode chatTokenizer GeminiDeprecation Oct 2026

Pricing

Input price
$0.00/ 1M tokens
Output price
$0.00/ 1M tokens
Compatible endpoints openaiVendor Google

Uptime

Performance

Loading performance data...

Usage & Ranking

Loading usage...

Supported parameters

All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.

ParameterProvidersDefault
frequency_penaltySome providersNot sent by default
include_reasoningAll providers-
max_tokensAll providers-
reasoningAll providers-
response_formatAll providers-
seedAll providers-
stopAll providers-
structured_outputsAll providers-
temperatureAll providersNot sent by default
tool_choiceAll providers-
toolsAll providers-
top_pAll providersNot sent by default

Frequently asked questions

Similar models