G
Google: gemini-2.5-flash-lite
gemini-2.5-flash-lite
100万 上下文6.6万 输出推理工具并行工具视觉音频输入视频文件缓存结构化网络搜索URL上下文系统消息
发布日期 Jul 22, 2025知识截止 Jan 2025更新时间 Sep 22, 2026
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.
模式 chat分词器 Gemini弃用日期 Oct 2026过期日期 Oct 2026
价格
输入价格
$0.09/ 100 万 token7% 优惠
输出价格
$0.37/ 100 万 token7% 优惠
兼容端点 openai供应商 Google
在线时长
性能
正在加载性能数据...
使用量与排名
正在加载使用量……
支持的参数
所有提供方:为该模型提供服务的每个上游都支持。部分提供方:取决于处理请求的上游。默认值:未设置时发送的值。
| 参数 | 提供方 | 默认值 |
|---|---|---|
| frequency_penalty | 部分提供方 | 默认不发送 |
| include_reasoning | 所有提供方 | - |
| max_tokens | 所有提供方 | - |
| reasoning | 所有提供方 | - |
| response_format | 所有提供方 | - |
| seed | 所有提供方 | - |
| stop | 所有提供方 | - |
| structured_outputs | 所有提供方 | - |
| temperature | 所有提供方 | 默认不发送 |
| tool_choice | 所有提供方 | - |
| tools | 所有提供方 | - |
| top_p | 所有提供方 | 默认不发送 |