G
Google: gemini-2.5-flash-lite
gemini-2.5-flash-lite
100萬 上下文6.6萬 輸出推理工具並行工具視覺音訊輸入影片檔案快取結構化網路搜尋URL上下文系統訊息
發布日期 Jul 22, 2025知識截止 Jan 2025更新時間 Sep 22, 2026
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.
模式 chat分詞器 Gemini棄用日期 Oct 2026過期日期 Oct 2026
價格
輸入價格
$0.09/ 100 萬 token7% 優惠
輸出價格
$0.37/ 100 萬 token7% 優惠
相容端點 openai供應商 Google
在線時長
效能
正在載入效能資料...
使用量與排名
正在載入使用量……
支援的參數
所有提供方:為該模型提供服務的每個上游都支援。部分提供方:取決於處理請求的上游。預設值:未設定時傳送的值。
| 參數 | 提供方 | 預設值 |
|---|---|---|
| frequency_penalty | 部分提供方 | 預設不傳送 |
| include_reasoning | 所有提供方 | - |
| max_tokens | 所有提供方 | - |
| reasoning | 所有提供方 | - |
| response_format | 所有提供方 | - |
| seed | 所有提供方 | - |
| stop | 所有提供方 | - |
| structured_outputs | 所有提供方 | - |
| temperature | 所有提供方 | 預設不傳送 |
| tool_choice | 所有提供方 | - |
| tools | 所有提供方 | - |
| top_p | 所有提供方 | 預設不傳送 |