G

Google: gemini-2.5-flash-lite

gemini-2.5-flash-lite
100萬 上下文6.6萬 輸出推理工具並行工具視覺音訊輸入影片檔案快取結構化網路搜尋URL上下文系統訊息
發布日期 Jul 22, 2025知識截止 Jan 2025更新時間 Sep 22, 2026

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.

模式 chat分詞器 Gemini棄用日期 Oct 2026過期日期 Oct 2026

價格

輸入價格
$0.09/ 100 萬 token7% 優惠
輸出價格
$0.37/ 100 萬 token7% 優惠
相容端點 openai供應商 Google

在線時長

效能

正在載入效能資料...

使用量與排名

正在載入使用量……

支援的參數

所有提供方:為該模型提供服務的每個上游都支援。部分提供方:取決於處理請求的上游。預設值:未設定時傳送的值。

參數提供方預設值
frequency_penalty部分提供方預設不傳送
include_reasoning所有提供方-
max_tokens所有提供方-
reasoning所有提供方-
response_format所有提供方-
seed所有提供方-
stop所有提供方-
structured_outputs所有提供方-
temperature所有提供方預設不傳送
tool_choice所有提供方-
tools所有提供方-
top_p所有提供方預設不傳送

常見問題

相似模型