G

Google: gemini-2.5-flash-lite

gemini-2.5-flash-lite
100万 コンテキスト6.6万 出力推論ツール並列ツールビジョン音声入力動画ファイルキャッシュ構造化Web検索URLコンテキストシステムメッセージ
リリース日 Jul 22, 2025知識のカットオフ Jan 2025更新日 Sep 22, 2026

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.

モード chatトークナイザー Gemini非推奨 Oct 2026有効期限 Oct 2026

料金

入力料金
$0.09/ 100万トークン7%オフ
出力料金
$0.37/ 100万トークン7%オフ
対応エンドポイント openaiベンダー Google

稼働率

パフォーマンス

パフォーマンスデータを読み込み中...

使用状況とランキング

使用状況を読み込み中...

対応パラメータ

すべてのプロバイダー = このモデルを提供するすべてのアップストリームが対応しています。一部のプロバイダー = リクエストを処理するアップストリームによって異なります。デフォルト = 未設定時に送信される値です。

パラメータプロバイダーデフォルト
frequency_penalty一部のプロバイダーデフォルトでは送信されません
include_reasoningすべてのプロバイダー-
max_tokensすべてのプロバイダー-
reasoningすべてのプロバイダー-
response_formatすべてのプロバイダー-
seedすべてのプロバイダー-
stopすべてのプロバイダー-
structured_outputsすべてのプロバイダー-
temperatureすべてのプロバイダーデフォルトでは送信されません
tool_choiceすべてのプロバイダー-
toolsすべてのプロバイダー-
top_pすべてのプロバイダーデフォルトでは送信されません

よくあるご質問

類似モデル