G

Google: gemini-2.5-flash-lite

gemini-2.5-flash-lite
100万 上下文6.6万 输出推理工具并行工具视觉音频输入视频文件缓存结构化网络搜索URL上下文系统消息
发布日期 Jul 22, 2025知识截止 Jan 2025更新时间 Sep 22, 2026

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.

模式 chat分词器 Gemini弃用日期 Oct 2026过期日期 Oct 2026

价格

输入价格
$0.09/ 100 万 token7% 优惠
输出价格
$0.37/ 100 万 token7% 优惠
兼容端点 openai供应商 Google

在线时长

性能

正在加载性能数据...

使用量与排名

正在加载使用量……

支持的参数

所有提供方:为该模型提供服务的每个上游都支持。部分提供方:取决于处理请求的上游。默认值:未设置时发送的值。

参数提供方默认值
frequency_penalty部分提供方默认不发送
include_reasoning所有提供方-
max_tokens所有提供方-
reasoning所有提供方-
response_format所有提供方-
seed所有提供方-
stop所有提供方-
structured_outputs所有提供方-
temperature所有提供方默认不发送
tool_choice所有提供方-
tools所有提供方-
top_p所有提供方默认不发送

常见问题

相似模型