G
Google: gemini-2.5-flash-lite
gemini-2.5-flash-lite
1M 컨텍스트65.5K 출력추론도구병렬 도구비전오디오 입력비디오파일캐시구조화웹 검색URL 컨텍스트시스템 메시지
출시일 Jul 22, 2025지식 기준일 Jan 2025업데이트됨 Sep 22, 2026
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance across common benchmarks compared to earlier Flash models. By default, "thinking" (i.e. multi-pass reasoning) is disabled to prioritize speed, but developers can enable it via the Reasoning API parameter to selectively trade off cost for intelligence.
모드 chat토크나이저 Gemini사용 중단 Oct 2026만료 Oct 2026
요금
입력 가격
$0.09/ 100만 토큰7% 할인
출력 가격
$0.37/ 100만 토큰7% 할인
호환 엔드포인트 openai공급자 Google
가동 시간
성능
성능 데이터 로딩 중...
사용량 및 순위
사용량 불러오는 중...
지원 파라미터
모든 프로바이더 = 이 모델을 제공하는 모든 업스트림에서 지원됩니다. 일부 프로바이더 = 요청을 처리하는 업스트림에 따라 다릅니다. 기본값 = 설정하지 않았을 때 전송되는 값입니다.
| 파라미터 | 프로바이더 | 기본값 |
|---|---|---|
| frequency_penalty | 일부 프로바이더 | 기본적으로 전송되지 않음 |
| include_reasoning | 모든 프로바이더 | - |
| max_tokens | 모든 프로바이더 | - |
| reasoning | 모든 프로바이더 | - |
| response_format | 모든 프로바이더 | - |
| seed | 모든 프로바이더 | - |
| stop | 모든 프로바이더 | - |
| structured_outputs | 모든 프로바이더 | - |
| temperature | 모든 프로바이더 | 기본적으로 전송되지 않음 |
| tool_choice | 모든 프로바이더 | - |
| tools | 모든 프로바이더 | - |
| top_p | 모든 프로바이더 | 기본적으로 전송되지 않음 |