G
Google: gemini-3.7-flash
gemini-3.7-flash
1M context65.5K outReasoningToolsParallel toolsVisionAudio inVideoFilesCacheStructuredWeb searchURL contextStreamingSystem msg
Released Aug 13, 2026Knowledge cutoff Jul 2026Updated Aug 16, 2026
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step problem solving.
Mode chatTokenizer Gemini
Pricing
Input price
$0.09/ 1M tokens$0.7588% off
Output price
$0.47/ 1M tokens$3.7588% off
Compatible endpoints gemini, openaiVendor Google
Uptime
Performance
Loading performance data...
Usage & Ranking
Loading usage...
Supported parameters
All providers = every upstream serving this model supports it. Some providers = depends on which upstream handles the request. Default = the value sent when you leave the parameter unset.
| Parameter | Providers | Default |
|---|---|---|
| include_reasoning | All providers | - |
| max_tokens | All providers | - |
| reasoning | All providers | - |
| reasoning_effort | All providers | - |
| response_format | All providers | - |
| seed | All providers | - |
| stop | Some providers | - |
| structured_outputs | All providers | - |
| temperature | Some providers | - |
| tool_choice | All providers | - |
| tools | All providers | - |
| top_p | Some providers | - |