GoogleBase plan
Gemini 3.5 Flash-Lite
Google's fastest Gemini 3.5 model, optimized for high-volume agentic work, document processing, translation, and classification. Uses 3 base requests per send before length multipliers.
You always get the exact model you pick — we never silently route you to another.
Specifications
| Provider | |
|---|---|
| Released | 2026-07-21 |
| Intelligence | Medium |
| Speed | Fast |
| Context window | 1,048,576 tokens |
| Max output | 65,536 tokens |
| Knowledge cutoff | March 2026 |
| Input price | $0.30 / 1M tokens |
| Output price | $2.50 / 1M tokens |
| Request cost | 3 base requests |
| Plan tier | Base |
| Input | Text, Image, Video, Audio, PDF |
| Output | Text |
| Features | 3 base requests per send before length multipliers, Gemini API search grounding, function calling, structured outputs, and code execution, Context caching supported |
| Model ID | gemini-3.5-flash-lite |