GoogleBase plan

Gemini 3.5 Flash-Lite

Google's fastest Gemini 3.5 model, optimized for high-volume agentic work, document processing, translation, and classification. Uses 3 base requests per send before length multipliers.

You always get the exact model you pick — we never silently route you to another.

Specifications

ProviderGoogle
Released2026-07-21
IntelligenceMedium
SpeedFast
Context window1,048,576 tokens
Max output65,536 tokens
Knowledge cutoffMarch 2026
Input price$0.30 / 1M tokens
Output price$2.50 / 1M tokens
Request cost3 base requests
Plan tierBase
InputText, Image, Video, Audio, PDF
OutputText
Features3 base requests per send before length multipliers, Gemini API search grounding, function calling, structured outputs, and code execution, Context caching supported
Model IDgemini-3.5-flash-lite

Related models