Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing. The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints.
Documentation
Visit the Latest model page for full coverage of features and capabilities.
For more information on knowledge cutoff for the Gemini 3.5 Flash-Lite model, see the model card.
gemini-3.5-flash-lite
| Property | Description |
|---|---|
| Model code | gemini-3.5-flash-lite |
| Supported data types |
Inputs Text, Image, Video, Audio, and PDF Output Text |
| Token limits[*] |
Input token limit 1,048,576 Output token limit 65,536 |
| Capabilities |
Not supported Supported Supported Not supported Supported Supported Supported Not supported Not supported Supported Supported Supported Supported |
| Consumption options |
Supported Supported Supported |
| Versions |
|
| Latest update | July 2026 |