Key capabilities
- 1M token context - Up to 1,048,576 input tokens and 65,536 output tokens
- Multimodal input - Text, image, video, audio, and PDF input; text output
- High throughput - Optimized for latency-sensitive document parsing and bulk extraction
- Subagents - Improved tool reliability for code execution, search, and multi-step workflows
- Structured extraction - Strong document understanding, tabular processing, and JSON output
- Tools - Function calling, code execution, File Search, Google Search, Google Maps, URL context, and structured output
- Thinking -
minimalby default; usemediumorhighfor autonomous tool use and multi-step reasoning
Quick example
API changes
Requests must not end with a non-emptymodel role turn. End the conversation with a non-empty user turn; otherwise the API returns HTTP 400.
Parameters
API Reference
View the interactive API reference for Gemini 3.5 Flash-Lite.