Google: Gemini 3.5 Flash-Lite API
Our most cost-efficient GA model, optimized for high-volume agentic tasks, translation, and simple data processing.
- Context window: 1,048,576 tokens
- Max output: 65,536 tokens
- Input: text, image, audio, video, file
- Output: text
- Reasoning: Supported
- Tool calling: Supported
- File input: Supported
- Released: 2026-03-02
- Knowledge cutoff: 2024-12-31
Frequently Asked Questions
What is the context window of Gemini 3.5 Flash-Lite?
Gemini 3.5 Flash-Lite supports a context window of up to 1,048,576 tokens.
Does Gemini 3.5 Flash-Lite support function calling?
Yes. Gemini 3.5 Flash-Lite supports tool / function calling.
Does Gemini 3.5 Flash-Lite support reasoning?
Yes. Gemini 3.5 Flash-Lite is a reasoning-capable model.
What is the knowledge cutoff of Gemini 3.5 Flash-Lite?
The knowledge cutoff of Gemini 3.5 Flash-Lite is 2024-12-31.