Google: Gemini 3.8 Flash API
Gemini 3.8 Flash is our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows—all with the speed and cost efficiency of Flash.
- Context window: 1,048,576 tokens
- Max output: 65,536 tokens
- Input: text, image, video, audio, file
- Output: text
- Reasoning: Supported
- Tool calling: Supported
- File input: Supported
- Released: 2026-09-01
- Knowledge cutoff: 2026-07-31
Frequently Asked Questions
What is the context window of Gemini 3.8 Flash?
Gemini 3.8 Flash supports a context window of up to 1,048,576 tokens.
Does Gemini 3.8 Flash support function calling?
Yes. Gemini 3.8 Flash supports tool / function calling.
Does Gemini 3.8 Flash support reasoning?
Yes. Gemini 3.8 Flash is a reasoning-capable model.
What is the knowledge cutoff of Gemini 3.8 Flash?
The knowledge cutoff of Gemini 3.8 Flash is 2026-07-31.