Back to Models
google logo
google/gemini-3-6-flash
Not Available

Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

7/21/2026
1,048,576 tokens

Specifications

Modalities

Input
text
image
video
file
audio
Output
text

Supported Parameters

include_reasoning
max_tokens
reasoning
reasoning_effort
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_p

Max Output Tokens

65,536

Reasoning Configuration

Default
Thinking on (effort: medium)
Selectable effort levels
high
medium
low
minimal
Turning thinking off
Not possible (always thinks)

How to control thinking, and what it costs

Frequently asked questions

Is Gemini 3.6 Flash available on FastMetal?
Not at the moment. GLM 5.3 Flash, from the same lab, is available on the FastMetal API today.
What is the context window of Gemini 3.6 Flash?
1,048,576 tokens, shared between the prompt and the response.