Back to Models
google logo
google/gemma-4-31b-it%3Afree
Not Available

Gemma 4 31B (free)

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

4/2/2026
262,144 tokens

Specifications

Modalities

Input
image
text
video
Output
text

Supported Parameters

include_reasoning
max_tokens
reasoning
response_format
seed
temperature
tool_choice
tools
top_p

Max Output Tokens

32,768

Reasoning Configuration

Default
Thinking off
Turning thinking off
Possible (reasoning.enabled: false)

How to control thinking, and what it costs

Frequently asked questions

Is Gemma 4 31B (free) available on FastMetal?
Not at the moment. GLM 5.3 Flash, from the same lab, is available on the FastMetal API today.
What is the context window of Gemma 4 31B (free)?
262,144 tokens, shared between the prompt and the response.