Back to Models
google logo
google/gemma-4-26b-a4b-it%3Afree
Not Available

Gemma 4 26B A4B (free)

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

4/3/2026
262,144 tokens

Specifications

Modalities

Input
image
text
video
Output
text

Supported Parameters

frequency_penalty
include_reasoning
logprobs
max_tokens
presence_penalty
reasoning
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_logprobs
top_p

Max Output Tokens

32,768

Reasoning Configuration

Default
Thinking off
Turning thinking off
Possible (reasoning.enabled: false)

How to control thinking, and what it costs

Frequently asked questions

Is Gemma 4 26B A4B (free) available on FastMetal?
Not at the moment. GLM 5.3 Flash, from the same lab, is available on the FastMetal API today.
What is the context window of Gemma 4 26B A4B (free)?
262,144 tokens, shared between the prompt and the response.