Back to Models
z-ai logo
z-ai/glm-4-5v
Not Available

GLM 4.5V

GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding, image Q&A, OCR, and document parsing, with strong gains in front-end web coding, grounding, and spatial reasoning. It offers a hybrid inference mode: a "thinking mode" for deep reasoning and a "non-thinking mode" for fast responses. Reasoning behavior can be toggled via the `reasoning` `enabled` boolean. [Learn more in our docs](https://openrouter.ai/docs/use-cases/reasoning-tokens#enable-reasoning-with-default-config)

8/11/2025
65,536 tokens

Specifications

Modalities

Input
text
image
Output
text

Supported Parameters

frequency_penalty
include_reasoning
max_tokens
presence_penalty
reasoning
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_p

Max Output Tokens

16,384

Frequently asked questions

Is GLM 4.5V available on FastMetal?
Not at the moment. GLM 4.7 Flash, from the same lab, is available on the FastMetal API today.
What is the context window of GLM 4.5V?
65,536 tokens, shared between the prompt and the response.