Back to Models
meta-llama logo
meta-llama/llama-4-maverick
Not Available

Llama 4 Maverick

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward pass (400B total). It supports multilingual text and image input, and produces multilingual text and code output across 12 supported languages. Optimized for vision-language tasks, Maverick is instruction-tuned for assistant-like behavior, image reasoning, and general-purpose multimodal interaction. Maverick features early fusion for native multimodality and a 1 million token context window. It was trained on a curated mixture of public, licensed, and Meta-platform data, covering ~22 trillion tokens, with a knowledge cutoff in August 2024. Released on April 5, 2025 under the Llama 4 Community License, Maverick is suited for research and commercial applications requiring advanced multimodal understanding and high model throughput.

4/5/2025
1,048,576 tokens
#89 Vision (Overall)
Specifications

Modalities

Input
text
image
Output
text

Supported Parameters

frequency_penalty
logit_bias
max_tokens
min_p
presence_penalty
repetition_penalty
response_format
seed
stop
structured_outputs
temperature
tool_choice
tools
top_k
top_p

Max Output Tokens

16,384
Leaderboard
Text
๐Ÿ†OverallELO: 1,327
#215
๐Ÿ‡ฏ๐Ÿ‡ตJapaneseELO: 1,254
#161
๐Ÿ‡จ๐Ÿ‡ณChineseELO: 1,330
#211
๐Ÿ‡ฐ๐Ÿ‡ทKoreanELO: 1,263
#176
๐Ÿ‡ฌ๐Ÿ‡งEnglishELO: 1,341
#223
frenchELO: 1,295
#208
germanELO: 1,334
#154
spanishELO: 1,318
#184
russianELO: 1,327
#194
๐Ÿ’ปCodingELO: 1,373
#204
๐ŸงฎMathELO: 1,318
#198
โœ๏ธCreative WritingELO: 1,307
#194
๐Ÿ“Instruction FollowingELO: 1,314
#212
๐ŸŒถ๏ธHard PromptsELO: 1,339
#209
๐Ÿ’ฌMulti-TurnELO: 1,324
#207
Vision
๐Ÿ†OverallELO: 1,146
#89