19 models
Mistral Large 4
mistral-large-4Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads. It offers a 512K-token context window with up...
Mistral Medium 3.5
mistralai/mistral-medium-3-5Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
Devstral 2 2512
devstral-2512Devstral 2 is a state-of-the-art open-source model by Mistral AI specializing in agentic coding. It is a 123B-parameter dense transformer model supporting a 256K context window.
Mistral Large 3 2512
mistralai/mistral-large-2512Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Medium 3.1
mistralai/mistral-medium-3.1Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost.
Devstral Medium
mistralai/devstral-mediumDevstral Medium is a high-performance code generation and agentic reasoning model developed jointly by Mistral AI and All Hands AI.
Mistral Medium 3
mistralai/mistral-medium-3Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost.
Mistral Small 3.1 24B
mistralai/mistral-small-3.1-24b-instructMistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities.
Mistral Small 3
mistralai/mistral-small-24b-instruct-2501Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks. Released under the Apache 2.0 license, it features both pre-trained and instruction-tuned versions designed for efficient local…
Mistral Large 2407
mistralai/mistral-large-2407This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more.
Mistral Large 2411
mistralai/mistral-large-2411Mistral Large 2 2411 is an update of [Mistral Large 2](/mistralai/mistral-large) released together with [Pixtral Large 2411](/mistralai/pixtral-large-2411) It provides a significant upgrade on the previous [Mistral Large 24.07](/mistralai/m…
Pixtral Large 2411
mistralai/pixtral-large-2411Pixtral Large is a 124B parameter, open-weight, multimodal model built on top of [Mistral Large 2](/mistralai/mistral-large-2411). The model is able to understand documents, charts and natural images.
Mistral 7B Instruct
mistralai/mistral-7b-instructA high-performing, industry-standard 7.3B parameter model, with optimizations for speed and context length. *Mistral 7B Instruct has multiple version variants, and this is intended to be the latest version.*
Mistral 7B Instruct v0.3
mistralai/mistral-7b-instruct-v0.3A high-performing, industry-standard 7.3B parameter model, with optimizations for speed and context length. An improved version of [Mistral 7B Instruct v0.2](/models/mistralai/mistral-7b-instruct-v0.2), with the following changes: - Extende…
Mixtral 8x22B Instruct
mistralai/mixtral-8x22b-instructMistral's official instruct fine-tuned version of [Mixtral 8x22B](/models/mistralai/mixtral-8x22b). It uses 39B active parameters out of 141B, offering unparalleled cost efficiency for its size.
Mistral Large
mistralai/mistral-largeThis is Mistral AI's flagship model, Mistral Large 2 (version `mistral-large-2407`). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more.
Mixtral 8x7B Instruct
mistralai/mixtral-8x7b-instructMixtral 8x7B Instruct is a pretrained generative Sparse Mixture of Experts, by Mistral AI, for chat and instruction use. Incorporates 8 experts (feed-forward networks) for a total of 47 billion parameters.
Mistral 7B Instruct v0.1
mistralai/mistral-7b-instruct-v0.1A 7.3B parameter model that outperforms Llama 2 13B on all benchmarks, with optimizations for speed and context length.
Voxtral Mini 3B
mistral-voxtral-mini-3b-2507A 3B-parameter speech-language model built on the Ministral-3B backbone with an audio encoder for state-of-the-art audio understanding. Supports speech transcription, translation, audio Q&A, and voice-to-function calling across 8 languages.