Mistral models
22 models on Infery.ai.
Codestral 2508
codestral-2508
Mistral's cutting-edge language model for coding released end of July 2025. 256,000 token context window.
256K ctx
Ministral 3 14B 2512
ministral-14b-2512
The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral…
262K ctx
Ministral 3 14B Instruct 2512
ministral-3-14b-instruct-2512
262K ctx
Ministral 3 3B 2512
ministral-3b-2512
The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.
131K ctx
Ministral 3 8B 2512
ministral-8b-2512
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
262K ctx
Mistral 7B Instruct v0.1
mistral-7b-instruct-v0.1
3K ctx
Mistral (7B) Instruct v0.3
mistral-7b-instruct-v0.3
33K ctx
Mistral Large
mistral-large
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). 128,000 token context window.
128K ctx
Mistral Large 2407
mistral-large-2407
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). 131,072 token context window.
131K ctx
Mistral Large 3 2512
mistral-large-2512
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters…
262K ctx
Mistral Medium 3
mistral-medium-3
Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly…
131K ctx
Mistral Medium 3.1
mistral-medium-3.1
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to…
131K ctx
Mistral Medium 3.5
mistral-medium-3-5
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. 262,144 token context window.
262K ctx
Mistral Nemo
mistral-nemo
A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. 131,072 token context window.
131K ctx
Mistral Small 3
mistral-small-24b-instruct-2501
Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks.
33K ctx
Mistral Small 3.1 24B
mistral-small-3.1-24b-instruct
Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal…
128K ctx
Mistral Small 3.2 24B
mistral-small-3.2-24b-instruct
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition…
128K ctx
Mistral Small 4
mistral-small-2603
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a…
262K ctx
Mixtral 8x22B Instruct
mixtral-8x22b-instruct
Mistral's official instruct fine-tuned version of Mixtral 8x22B. 65,536 token context window.
66K ctx
Mixtral-8x7B Instruct v0.1
mixtral-8x7b-instruct-v0.1
33K ctx
Saba
mistral-saba
Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and…
33K ctx
Voxtral Small 24B 2507
voxtral-small-24b-2507
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class…
32K ctx