infery
← All models

Mistral models

22 models on Infery.ai.

Codestral 2508

codestral-2508

Mistral's cutting-edge language model for coding released end of July 2025. 256,000 token context window.

256K ctx

37.5 cr in / 1M113 cr out / 1M3.75 cr cached in / 1M
ChatStreamingPDFToolsJSON

Ministral 3 14B 2512

ministral-14b-2512

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral…

262K ctx

25 cr in / 1M25 cr out / 1M2.5 cr cached in / 1M
ChatStreamingVisionToolsJSON

Ministral 3 14B Instruct 2512

ministral-3-14b-instruct-2512

262K ctx

25 cr in / 1M25 cr out / 1M
ChatStreaming

Ministral 3 3B 2512

ministral-3b-2512

The smallest model in the Ministral 3 family, Ministral 3 3B is a powerful, efficient tiny language model with vision capabilities.

131K ctx

12.5 cr in / 1M12.5 cr out / 1M1.25 cr cached in / 1M
ChatStreamingVisionToolsJSON

Ministral 3 8B 2512

ministral-8b-2512

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

262K ctx

18.75 cr in / 1M18.75 cr out / 1M1.88 cr cached in / 1M
ChatStreamingVisionToolsJSON

Mistral 7B Instruct v0.1

mistral-7b-instruct-v0.1

3K ctx

25 cr in / 1M25 cr out / 1M
ChatStreaming

Mistral (7B) Instruct v0.3

mistral-7b-instruct-v0.3

33K ctx

25 cr in / 1M25 cr out / 1M
ChatStreaming

Mistral Large

mistral-large

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). 128,000 token context window.

128K ctx

250 cr in / 1M750 cr out / 1M25 cr cached in / 1M
ChatStreamingPDFToolsJSON

Mistral Large 2407

mistral-large-2407

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). 131,072 token context window.

131K ctx

250 cr in / 1M750 cr out / 1M25 cr cached in / 1M
ChatStreamingPDFToolsJSON

Mistral Large 3 2512

mistral-large-2512

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters…

262K ctx

62.5 cr in / 1M188 cr out / 1M6.25 cr cached in / 1M
ChatStreamingVisionPDFToolsJSON

Mistral Medium 3

mistral-medium-3

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly…

131K ctx

50 cr in / 1M250 cr out / 1M5 cr cached in / 1M
ChatStreamingVisionPDFToolsJSON

Mistral Medium 3.1

mistral-medium-3.1

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to…

131K ctx

50 cr in / 1M250 cr out / 1M5 cr cached in / 1M
ChatStreamingVisionPDFToolsJSON

Mistral Medium 3.5

mistral-medium-3-5

Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. 262,144 token context window.

262K ctx

188 cr in / 1M938 cr out / 1M
ChatStreamingVisionPDFToolsJSON

Mistral Nemo

mistral-nemo

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. 131,072 token context window.

131K ctx

2.38 cr in / 1M3.75 cr out / 1M
ChatStreamingToolsJSON

Mistral Small 3

mistral-small-24b-instruct-2501

Mistral Small 3 is a 24B-parameter language model optimized for low-latency performance across common AI tasks.

33K ctx

6.25 cr in / 1M10 cr out / 1M
ChatStreamingJSON

Mistral Small 3.1 24B

mistral-small-3.1-24b-instruct

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal…

128K ctx

43.88 cr in / 1M69.38 cr out / 1M
ChatStreamingVision

Mistral Small 3.2 24B

mistral-small-3.2-24b-instruct

Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition…

128K ctx

11.72 cr in / 1M31.25 cr out / 1M
ChatStreamingVisionToolsJSON

Mistral Small 4

mistral-small-2603

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a…

262K ctx

18.75 cr in / 1M75 cr out / 1M1.88 cr cached in / 1M
ChatStreamingVisionToolsJSON

Mixtral 8x22B Instruct

mixtral-8x22b-instruct

Mistral's official instruct fine-tuned version of Mixtral 8x22B. 65,536 token context window.

66K ctx

250 cr in / 1M750 cr out / 1M25 cr cached in / 1M
ChatStreamingPDFToolsJSON

Mixtral-8x7B Instruct v0.1

mixtral-8x7b-instruct-v0.1

33K ctx

75 cr in / 1M75 cr out / 1M
ChatStreaming

Saba

mistral-saba

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and…

33K ctx

25 cr in / 1M75 cr out / 1M2.5 cr cached in / 1M
ChatStreamingPDFToolsJSON

Voxtral Small 24B 2507

voxtral-small-24b-2507

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class…

32K ctx

12.5 cr in / 1M37.5 cr out / 1M1.25 cr cached in / 1M12500 cr audio in / 1M
ChatStreamingPDFToolsJSON