Mistral Small 4 unifies instruct, reasoning, and coding in one open model
Original titlemistralai/Mistral-Small-4-119B-2603
Mistral Small 4 is a 119B-parameter MoE model with 6.5B active per token and a 256k context window, combining instruct, reasoning, and Devstral-style coding in one model.
It accepts text and image input, lets users set reasoning_effort per request, and is released under Apache 2.0.
The model card reports a 40% latency reduction and 3x throughput versus Mistral Small 3 in its tested setups, and its benchmark chart shows reasoning scores on GPQA Diamond, MMLU Pro, AIME-style text tasks, and MMMU-Pro.
The model card names concrete architecture, context, and licensing details, letting readers compare its reasoning toggle and efficiency claims against other open models.
Source: Mistral AI · new models on Hugging Face · huggingface.coPublished · added here