Mistral AI releases Mistral Small 3.1 with multimodal support and 128k context
Mistral Small 3.1
Mistral AI released Mistral Small 3.1, which improves text performance and multimodal understanding over Mistral Small 3 and extends the context window to 128k tokens. Inference runs at 150 tokens per second, and the model is open-sourced under Apache 2.0.
Why it matters: Mistral gives the multimodal, 128k-context and 150 tokens/s figures for Mistral Small 3.1, letting readers compare it with small models of the same class.