跳到正文
Mistral AI

Mistral AI 发布 Mistral Small 3,24B 模型采用 Apache 2.0 许可

Mistral Small 3

Mistral AI 发布面向低延迟任务的 24B 模型 Mistral Small 3,以 Apache 2.0 许可开放预训练和指令微调检查点。官方称其性能与 Llama 3.3 70B instruct 相当,同硬件下速度超过 3x,MMLU 准确率超过 81%;模型未使用 RL 或合成数据训练。

推荐理由:Mistral AI 将 Small 3 的性能对标 Llama 3.3 70B,并报告同硬件下超过 3x 的速度优势。

读原文 ↗导出 Markdown