Xiaomi open-sources MiMo-V2.6, using RL to let models teach themselves
小米发布并开源 MiMo-V2.6 系列,扩展强化学习规模迈向自我提升
Xiaomi released and open-sourced the MiMo-V2.6 series, focusing on scaling reinforcement learning for self-improvement. The post is blocked by WeChat and does not disclose specific parameters, performance, or repo links.