
Xiaomi has released and open-sourced the MiMo-V2.6 series. The company says the release focuses on scaling reinforcement learning (RL) compute on verifiable and complex tasks through exploration and feedback. Continue reading “Xiaomi releases MiMo-V2.6-Pro and MiMo-V2.6-Flash open-source AI models with scaled RL training”
