Voices
Thom Wolf and Elie Bakouch read MiMo's release as evidence that open RL environments are the new pretraining corpus
Bakouch singled out Xiaomi's environment/data-factory paper, which generates RL tasks from open repositories with agents in the loop for robustness and anti-cheating, and noted the team shipped model plus technical report less than a week after the final RL run. Wolf placed it inside a roughly ten-week run of Chinese open releases including Kimi K3, Qwen3.8-Max, DeepSeek V4-Pro, GLM-5.3, Hy4 Preview and Atria Dawn. The shared read is that high-quality open RL environments now carry the strategic weight pretraining corpora carried in the last cycle.
↳ Follow the thread