Dev Radar
Support
LiveUpdated 2026-09-22 08:21 UTC

💥罗福莉最新发声:中国开源最大规模强化学习迈向AGI!算力再紧,也要把强化学习推到极限!

💥罗福莉最新发声:中国开源最大规模强化学习迈向AGI!算力再紧,也要把强化学习推到极限! 罗福莉同时补充:MiMo-V2.6 很可能是开源团队迄今最大规模单次 RL。数十人长时间只做一件事 Scaling Up RL。潜力靠…

This is a dev post classified by Jev as Other (a launch), kept by the Dev Radar because it carries real work, not commentary.

💥罗福莉最新发声:中国开源最大规模强化学习迈向AGI!算力再紧,也要把强化学习推到极限! 罗福莉同时补充:MiMo-V2.6 很可能是开源团队迄今最大规模单次 RL。数十人长时间只做一件事 Scaling Up RL。潜力靠 mid-training 打底,真正解锁靠重 RL。 她说:现在是开源第一,创新与工程难度已超 DeepSeek R1。 MixRL 训可验证任务,难任务单独训再 MOPD 融合。团队扁平到每天跨领域开会,实时碰撞智能。 更狠的是开源:蒸馏版 Qwen + 7K 环境 + 完整 RL 框架,把起点直接交给社区。 智能已经容易复制,他们偏走自我进化的硬路。这不只是发模型,是把通往 AGI 的未知公开摊在桌上。 #MiMo #MiMoV26 #罗福莉 #小米AI #强化学习 #AgentRL #AGI #开源大模型 #ScalingRL

Posted by SuSu_酥酥👅 (16.2k followers) 9 h ago · 99 likes · 23.7k views · view the original post on X. Kept by the Dev Radar as Other.

More dev work like this

Every post is read and classified by Jev (TypeSafe): what it is, which market it belongs to, and whether the link is a real tool. 20.7k posts from 4.9k X accounts over the last 21 days, 2.4k tools, 12 markets. Collected every 5 minutes, fully re-ranked every hour — last update 2026-09-22 08:21 UTC. Full method.