Luo Fuli: The R&D challenges of MiMo-V2.6 have surpassed those of DeepSeek R1 that I have been involved in before.
2026-09-22 11:57:44
According to CoinMeta, on-chain analyst Luo Fuli stated that the development challenges of MiMo-V2.6 have surpassed those of DeepSeek R1 in which she participated. After the release of MiMo-V2.6, she further explained the innovations and engineering challenges of this round of reinforcement learning. Luo Fuli mentioned that MiMo uses both mixrl and mopd simultaneously, while mixrl combines verifiable tasks such as code, general agent, vision, and network security in the same round of reinforcement learning training. mopd, on the other hand, deals with extremely long, difficult-to-verify, or subjectively rewarding tasks; these are trained separately before being integrated back into the main model. She pointed out that games and 3D tasks have longer execution times and it is difficult to automatically determine right from wrong, therefore MiMo trains such tasks separately before integrating their capabilities through mopd.
Source:Internet
This content is for market information only and does not constitute investment advice.
Follow HQYC official accounts to stay updated

Hot Articles
Refresh

What is Bybit Exchange? Is Bybit Safe with EU Dual Licenses?
09-21 18:42

Bitcoin 5-Year Outlook: $75.5K Miner Cost, Crash or Floor?
09-20 19:23

Legit Bitcoin Trading Apps 2026: Top 4 Safe & Regulated Picks
09-18 18:52

Zcash Jumps 23% After Fed Hike, Beats Bitcoin: How Far Can It Go?
09-17 18:03

Which Crypto Wallet Is Best? 2026 Ranking & Review
09-16 18:34



