Instead of chasing viral news headline, Xiaomi literally made a dedicated hack agent to probe the environments as hard as possible before letting LLMs to train in the environments.
"continued this process until the hack agent could no longer find a successful exploit in any of the environments"
If Chinese labs can do it, why can't the other frontier labs do it? Makes you wonder if they are just being reckless or its simply some viral marketing stunt
This is open model SoTA btw, finding exploits is def possible, it's just if you want to "let" it escape or not