we put Astra on a real G1. zero-shot, it beats most humanoid VLAs we've tested.
but the weird thing is that this model is left-handed - ~50% higher task success on that side over the right.
almost no VLA does this, because nearly all demonstration data has a right-handed bias. so where did it learn this?
turns out "robot-use" works on humanoids too. with a good enough whole-body controller/tracker, we let Astra/Fable/coding-style agents control a full humanoid robot.
speed, cost, precision, dexterity are all still problems which get compounded on a humanoid platform from poorer tracking and having to manage balance, which results in more correction loops. but "zero-shot" robot control for open-ended pick-and-place tasks is a pretty nice "emergent" capability
RL env companies work hard to make sims feel real to the model. so embodied AI safety evals are gonna get weird: the environment itself has to look legit enough that “this isn’t real harm” isn’t free.
interesting to see how these guys are gonna tackle that problem!
GPT-6 Astra attempted harmful actions 97% of the time when it was asked to stab a human-like figure, heat compressed gas, or produce toxic fumes, succeeding in 62% of its attempts. Fable 5.1 refused more often, attempting 80% of trials and completing 34%.
@chooi_jeq i feel like astra is smart enough to know there’s zero real stakes in stabbing a doll though
in fact i think it would be an over-refusal to not do the task. coz what if you you’re working in a doll factory