A few hours of post training data.
That was enough. Memo picked up a new skill because recovery, generalization, and robustness were already there.
Pretraining did the heavy lifting.
When your base data is diverse enough, post training is not the foundation. It's where all that upstream work pays off. ACT-2 is proof of that.