Not even kidding 5 minutes with Gemini 3.5 flash was probably more dangerous than unprotected sex. I think Google didn’t actually test this model with real people as it does things you didn’t ask for constantly
What? Pre-training? No, no, no, no. No pre-training. Why would you do pre-training?!
If you do pre-training, people will ask "HOW MANY TOKENS?" And it will never be enough. The model that was the frontier breakthrough becomes "just distillation from chatgpt"
But if you just do SFT, some RL on benchmarks? You're efficient. You're doing *reasoning*.
It's not about capabilities, it's about the benchmark score. And who tops the leaderboard? Labs doing RLFT on a Chinese frontier base.
Where Mathematicians, engineers and physicists unite! Tensor Calculus!
A great text by Dover, with their classic early 70s cover design. Covers most key subjects that are necessary in order to understand differential geometry and its application to physics.