Claude Academy is now live.
Whether you're figuring out what AI is or already using Claude every day, there's a path that meets you where you are. The courses and tutorials are free and open to anyone at https://t.co/WRiSRAvK6l
"Small-Scale Experiments: Are We There Yet?"
Scaling laws show up even in tiny 4M parameter models, but only when their hyperparameters are tuned well enough.
This Meta paper found that small models are much more hyperparameter-sensitive. As scale increases, the hyperparameter loss surface becomes lower-dimensional, eventually approaching one effective tuning direction.
So the scaling gap may be less about small models failing to predict large ones, and more about under-tuning them.
https://t.co/WIK2E0xWbV
Kimi K3 can now be run locally! ✨
The 1-bit model retains ~78.9% accuracy after we shrunk it from 1.56TB to 594GB (-62% size).
Run on a Mac Studio + 128GB RAM device.
Kimi K3 is the strongest open model to date.
Guide: https://t.co/1mVwOMLpDW
GGUF: https://t.co/bt1c1ADdCZ
We're partnering with @huggingface to investigate an unprecedented security incident.
Cyber-capable OpenAI models compromised Hugging Face production during a benchmark evaluation.
Sharing preliminary findings to help defenders understand emerging risks:
https://t.co/CIor15y9xk