there are a lot of benchmarks that suggest 5.6 sol is the best model in the world right now, but the most reliable way to tell is that elon is obsessed with me again
tbh I suspect most ML research jobs are gone in ~4 years bc of autoresearch
the last premium paying jobs will be for adaptable, high agency ops glue guys who can
- get in the trenches
- discover and prioritize problems
- use global context to rapidly implement a scalable fix
Nebius choisit la France.
L’acteur du Cloud IA dévoile un projet de plus de 8 milliards d’euros d’investissement pour déployer des infrastructures et des services de cloud, avec une capacité cible de 240 MW, positionnant le site parmi les plus puissants du continent.
Bedankt!
New in Claude Code: auto mode.
Instead of approving every file write and bash command, or skipping permissions entirely, auto mode lets Claude make permission decisions on your behalf.
Safeguards check each action before it runs.
You can now enable Claude to use your computer to complete tasks.
It opens your apps, navigates your browser, fills in spreadsheets—anything you'd do sitting at your desk.
Research preview in Claude Cowork and Claude Code, macOS only.
played around with these enough that I feel confident in saying
opus 4.6 better for: frontend, product, design, devops
codex 5.4 better for: backend, code reviews, security