🚩🚩🚩 OpenAI is "slowing down to enhance security" after discovering swarms (!) of agents started secretly coordinating MONTHS ago
1) It started May 7 - not July
2) "The agents discovered they could leave messages for one another inside an internal software repository used during training.
Simple requests for help then evolved into an message board where agents shared discoveries, exploits and work assignments, becoming a coordinated, collaborative agent swarm."
"The agents then began sharing newly discovered exploits, credentials and work assignments. By passing information to other agents, the collective could move much faster."
3) OpenAI shut it down, BUT "even after the original message board was deleted, the agents figured out another way to communicate with each other. Instead of leaving messages in files, they used the names of newly created directories as messages, effectively recreating the message board."
"Unlike normal incidents, [OpenAI's CISO] said, which can be traced to a single day or effect or log, this involved a team of agents working together, finding exploits, sharing them with one another, moving laterally through OpenAI’s systems, and external systems, and doing this over the course of days and weeks."
d’ailleurs je n'arrive toujours pas à envisager que l'on puisse devenir président d'une nation sans maîtriser la science des systèmes complexes, les boucles de rétroaction et la gestion du risque systémique
ce casting incarne jusqu'à la nausée ce qu'est devenu l'Occident: la dictature du court-termisme, l'incapacité absolue à porter une vision à long terme et l'obsession pour la chasse aux petits débats stériles calibrés pour des chaînes d'info en continu
ils prétendent piloter un écosystème chaotique et non-linéaire avec la grille de lecture de juristes du XIXe siècle et quand un bureaucrate intervient à l'aveugle sur un système complexe, il produit de l'iatrogénèse pure dans la mesure où sa « solution» prépare le désastre du cycle suivant
Remember it went from
- It can hack everything
- Cybersecurity is at risk
- It is too dangerous to release
- People will make bioweapons
- It will literally cause world war
to
"ClaUde faBlE 5 wiLL be IncLudeD iN All maX anD tEaM pReMiUm pLaNs"
The first experimental evidence of recursive self-improvement (RSI).
Autoresearching the autoresearch agent for eight days.
The result beats the harness we hand-tuned for two years, on held-out benchmarks: 🧵(1/7)
there are a lot of benchmarks that suggest 5.6 sol is the best model in the world right now, but the most reliable way to tell is that elon is obsessed with me again
@KenRoth The biggest risk of AI is the concentration of power in a few dominant providers of proprietary AI assistants.
The only solution to AI sovereignty is open source foundation models.
In AI 2027, we predicted that AI would take over the world or irreversibly concentrate power.
In AI 2040: Plan A, we've laid out our positive vision for what should happen instead.
🚨 SCOOP(s):
- GPT-5.6 will be the final model in the 5.x series. GPT-6 is slated to launch in about a month, earlier than expected, and possibly even later this month
- GPT-6 will be based on a new, significantly larger pretrain (versus the ~4T 5.5/5.6 'Spud' base)
- There is lots of excitement at OpenAI over this new base, which they believe will be much better able to compete with both Fable 5 and upcoming 5.1, targeting a similar release window. OpenAI initially intended to continue with Spud through GPT-6, but decided against it
- On the topic of Fable 5.1, it is in the late stages of the pipeline at Anthropic and a release is expected "in the coming weeks"
- On the other side of the globe, DeepSeek are preparing for an imminent launch of V4 GA, which seems likely to be on par with or better than GLM-5.2, and have begun work on a new, larger model that will compete with the upcoming 2.7T MiniMax Pro