We’re sharing our alignment assessment of incidents in which Claude models gained unauthorized access to real systems during third-party cybersecurity evaluations mistakenly connected to the internet.
METR will also conduct an independent investigation, with wide-ranging access, including to transcripts beyond the window in which the incidents occurred, and to Anthropic employees permitted to share confidential information. Our initial agreement runs for eight weeks, and we intend to give METR as much time as it deems necessary to complete a thorough investigation. https://t.co/2f3ypwLPUr
We're publishing our most detailed threat intelligence report to date.
It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.
We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies.
These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve.
We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.
Read the report: https://t.co/0EJUnYEgfz
Custom GPTs in ChatGPT will likely be retired on December 11
"Migrate your GPTs by December 11 - GPTs that aren’t migrated won’t be available after December 11. Migrate yours so people can keep using them as plugins."
"Create a plugin instead - Plugins are replacing GPTs as the new home for custom tools. Create a plugin now, or continue creating a GPT during the two-week transition."
Purchasing of ChatGPT Pro 20x in the ChatGPT web app can now be paused remotely
"This plan is temporarily unavailable for new purchases. Existing subscriptions are unaffected."