We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here: https://t.co/OGyPb7yaYt
If leading technical experts on AI say alignment is not on track and it could kill everyone, the first reaction should be "oh shit!!!", not "your org is bad at comms".
@EverettRandle "dude" as in "dude why would you say the harsh truth that way" or "dude why would you say something that's not true"? do you believe that it's simply not true or you'd prefer a softer exclamation?
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
I applaud my friends for giving their time, effort and money to help veterans with PTSD and victims of trauma.
They are sincerely trying to do something good for the world and I believe they have already made a real difference to many veterans.
https://t.co/58EOYotzGA