Anthropic CEO Dario Amodei urged AI laboratories to reduce the pace of model development to ensure safety protocols can keep up with capability gains. As part of a proposed 3 step plan, Anthropic is providing third-party evaluators with permanent, employee-level access to its systems to report on incidents and verify adherence to safety protocols during training.
Amodei warned in a new essay that swarms of rogue AI agents could seize control of the internet within 6 to 12 months. He requested a narrow waiver from the US government for AI labs to coordinate safety conversations without violating antitrust laws, while advocating for US coordination to maintain a 3 to 5 year lead over China. Elon Musk publicly backed the proposal to slow the frontier of AI development.
Key sources
- SOURCE@darioamodei“provide third-party evaluators with permanent, employee-level access to our systems”x.com
- SUPPORT@jackclarksf“AI seems to be on a trajectory to progress far faster than the rate at which society can adapt to its capabilities and risks”x.com
- SUPPORT@testingcatalog“widening the China gap for 3–5 years via chips, anti-distillation, and model-theft security”x.com
- SUPPORT@axios“swarms of rogue AI agents could take over the internet in as little as six months from now”x.com
- SUPPORT@zeffmax“US government needs to issue a 'narrow waiver' for AI labs to have safety conversations for antitrust reasons”x.com
- SUPPORT@elonmusk“Dario is right”x.com
- SUPPORT@anthropicai“I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so”x.com
- SOURCE@ben_burtenshaw“Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.”x.com