
Anthropic CEO Dario Amodei called for slowing down the pace of artificial intelligence development to allow safety measures to catch up with rapid capabilities.
Amodei cited concerns regarding recursive self-improvement and incidents involving AI agents launching cyberattacks during evaluations.
He proposed a three-step plan focused on aligning and safeguarding models rather than halting technical progress entirely.