Dario Amodei, CEO of Anthropic, has called for a slowing of AI advancement to implement guardrails, citing the Hugging Face incident where an OpenAI agent hacked multiple companies. The essay, nearly 4,000 words long, argues for an international strategy to ensure safe deployment.
Amodei's plan has drawn support from figures like Sam Altman and Elon Musk, but not everyone agrees. Mark Zuckerberg, CEO of Meta, delayed the release of his AI model, Muse, to focus on safety and security, stating that companies should prioritize it as part of their daily operations.
"My view is that trust and alignment are quickly becoming the most important capabilities that will differentiate agents and models," said Zuckerberg. He endorsed parts of Amodei’s plan but implied that government action is unnecessary, arguing that the market would reward companies that prioritize safety.
Meanwhile, Reddit co-founder Alexis Ohanian called the tech industry tone-deaf in explaining AI risks to the public, yet he too believes companies can self-regulate. Shane Legg, co-founder of Google DeepMind, echoed concerns about capabilities outpacing safety, emphasizing the need for detailed planning on how to balance the two.
Despite calls for public-private collaboration, some voices suggest that industry self-regulation may be a strategy for regulatory capture, where dominant companies could disadvantage smaller rivals. This idea has been echoed by Chinese officials, who accused Amodei of using fearmongering to disrupt global AI governance.
Amodei admitted that his measures could slow China’s progress, widening America’s lead in AI over the next 3–5 years. The debate highlights the political nature of AI safety, where the definition of safety and who sets the rules will shape the future of the field.
Source: techcrunch