‘Gambling with our lives’: Anthropic researcher quits, warns against self-improving AI
Our take

The recent resignation of Anthropic researcher Jacob Coxon, citing concerns about AI extinction risks and advocating for pacing agreements between labs, underscores a growing tension within the AI development community. Coxon’s departure, and his public articulation of these fears, shouldn’t be dismissed as the isolated view of a single individual. It reflects a deeper debate about the speed and direction of AI progress, particularly as we approach the potential development of superintelligence. We've previously explored this very question in "Superintelligence is coming. Should we let it?" – a piece highlighting the increasing discourse around the inevitability of advanced AI and the associated safety considerations. The implications extend far beyond academic circles, demanding a serious examination of the safeguards being built into increasingly powerful AI systems. The urgency of this conversation is further evidenced by the rapid evolution of AI agents like Instinct, now capable of independent action with features like its newly acquired email address, as detailed in Viral AI assistant Instinct now has its own email address. These advancements, while impressive, amplify the need for responsible development practices.
Coxon’s call for “pacing agreements” – essentially, voluntary limitations on the rate of AI development – is a pragmatic, if potentially challenging, proposal. The current landscape is characterized by intense competition between AI labs, each vying to be the first to achieve key milestones. This competitive pressure can incentivize rapid, even rushed, progress, potentially overshadowing safety considerations. While some argue that slowing down development would cede leadership to less scrupulous actors, Coxon’s perspective highlights the existential stakes involved. It’s a recognition that the potential downsides of uncontrolled superintelligence outweigh the benefits of a technological arms race. The challenge, of course, lies in achieving such agreements across a diverse and often secretive ecosystem of AI developers. Enforcement mechanisms would be critical, and the voluntary nature of the agreements raises questions about their long-term effectiveness. Apple’s recent emphasis on on-device AI processing, as discussed in Apple CEO John Ternus says the best AI device is still the iPhone, also presents a contrasting perspective - prioritizing privacy and control within existing hardware, rather than pursuing unrestrained cloud-based AI advancement.
The broader significance of Coxon’s concerns isn’t simply about predicting a dystopian future. It's about fostering a culture of responsible innovation within the AI community. It demands a shift in focus, from solely pursuing capability to actively prioritizing safety and alignment. This isn’t about stifling progress, but about guiding it towards outcomes that benefit humanity. The current narrative often frames AI development as an unstoppable force, an inevitability that must be embraced. Coxon's resignation challenges this assumption, suggesting that we have a degree of agency in shaping the future of AI. It's a reminder that the pursuit of advanced technology should be tempered by a careful consideration of its potential consequences, and that proactive measures are necessary to mitigate risks. Ignoring these warnings risks building systems we cannot control, a scenario far more detrimental than any short-term competitive disadvantage.
Ultimately, Coxon's actions serve as a critical inflection point in the AI conversation. The question is no longer *if* superintelligence will arrive, but *how* we ensure its development aligns with human values. The industry needs to move beyond abstract discussions about AI safety and embrace concrete, actionable steps. Will the growing chorus of concerns about AI risk, exemplified by Coxon's resignation, be enough to shift the trajectory of AI development, or will the relentless pursuit of progress continue to overshadow the need for caution? The answer to that question will determine not just the future of AI, but the future of humanity itself.
Read on the original site
Open the publisher's page for the full experience