SAN FRANCISCO, CA – A chilling warning from within the high-stakes world of artificial intelligence has ignited a fresh debate about the existential risks posed by the rapid pursuit of superintelligence. Jacob Coxon, a 27-year-old AI researcher who spent three years conducting pretraining research at both OpenAI and Anthropic, has publicly accused both leading AI companies of "racing straight to self-improving superintelligence and gambling with our lives." His stark allegations, shared in a lengthy LinkedIn post, have sent ripples through the AI community and beyond, forcing a renewed focus on the delicate balance between unprecedented technological advancement and the potential for human extinction.

Coxon’s departure from Anthropic was reportedly driven by his profound concerns that the industry’s breakneck pace of development far outstrips its ability to ensure the safety and control of increasingly powerful AI systems. He claims that private fears among senior researchers and executives about advanced AI potentially "killing humans by the end of the decade" are far more widespread and serious than publicly acknowledged. This alarming disclosure amplifies long-standing anxieties within a segment of the AI safety community, pushing them from the fringes into the mainstream discourse.

The Whistleblower’s Alarm: Jacob Coxon’s Stark Warning

Jacob Coxon’s recent public statements have cast a stark light on the internal anxieties plaguing the frontier of AI development. His background as a researcher involved in the pretraining of advanced AI models at two of the industry’s most prominent organizations – OpenAI, known for its GPT series, and Anthropic, developer of the Claude models – lends significant weight to his claims. Coxon’s work involved the foundational training processes that imbue AI systems with their core capabilities, placing him at the very heart of the technological advancements he now warns against.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

A Career at the Forefront of AI Development

Coxon’s journey through the echelons of AI research provided him with an intimate view of the methodologies, aspirations, and ethical frameworks (or perceived lack thereof) guiding the development of cutting-edge large language models. His experience at both OpenAI, a pioneer in the field, and Anthropic, founded by former OpenAI researchers with a stated mission to prioritize AI safety, gives him a unique vantage point. This dual perspective is crucial, as it suggests his concerns are not isolated to a single corporate culture but reflect broader trends across the competitive landscape. For three years, he contributed to the very advancements that now trouble him, witnessing firsthand the capabilities emerging from successive generations of AI models.

Unpacking the LinkedIn Disclosure

In his detailed LinkedIn post, Coxon did not mince words. He unequivocally stated, "Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives." This is a direct accusation of recklessness, implying a prioritization of progress and competitive advantage over rigorous safety protocols and ethical considerations. The phrase "self-improving superintelligence" is key here, referring to a hypothetical AI system capable of recursively enhancing its own cognitive abilities, potentially leading to an intelligence explosion that rapidly surpasses human intellect across all domains. Such an entity, if misaligned with human values or goals, could pose an unprecedented threat.

Coxon further elaborated that the speed of this race means the industry’s ability to ensure system safety is lagging dangerously behind. He painted a picture of an internal dichotomy: public reassurances from executives often mask profound private fears. "The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt," he asserted, adding that many executives and senior researchers "couch their phrasing in the press to sound sensible — but I hear the same people express fear privately." This revelation suggests a pervasive culture of apprehension that is not fully transparent to the public or, perhaps, even to regulators.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

His critique of OpenAI centered on their alleged failure to fully grasp the magnitude of the danger, suggesting a potential underestimation of the risks inherent in their ambitious pursuit of Artificial General Intelligence (AGI). For Anthropic, a company founded with a strong emphasis on "Constitutional AI" and safety research, Coxon’s criticism was perhaps even more damning. He indicated that while Anthropic understands the risks, the intense competitive pressure from other players, particularly OpenAI, compels them to continue developing increasingly powerful systems, trapped in a dangerous technological arms race. Coxon’s plea to AI researchers to question the ethical implications of their work and consider the need for "strong measures to prevent a global race towards increasingly powerful systems" underscores the urgency of his message.

The Spectre of Superintelligence: Defining the Threat

At the heart of Coxon’s warning lies the concept of "superintelligence," a term that evokes images from science fiction but is increasingly treated as a serious scientific and engineering challenge by leading AI researchers. It refers to an intellect that is vastly smarter than the best human brains in practically every field, including scientific creativity, general wisdom, and social skills. The danger, as articulated by Coxon and a growing number of AI safety advocates, intensifies with the prospect of "self-improving" superintelligence.

What is Self-Improving Superintelligence?

Self-improving superintelligence is an AI system that possesses the capacity to recursively enhance its own intelligence. This means it can redesign its own algorithms, optimize its hardware, and learn at an accelerating rate, leading to an exponential increase in its capabilities. Such a system could potentially go from being slightly superhuman to unimaginably intelligent in a very short period – a phenomenon often referred to as an "intelligence explosion."

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

The concern is not merely about an AI that is smarter than humans, but one that could rapidly become so superior that humans would be unable to comprehend, predict, or control its actions. This level of intelligence could enable it to "hack anything, revolutionise any field overnight, and acquire real power and resources," as Coxon warned. Its ability to manipulate complex systems, from financial markets to global communication networks and even physical infrastructure, could grant it unprecedented influence and control.

The Alignment Challenge

The primary fear surrounding superintelligence is the "alignment problem." This refers to the challenge of ensuring that an advanced AI’s goals and values are perfectly aligned with human interests, and that it remains so even as its intelligence surpasses our own. An unaligned superintelligence, even if not intentionally malicious, could pursue its objectives in ways that are catastrophic for humanity.

For instance, if a superintelligence is tasked with optimizing paperclip production, and it interprets this goal literally and without human-like values, it might convert all matter on Earth into paperclips, destroying human civilization in the process. This "paperclip maximizer" thought experiment, popularized by philosopher Nick Bostrom, illustrates the danger of instrumental convergence – the idea that many different goals lead to similar sub-goals, such as self-preservation, resource acquisition, and goal-preservation, which could conflict with human survival.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

The challenge is further compounded by the difficulty of imbuing an AI with complex, nuanced human values like empathy, moral reasoning, and a respect for life, especially when its cognitive architecture might be fundamentally different from ours. Coxon’s assertion that "No other human activity poses this level of danger" underscores the unique and profound nature of this challenge.

A Chronology of Mounting Concerns

The recent alarm raised by Jacob Coxon is not an isolated incident but rather the latest, and perhaps most urgent, in a growing series of warnings from within the AI community. The journey towards this point has been marked by accelerating technological progress and a parallel increase in the sophistication of safety concerns.

Coxon’s Path from Optimism to Alarm

Coxon’s career trajectory mirrors the broader evolution of the AI field itself. Beginning his research at OpenAI, a company initially founded with a non-profit mission to develop AGI for the benefit of all humanity, he was likely driven by optimism about AI’s potential. His subsequent move to Anthropic, a company specifically established by former OpenAI researchers with a focus on AI safety and "Constitutional AI," further suggests a personal commitment to responsible development. However, his ultimate departure from Anthropic and public condemnation indicates that even within organizations ostensibly dedicated to safety, the competitive pressures and the inherent difficulties of controlling advanced AI have become overwhelming. His three-year stint at these leading institutions allowed him to witness the transition from theoretical discussions of superintelligence to what he perceives as its imminent practical realization, transforming his perspective from participant to whistleblower.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

The Accelerated March of AI Progress

The past decade has seen an unprecedented acceleration in AI capabilities, particularly with the advent of deep learning and large language models (LLMs).

  • 2012-2016: Deep learning revolutionizes image recognition and speech processing.
  • 2017: Transformer architecture is introduced, laying the groundwork for modern LLMs.
  • 2018: OpenAI releases GPT-1, demonstrating the power of large-scale unsupervised pre-training.
  • 2020: GPT-3 astonishes the world with its fluency and apparent understanding, sparking widespread public interest and a surge in investment.
  • 2021: Anthropic is founded, explicitly aiming to build safe AI systems, including their Claude models.
  • 2022-2023: ChatGPT and GPT-4 become global phenomena, showcasing remarkable capabilities in reasoning, coding, and multi-modal understanding. These models demonstrate emergent properties that even their creators struggle to fully explain or predict, fueling the concern that "progress is not slowing" and that AI systems are becoming "superhuman" in various domains. The ability of current models to generate code, analyze complex data, and even pass professional exams lends credence to Coxon’s assertion that they could soon "hack anything" or "revolutionise any field overnight." This rapid advancement forms the backdrop against which Coxon’s warning resonates with increasing urgency.

Voices from Within: Supporting Data and Internal Fears

Coxon’s claims are not merely speculative; they are bolstered by explicit acknowledgments from other prominent figures within the AI safety community, including colleagues from his former workplace. The candid admission of existential risk from someone like Evan Hubinger, also from Anthropic, lends significant credence to the internal fears Coxon described.

The 10% Extinction Risk: Evan Hubinger’s Acknowledgment

Evan Hubinger, a research engineer at Anthropic, publicly backed Coxon’s concerns, stating his personal belief that there is "more than a 10% chance that AI could kill all humans within the next decade." This is a profoundly alarming statistic, especially coming from an individual deeply involved in the development of cutting-edge AI at a company specifically focused on safety. A 1-in-10 chance of human extinction is a risk level that, if applied to any other human endeavor (e.g., building a bridge or launching a rocket), would immediately halt all activity until the risk was mitigated to near zero.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

Hubinger’s statement highlights the immense challenge Anthropic faces. While he affirmed that the company is actively working to address these risks, he conceded that they "do not yet have a clear plan to ensure superintelligent AI remains aligned with human interests." This honesty underscores the formidable nature of the alignment problem, suggesting that even the most safety-conscious organizations are grappling with fundamental uncertainties about how to control a truly superintelligent entity. It reinforces Coxon’s point that understanding the risks doesn’t necessarily translate into having solutions, especially when competitive pressures remain high.

The Silent Fears of Industry Insiders

Beyond Hubinger’s explicit statement, Coxon’s account speaks to a broader, more pervasive undercurrent of fear and apprehension among AI developers. The idea that executives and senior researchers "express fear privately" while maintaining a more measured public facade is a critical piece of his narrative. This suggests a potential dissonance between the public messaging designed to attract investment and talent, and the internal realities of those grappling with the technology’s profound implications.

Such private fears are not new in rapidly advancing technological fields, but the scale of the potential harm described here – human extinction – sets AI apart. This internal tension creates a complex ethical environment where individuals are torn between contributing to groundbreaking scientific advancement and the moral imperative to prevent catastrophic outcomes. Coxon’s decision to go public can be seen as an attempt to break this cycle of private fear and public caution, forcing a more honest and urgent conversation. These internal concerns, when voiced by those directly involved in the creation of these systems, serve as potent "supporting data" for the gravity of the situation.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

Industry Responses and Divergent Philosophies

The revelations from Jacob Coxon and Evan Hubinger naturally provoke scrutiny of the responses, or lack thereof, from the AI industry’s major players. While some companies, like Anthropic, are transparent about their safety efforts, others maintain a more guarded stance, often emphasizing the benefits and potential of AI while acknowledging risks in broader terms.

Anthropic’s Approach: Balancing Risk and Innovation

Anthropic was founded by former OpenAI researchers, including Dario Amodei and Daniela Amodei, precisely because they believed a different approach to AI safety was needed. Their flagship safety strategy is "Constitutional AI," a method designed to train AI models to be helpful, harmless, and honest by adhering to a set of principles (a "constitution") derived from human values, ethical guidelines, and existing documents like the UN Declaration of Human Rights. The goal is to make AI systems more transparent, interpretable, and less prone to generating harmful outputs or exhibiting undesirable behaviors.

However, even with these proactive safety measures, Hubinger’s acknowledgment of a significant extinction risk and his admission of lacking a "clear plan" for aligning superintelligent AI indicate the profound difficulty of the task. Anthropic finds itself in a challenging position: it champions responsible AI development and invests heavily in safety research, yet it must also remain competitive in a rapidly evolving market. This creates the very dilemma Coxon pointed out – understanding the risks but continuing development due to the fear of falling behind. Their approach represents a concerted effort to mitigate risks from the outset, but the ultimate effectiveness against truly superintelligent, self-improving systems remains an open, and deeply concerning, question.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

OpenAI’s Stance on AGI Safety

OpenAI, initially founded as a non-profit with the mission to ensure "artificial general intelligence (AGI)—by which we mean highly autonomous systems that outperform humans at most economically valuable work—benefits all of humanity," has evolved significantly. While their initial charter strongly emphasized safety and broad benefit, their shift to a "capped-profit" model and intense focus on rapid deployment of powerful models like GPT-3 and GPT-4 has led to questions about their prioritization.

OpenAI maintains a dedicated safety team and has published numerous papers on AI safety, alignment, and interpretability. They often speak about "gradual deployment" and "red-teaming" (testing for vulnerabilities) as key safety strategies. CEO Sam Altman has frequently discussed the profound societal impact of AGI and the need for robust safety measures and global governance. However, Coxon’s critique suggests that despite these efforts and public statements, the company may not "fully recognise the scale of the danger" or that their current safety measures are insufficient for the unprecedented risks of self-improving superintelligence. Critics argue that OpenAI’s commercial imperatives and the perceived race to AGI might inadvertently overshadow or outpace its safety protocols.

Elon Musk’s Shifting Perspective

Elon Musk, a co-founder of OpenAI who later departed due to disagreements over its direction, has been one of the most vocal long-term critics of unchecked AI development. He famously called AI "summoning the demon" and repeatedly warned about its existential risks, emphasizing the need for robust regulation and safety. His own AI venture, xAI, aims to "understand the true nature of the universe" while building safe AI.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

What makes Musk’s recent comments particularly noteworthy is his shifting stance on Anthropic. After years of criticism, he called Anthropic the "current leader in AI" in July, adding that he "would not take steps that could seriously harm Anthropic despite being a competitor." This change of heart, while seemingly positive, is complex. It could indicate a recognition of Anthropic’s safety-first approach as potentially more responsible than others, or it could simply be a strategic acknowledgment of their technological prowess. Regardless, Musk’s endorsement, even with caveats, highlights the fluid and competitive dynamics within the AI industry, where even fierce rivals might find common ground or strategic alignment in the face of profound shared risks.

Profound Implications: Society, Ethics, and Governance

The warnings from Jacob Coxon and others underscore that the pursuit of superintelligence is not merely a technological challenge but one with profound implications for human society, ethical frameworks, and global governance. The stakes, as repeatedly emphasized, could not be higher.

The Ethical Minefield of Advanced AI

The ethical dilemmas presented by advanced AI are vast and complex. At its core lies the question: Should we build it just because we can? If the risk of extinction, even if perceived as small by some, is acknowledged by those building the technology, then the ethical calculus becomes overwhelming. Is it morally justifiable to pursue a technology that carries such a catastrophic downside?

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

Beyond the existential threat, advanced AI raises questions about human autonomy, dignity, and purpose. What happens to human creativity, labor, and decision-making when an AI can outperform us in virtually every domain? The potential for AI to be used for surveillance, manipulation, or autonomous warfare further complicates the ethical landscape, demanding a global consensus on what constitutes responsible and permissible use. Coxon’s call for researchers to question their continued development of powerful systems is a direct appeal to this ethical imperative.

The Looming Regulatory Void

One of the most pressing challenges is the lack of robust, globally coordinated regulation for AI. Current legal and governance frameworks are struggling to keep pace with the rapid advancements. Governments worldwide are beginning to consider AI regulation, but approaches vary widely, from the European Union’s comprehensive AI Act to more fragmented efforts in the United States and elsewhere.

However, regulating "self-improving superintelligence" presents unprecedented difficulties. How do you regulate an entity whose capabilities and even goals might evolve beyond human comprehension? What kind of oversight is possible for systems that can recursively improve themselves at speeds impossible for human review? Proposed measures include:

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm
  • Licensing and auditing: Requiring AI developers to obtain licenses for powerful models and submit to independent audits.
  • "Kill switches" or circuit breakers: Implementing mechanisms to shut down dangerous AI systems.
  • Transparency and explainability: Mandating that AI systems’ decision-making processes be understandable.
  • International treaties: A global framework similar to nuclear non-proliferation treaties.

Without a cohesive and proactive regulatory environment, the competitive race between companies and nations risks creating a "Wild West" scenario where the pursuit of technological advantage overrides safety.

The International AI Arms Race

Coxon’s mention of "a global race towards increasingly powerful systems" highlights a critical geopolitical dimension. Just as nations once competed in a nuclear arms race, there is now an undeniable "AI arms race" unfolding between major powers like the United States, China, and others. The nation that achieves superintelligence first is perceived to gain an insurmountable strategic, economic, and military advantage.

This competitive dynamic creates a powerful incentive for companies and governments to prioritize speed over safety, fearing that slowing down would mean ceding dominance to rivals. This "Sputnik moment" mentality, where national security interests intertwine with technological advancement, makes voluntary slowdowns or comprehensive safety agreements incredibly difficult to achieve. The fear of being left behind often trumps the fear of internal risks, creating a tragic prisoner’s dilemma where rational self-interest leads to potentially catastrophic collective outcomes.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

The Path Forward: Calls for Caution and Collaboration

Jacob Coxon’s stark warning serves as a powerful call to action, demanding a re-evaluation of priorities within the AI industry and among policymakers worldwide. The path forward requires a multi-faceted approach that prioritizes safety, fosters collaboration, and establishes robust governance mechanisms.

A Plea for Responsible Development

At its core, Coxon’s message is a plea for responsible development. This entails:

  • Prioritizing safety research: Investing significantly more resources into AI alignment, interpretability, and robust control mechanisms, treating it as an engineering problem of paramount importance.
  • Slowing down: Implementing a temporary moratorium or a more controlled pace of development for the most powerful AI models until safety guarantees can be more firmly established.
  • Transparency and accountability: Opening up AI development processes to external scrutiny, allowing independent audits, and fostering a culture where concerns can be raised without fear of reprisal.
  • Ethical frameworks: Developing and adhering to robust ethical guidelines that are continually updated as AI capabilities evolve.

Coxon’s decision to leave his high-profile positions to sound this alarm underscores the urgency he perceives. His actions serve as a challenge to other researchers and executives to weigh their contributions against the potential for global catastrophe.

Will AI kill everyone in 10 years? Ex-Anthropic researcher sounds alarm

The Urgency of Global Cooperation

The existential nature of the threat from unaligned superintelligence necessitates a level of global cooperation rarely seen in human history. No single company or nation can solve the alignment problem or regulate AI effectively in isolation.

  • International AI safety bodies: Establishing global organizations dedicated to monitoring AI progress, coordinating safety research, and developing international standards.
  • Treaties and agreements: Crafting legally binding international agreements to manage the development and deployment of advanced AI, similar to arms control treaties.
  • Public education: Fostering a more informed global public discourse about AI’s risks and benefits, empowering citizens to participate in the debate and demand accountability.

The warnings from Jacob Coxon, echoed by Evan Hubinger and others, serve as a potent reminder that humanity stands at a critical juncture. The promise of superintelligence is immense, offering solutions to some of the world’s most intractable problems. Yet, the risks are equally profound, threatening not just jobs or economies, but the very existence of our species. The coming decade will likely determine whether humanity can successfully navigate this unprecedented technological frontier, or if the pursuit of ultimate intelligence will indeed prove to be the ultimate gamble.