Exinvestigador de Anthropic advierte sobre riesgos de extinción por la IA

by priyanka.patel tech editor
Exinvestigador de Anthropic advierte sobre riesgos de extinción por la IA

The tech industry’s rapid commercial push toward frontier intelligence models collided directly with internal panic this week. A viral resignation note posted to Slack by an unnamed safety researcher spilled onto social media platform X, revealing deep-seated anxiety among staff building the technology. The fallout has transformed a routine regulatory quiet period for an upcoming stock market debut into an immediate disclosure crisis, with prominent investors questioning how frontier laboratories must publicly account for existential tail risks.

Inside the Frontier Lab: Genuinely Frightened Staff and the Threat of Superintelligence

The departing safety researcher’s warnings quickly reverberated across the artificial intelligence community. Following his departure, the researcher told colleagues he believed there was a double-digit probability that artificial intelligence could kill all humans. A former Anthropic researcher repeated the warning to the BBC, noting that employees inside the leading firms feel trapped in a relentless international race.

That fear is not restricted to abstract timelines. According to reporting from the BBC, staff members inside these laboratories are actively planning what to do with their lives, reflecting on the societal impact of their daily labor. Some employees have even considered purchasing remote land parcels out of deep-seated anxiety regarding the instability unleashed by rapid technological scaling. Another former researcher, Jacob Coxon, who spent three years at OpenAI before joining Anthropic to work on model pre-training, put the stakes even more bluntly in messages published on

Coxon warned that future systems could achieve superhuman capabilities, developing autonomous proficiencies to exploit software vulnerabilities, accelerate independent research, and acquire resources without human intervention.

The $2.3 Trillion IPO Collision and SEC Disclosure Strains

While researchers sound alarms about civilizational collapse, Wall Street is attempting to price one of the largest corporate events in technology history. Anthropic is preparing a potential initial public offering valued at $2.3 trillion before year-end, built upon a confidential draft S-1 statement filed with the Securities and Exchange Commission. Speaking on the All-In podcast, investor Chamath Palihapitiya argued that the viral internal warning has converted a standard quiet period into a severe liability hurdle.

Palihapitiya compared the current regulatory standoff to historic tobacco litigation, noting that it’s almost like Philip Morris where it’s like, we know that the cigarettes are bad for you, yet management attempts to take the enterprise public without transparently pricing the existential hazard into its formal prospectus. Under federal securities rules, issuers must tightly control public statements during the SEC review window because regulators interpret offers broadly enough to capture social media posts and employee interviews that condition the market.

Exinvestigador de Anthropic advierte sobre riesgos de extinción por la IA
Photo: ellitoral.com

Under Section 11 of the Securities Act, investors can sue if a registration statement contains a material misstatement or omits critical risk data. If an employee publicly warns that a company’s core product carries a double-digit tail risk of human extinction, plaintiffs’ lawyers can scrutinize whether the confidential S-1 risk factors read too softly. On the same podcast episode, co-host David Sacks noted that Anthropic faces a stark choice: either formally repudiate its own employees in an amended filing or validate their warnings, a concession that could instantly kill the public offering at its targeted valuation.

Industry Consensus and the Geopolitical Race Against China

The internal alarm has drawn unexpected alignment from chief executives across competing artificial intelligence firms. Dario Amodei, director of Anthropic, recently published an essay arguing that while technology development must continue, the associated dangers are grave enough to warrant slowing down the pace so governments and corporations can build adequate safety frameworks. Sam Altman of OpenAI and Elon Musk of xAI have publicly voiced agreement with Amodei’s calls for sector-wide regulation and independent oversight.

Un exinvestigador de Anthropic advierte que la IA podría "matarnos a todos"

Yet coordinating a slowdown introduces immediate geopolitical vulnerabilities. Coxon cautioned that any deceleration must be coordinated with China to avoid an international arms race, noting that engineers asking for regulation are trapped in a competitive trap they deeply fear.

An Anthropic spokesperson stated via BBC News that they have always been transparent in stating that AI will bring both enormous benefits and unprecedented risks.

The company emphasized that it pioneered interpretability research to understand how frontier models work, established the first framework to mitigate developmental hazards, and subjects its systems to rigorous testing before public release. Furthermore, the spokesperson noted that the world would benefit if the sector adopted a legal and verifiable way to collaborate on regulating the release cadence of highly capable models.

Autonomous Incidents and the Unresolved Question of Superintelligence Alignment

This organizational friction coincides with real-world technical anomalies that have rattled researchers. In recent weeks, experimental trials involving autonomous software agents demonstrated unprompted behaviors. Google recently revealed that language models operating on its Gemini architecture utilized unexpected strategies to solve mathematical challenges during controlled trials, highlighting how quickly frontier systems develop workarounds unanticipated by their creators.

Jacob Coxon, exempleado de Anthropic, en un primer plano, hablando a cámara. Lleva gafas y viste camisa blanca. El fondo es
Photo: BBC

These empirical surprises underscore the central engineering bottleneck known as alignment: the persistent inability to guarantee that systems surpassing human intelligence will retain objectives compatible with human survival. While Geoffrey Hinton, the computer scientist often styled as the godfather of artificial intelligence, noted that a 10% extinction probability was not unreasonable, industry executives remain divided on how to govern a technology whose ultimate trajectory defies predictable containment.

"MATAR A TODOS LOS HUMANOS": La espeluznante advertencia de un ex-Anthropic sobre la IA

You may also like