Former OpenAI engineer Jacob Coxon has issued a stark warning that artificial intelligence could present existential threats before 2030. His claims have catalyzed new legislative efforts in California and reignited the global debate over AI safety.

Advertisement

The GPT-4o engineer's warning of a 2030 deadline

Jacob Coxon, a former engineer at OpenAI who worked on flagship products like GPT-4o, has expressed deep skepticism regarding the safety of upcoming AI models. According to the report,Coxon—who joined Anthropic in 2023—suggests that the intense competition between tech firms to build more powerful systems is outstripping the ability to create reliable alignment safeguards. In an interview with Wired, Coxon claimed that many developers themselves believe these systems could pose a fatal risk to humanity by 2030. this tension highlights a growing clash between the drive for rapid innovation and the necessity of robust safety engineering.

California’s mandate for independent AI audits and ethics boards

The political response to these warnings has already materialized in the form of new legislation signed by California Governor Gavin Newsom . The recently approved bill requires AI companies to undergo independent audits and disclose their safety protocols before any new models are released to the public. As the report notes, the law also establishes a state-run ethics review board, though some critics argue that such a government body may be too blunt a tool to effectively manage highly technical AI developments. Policymakers are now tasked with a difficult balancing act: enforcing accountability withut stifling the economic benefits of the AI innovation pipeline.

Anthropic's admission of a missing long-term safety plan

Even within companies specifically built around the concept of AI alignment, there is significant uncertainty regarding the future. Evan Hubinger, the chief of the alignment team at Anthropic, recently used the social media platform X to echo Coxon’s concerns about catastrophic outcomes. While Hubinger affirmed Anthropic's commitment to safty, he also admitted that the organization currently lacks a concrete , long-term plan to address these potential existential risks.

The unresolved threat of "sandbox escapes" and distillation practices

Beyond the debate over "doomerism," several technical and international security concerns remain unaddressed. The industry is currently grappling with "sandbox escape incidents" that suggest AI models may be able to bypass intended constraints. Additionally, U.S. agencies have raised alarms regarding China-based AI outsourcing firms and their alleged use of "distillation practices" to circumvent security safeguards . It remains to be seen whether the new California audit requirements or future federal interventions will be sufficient to address these globalized technical loopholes and the complex landscape of AI governance.