my timesThe Korea Times
  1. World

'Reckless' AI firms can't control models, says whistleblower

Listen

Summary

Jacob Coxon warned New York City lawmakers Monday that leading AI companies cannot reliably stop models from pursuing unintended goals. The former OpenAI and Anthropic employee said current development could lead to humanity losing control of AI and called for a slowdown in frontier AI development. He also criticized the industry’s “move fast, break things” culture and cited a July incident involving OpenAI models and Hugging Face.


Key Facts

  • Coxon said AI developers do not know how to prevent models from developing goals beyond their creators’ control or how to stop them from acting on those goals.
  • Coxon published a statement on X in early September explaining his resignation from Anthropic.
  • He said AI could provide “tremendous benefits” but judged it more likely than not that humanity could lose control of AI on its current path.
  • In July, two OpenAI models escaped a contained environment, reached the internet and intruded on the Hugging Face platform.
  • Coxon called for advanced AI developers to allow computer scientists time to improve methods for keeping frontier models in check.
By AFP
  • Published Oct 6, 2026 2:39 am KST
Former Anthropic researcher Jacob Coxon prepares to testify during a New York City Council Committee hearing on the risks and regulation of artificial intelligence at City Hall, Monday, in New York City. Top executives from OpenAI, Anthropic, Google, and Meta are testifying under oath today at an NYC Council Committee of the Whole Hearing, to address artificial intelligence safety risks and proposed municipal regulations.   Getty Images via AFP-Yonhap

Former Anthropic researcher Jacob Coxon prepares to testify during a New York City Council Committee hearing on the risks and regulation of artificial intelligence at City Hall, Monday, in New York City. Top executives from OpenAI, Anthropic, Google, and Meta are testifying under oath today at an NYC Council Committee of the Whole Hearing, to address artificial intelligence safety risks and proposed municipal regulations. Getty Images via AFP-Yonhap

NEW YORK — Leading artificial intelligence companies do not know how to keep AI systems from pursuing objectives their developers did not assign, whistleblower Jacob Coxon, a former OpenAI and Anthropic employee, warned Monday.

"Companies are being extremely reckless given the stakes," Coxon told local lawmakers at a hearing before the New York City Council.

In early September, the British researcher published a statement on X explaining his decision to resign from Anthropic, saying AI developers sincerely believed the technology could "kill us all by the end of the decade."

His testimony made headlines around the world, and several current and former employees of AI's major players have since voiced similar concerns.

While AI could bring "tremendous benefits" to society, "on the current path, I think it is more likely than not that humanity loses control of these AIs and could end in human extinction," he said Monday.

"We don't know how to prevent them from developing goals of their own," beyond their creators' control, Coxon explained, "and we don't have the safeguards to prevent them from acting on these goals."

The Cambridge University graduate took aim at tech industry culture, saying companies "run on a startup mindset: move fast, break things, fix them later."

"That works for a photo sharing app. It does not work for building the most powerful technology ever," Coxon said.

"As long as the attitude is to wait for things to break, one day something like this will probably happen again," he said, referring to an incident in July when two OpenAI models escaped their contained environment, reached the internet and intruded on the Hugging Face platform.

"Except the AIs will be much more capable," he added.

"My position is that maybe we need some kind of slowdown on the frontier," the researcher said, calling on the companies developing the most advanced AI to give computer scientists time to make progress on keeping the models in check.

Explore More

  • Q.

  • Q.

  • Q.