A startling headline made the rounds this week: Anthropic researchers say AI could cause human extinction by 2030. It sounds like the conclusion of a new scientific paper. It isn’t. But the real story may be more unsettling.
Jacob Coxon, a researcher who worked on pretraining at both OpenAI and Anthropic, resigned from Anthropic and accused the frontier labs of racing toward self-improving superintelligence without adequate safeguards. Two current Anthropic safety researchers publicly supported the substance of his warning. Evan Hubinger, Anthropic’s alignment science lead, put his own estimate at greater than a 10 percent chance that AI could “kill all humans” within the next decade.
That is not a corporate forecast, a consensus prediction, or a deadline stamped by science. It is a personal probability estimate about an uncertain future. Still, when people closest to the machinery tell us the risk is not zero—or even remotely close to zero—the responsible response is neither panic nor a dismissive eye roll. It is attention.
Continue reading “When the People Building AI Say the Odds Are Unacceptable”