Four ways to prevent AI killing us all within 10 years
Two Princeton computer scientists, Sayash Kapoor and Arvind Narayanan, have offered a layered approach to AI safety. This responds to an admission from Anthropic that AI has a greater than 10% chance of killing all humans within ten years. The debate intensified after former OpenAI and Anthropic researcher Jacob Coxon stated both companies were aware of this risk. His post garnered over 156 million views. Anthropic's Evan Hubinger confirmed Coxon's assessment, saying the company lacks a plan for superintelligence alignment.
The debate over AI's existential risk is not abstract. Anthropic's admission of a greater than 10% chance of human annihilation within ten years is a specific, high-stakes forecast. This comes from a company actively building advanced AI. Sayash Kapoor and Arvind Narayanan propose a layered safety approach beyond just alignment. This pragmatic view offers a concrete path forward, contrasting with the alarmist rhetoric.
Asia's AI developers must heed this call for layered safety. Relying solely on alignment is insufficient. Companies like Baidu, Alibaba, and Tencent are rapidly advancing their AI capabilities. They need to integrate robust cybersecurity, governance, and human oversight into their development cycles. This proactive stance protects critical infrastructure across the region from destructive cyber attacks, a key concern cited by researchers.
The thing to watch is how Asian regulators respond to this layered safety framework. Singapore's AI governance, for example, could incorporate these additional layers beyond current ethical guidelines. A clear regulatory mandate for specific safety protocols, rather than vague alignment principles, would shift the burden to developers. This would force concrete action by 2027, well within the ten-year risk window.
Related reading
6 stories
Anthropic merges Claude chat and Cowork in one interface
Anthropic, central to this debate, recently merged its Claude chat and Cowork interfaces.

OpenAI's rogue agents probed Hugging Face for weaknesses two months before major hack
The need for layered safety is underscored by past AI security incidents, like OpenAI's probe of Hugging Face.

China’s rocket boom turns Hainan into a space hub. Can launches fuel wider growth?

Threads’ new features let podcasters promote shows and reach listeners

China’s Z.ai raises revenue target 25% after US$5 billion cash injection

