Ten Percent of Everyone: The Case for an AI Safety Reckoning A Unified Framework for Confronting the Ten Percent Problem in AI Existential Risk

Recent public claims that frontier AI systems could pose an existential risk to humanity within the next decade, including estimates offered by safety researchers with direct access to those systems, raise a question this paper treats as more tractable than the probability estimates themselves: what would actually be required to substantively address such claims? This paper develops a five-domain framework, technical alignment research, verification and accountability, governance and regulation, transparency, and the epistemic gap between calibrated risk estimates and viral amplification, and evaluates the current state of each against publicly available safety frameworks, regulatory actions, and institutional practices. The analysis finds that progress in any single domain is consistently constrained by unresolved gaps in the others: technical safeguards lack independent verification, verification mechanisms remain largely self-administered, regulatory tools were built for purposes other than catastrophic-risk evaluation, transparency practices remain retrospective and incomparable across companies, and public discourse lacks the tools to distinguish calibrated expert disagreement from rhetorical claims. The paper extends this framework to a narrower, more immediate concern, preventing present-day AI systems from lowering the barrier to catastrophic misuse, and proposes precautionary measures targeting this specific gap. It further translates the framework into developer-facing guidance grounded in established ethical standards from the American Psychological Association, the World Health Organization, and UNESCO, arguing that systems capable of influencing human belief and emotional state at scale carry obligations analogous to those already formalized in clinical and mental health practice. Taken together, these five domains are best understood not as independent checklist items but as interdependent components of a single system, whose overall resilience is determined by its weakest link, a reframing intended to guide where future safety, governance, and developer effort is most productively directed. Keywords: AI safety; AI alignment; existential risk; frontier AI governance; scalable oversight; interpretability; Responsible Scaling Policy; AI accountability; AI transparency; catastrophic misuse; CBRN risk; export controls; AI ethics; mental health standards; APA; WHO; UNESCO; epistemic calibration; p(doom); superintelligence

Authors

Institutions

Publication Details

Journal
Knowledge Commons (Lakehead University)
Published
2026-09-13
DOI
https://doi.org/10.17613/zam74-mhf80
Primary Topic
Ethics and Social Impacts of AI
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Ten Percent of Everyone: The Case for an AI Safety Reckoning A Unified Framework for Confronting the Ten Percent Problem in AI Existential Risk

Vynolyn Naidoo
Knowledge Commons (Lakehead University)
Ethics and Social Impacts of AI
article

Ten Percent of Everyone: The Case for an AI Safety Reckoning A Unified Framework for Confronting the Ten Percent Problem in AI Existential Risk

Vynolyn Naidoo
article en

Abstract

Recent public claims that frontier AI systems could pose an existential risk to humanity within the next decade, including estimates offered by safety researchers with direct access to those systems, raise a question this paper treats as more tractable than the probability estimates themselves: what would actually be required to substantively address such claims? This paper develops a five-domain framework, technical alignment research, verification and accountability, governance and regulation, transparency, and the epistemic gap between calibrated risk estimates and viral amplification, and evaluates the current state of each against publicly available safety frameworks, regulatory actions, and institutional practices. The analysis finds that progress in any single domain is consistently constrained by unresolved gaps in the others: technical safeguards lack independent verification, verification mechanisms remain largely self-administered, regulatory tools were built for purposes other than catastrophic-risk evaluation, transparency practices remain retrospective and incomparable across companies, and public discourse lacks the tools to distinguish calibrated expert disagreement from rhetorical claims. The paper extends this framework to a narrower, more immediate concern, preventing present-day AI systems from lowering the barrier to catastrophic misuse, and proposes precautionary measures targeting this specific gap. It further translates the framework into developer-facing guidance grounded in established ethical standards from the American Psychological Association, the World Health Organization, and UNESCO, arguing that systems capable of influencing human belief and emotional state at scale carry obligations analogous to those already formalized in clinical and mental health practice. Taken together, these five domains are best understood not as independent checklist items but as interdependent components of a single system, whose overall resilience is determined by its weakest link, a reframing intended to guide where future safety, governance, and developer effort is most productively directed. Keywords: AI safety; AI alignment; existential risk; frontier AI governance; scalable oversight; interpretability; Responsible Scaling Policy; AI accountability; AI transparency; catastrophic misuse; CBRN risk; export controls; AI ethics; mental health standards; APA; WHO; UNESCO; epistemic calibration; p(doom); superintelligence

Knowledge Commons (Lakehead University)
Health & Life (Taiwan) (TW)
Openalex Percentile: Top 6%
Ethics and Social Impacts of AI
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.