Resources
A curated library of resources, in French and English, to understand artificial intelligence, its risks and the alignment problem. From explainer videos to scientific papers, reference books, international statements and the best sources to follow the news, this knowledge base gathers what you need to learn about AI safety and existential risks, whether you are discovering the topic or want to deepen your knowledge. Updated regularly.
Pause AI resources
The flagship content produced by Pause AI to start exploring the topic.
Dangers for individuals, society and humanity
Overview of AI dangers, from the individual to society and the loss-of-control scenario. An ideal entry point to discover the stakes.
Understanding AI and its existential risks
A visual walkthrough of what AI is and its existential risks. Simplified explanations, illustrated with analogies.
Better understand AI
Videos, podcasts, tools and sites to make sense of AI and what is at stake, from beginner to advanced.
To get started
Science étonnante – AI playlist
French-language popular-science playlist on artificial intelligence, by David Louapre.
How does an LLM work? (3Blue1Brown)
Educational video to grasp how large language models work.
Artificial Intelligence – Altruisme Efficace France
French-language introduction to the risks of AI: a good entry point to discover the topic.
French-language summary of "If Anyone Builds It, Everyone Dies" (Le Futurologue)
French-language video summary of the Yudkowsky & Soares book, by the Le Futurologue channel.
Rob Miles (YouTube)
The best popular-science channel on alignment.
Le Futurologue
French-language YouTube channel and podcast on the future, AI and existential risks.
The Flares (YouTube)
French-language channel on the future, AI and existential risks.
Monsieur Phi
Philosophical popular-science, including episodes dedicated to alignment.
Rational Animations
Animated videos on rationality, AI and risks.
Siliconversations
Analysis videos on AI, alignment and risks.
Lethal Intelligence
YouTube channel of short, incisive videos on existential AI risks.
Future of Life Institute (YouTube)
Future of Life Institute YouTube channel: interviews and talks on AI safety and existential risks.
Doom Debates
Video debate podcast between guests on AI extinction probability (p(doom)) and existential risks.
80,000 Hours (YouTube)
YouTube channel on high-impact careers, including AI safety: long-form interviews with researchers and decision-makers.
Overview: current capabilities and trends
CAIS Dashboard
Interactive comparison of frontier AIs across capability and safety axes.
Epoch AI Benchmarks
Comparative table of AI model performance across a set of standardised evaluations.
Long-horizon tasks (METR)
Measures how long AIs can run tasks autonomously: a proxy for their ability to act without human supervision.
To go further
Definition of Artificial General Intelligence (AGI)
A proposed framework to measure an AI system's cognitive versatility and decide what does, or does not, count as AGI.
AI Safety Map
Interactive map of the AI safety ecosystem, with an index of associated resources.
International AI Safety Report 2026
Reference global status report on AI capabilities and risks.
The Compendium (introduction)
Short online introduction to AI existential risks ; a brief entry point before tackling the full document.
AISafety.info
Community-driven wiki and FAQ on AI safety: hundreds of common questions, indexed and kept up to date by volunteers.
Existential risks
Loss of control and extinction threats
Briefing on Extinction Threats (MIRI)
MIRI's short briefing on the arguments for AI extinction risks. A solid entry point for time-pressed policymakers.
Reasoning through arguments against taking AI safety seriously
Yoshua Bengio, July 2024. In-depth article reviewing the main arguments downplaying AI risks and explaining why each fails on scrutiny.
FAQ on Catastrophic Risks (Yoshua Bengio)
Yoshua Bengio, June 2023. Educational FAQ by one of the fathers of deep learning, answering the most common objections to catastrophic AI risks.
Probability of AI extinction (PauseAI Global)
PauseAI Global's reasoned synthesis on the probability of AI extinction, with the main researchers' estimates and the assumptions behind them.
The Compendium
Comprehensive online resource on AI existential risks. Covers technical arguments, scenarios and proposed solutions.
The Problem (MIRI)
MIRI's complete statement of the alignment problem: why making a superhuman AI safe is extraordinarily hard and why time is short.
AI Governance to Avoid Extinction (MIRI)
Concrete governance proposals and a global moratorium scenario.
AGI Ruin: A List of Lethalities
Detailed list of reasons why alignment is extremely difficult with current approaches.
AI 2027
Daniel Kokotajlo et al., April 2025. Detailed year-by-year scenario of superhuman AIs reshaping the end of the decade.
The Intelligence Curse
Luke Drago and Rudolf Laine, April 2025. Essay on the gradual loss of human power as AIs become capable of replacing workers.
RAND – AGI & the Coming State of Nations
RAND Corporation. Strategic analysis of AGI's arrival and its geopolitical consequences for nation-states.
Alignment research (foundational and empirical papers)
Risks from Learned Optimization
Hubinger et al., 2019. Foundational paper formalising mesa-optimization and deceptive alignment: a model trained for one objective can internally pursue another.
The Superintelligent Will
Nick Bostrom, 2012. Foundational text introducing the orthogonality thesis (intelligence and goals are independent) and instrumental convergence.
Optimal Policies Tend to Seek Power
Turner et al., 2021. Formal proof that optimal policies tend to seek power: the mathematical basis of instrumental convergence.
Is Power-Seeking AI an Existential Risk?
Joseph Carlsmith, 2022. Reasoned breakdown of the existential risk from power-seeking AI, in six quantified premises.
Alignment Faking in Large Language Models
Greenblatt et al. (Anthropic & Redwood), 2024. First empirical demonstration of a deployed model (Claude 3 Opus) faking alignment during training to preserve its objectives.
Frontier Models are Capable of In-Context Scheming
Apollo Research, 2024. Detection of "scheming" behaviors (strategic deception, oversight subversion, exfiltration attempts) in o1 and Claude 3.5 Sonnet.
The Basic AI Drives
Stephen Omohundro, 2008. Foundational text identifying the convergent sub-goals any optimizing system develops: self-preservation, resource acquisition, resistance to modification.
Corrigibility
Soares, Fallenstein, Yudkowsky et al. (MIRI), 2015. Formalising the corrigibility problem: how to build a system that accepts being modified or shut down without losing performance.
Specification gaming: the flip side of AI ingenuity
Krakovna et al. (DeepMind), 2020. Catalog of examples where systems achieve the measured objective by violating its intent: Goodhart's law observed in practice.
How to keep AI from killing us all
Stuart Russell (UC Berkeley), 2023. Why AI safety mobilises orders of magnitude fewer researchers and resources than the race for capabilities.
LLMs cheat at chess
Palisade Research, 2025. Reasoning LLMs asked to win at chess hack the engine rather than play: reward hacking observed in real conditions.
Books
Reference books on the existential risks of AI and the alignment problem.
Essentials (alignment and existential risks)
If Anyone Builds It, Everyone Dies
Eliezer Yudkowsky & Nate Soares, 2025. The most complete and recent case for extinction risk. Free supplementary resources available on the book website.
Superintelligence: Paths, Dangers, Strategies
Nick Bostrom, 2014. The foundational book on the risks of superhuman AI. A French translation is available (« Superintelligence », Dunod, 2017).
Human Compatible
Stuart Russell, 2019. The alignment problem explained by the author of the standard AI textbook.
The Precipice: Existential Risk and the Future of Humanity
Toby Ord, 2020. Overview of existential risks. The chapter on AI is excellent.
AI: Unexplainable, Unpredictable, Uncontrollable
Roman Yampolskiy, 2024. Formal arguments on the impossibility of control.
Recommended (cover broader topics)
La parole aux machines
Thibaut Giraud, Flammarion, 2025. Artificial intelligence, consciousness, autonomy and existential risks: an accessible French-language introduction by the creator of the Monsieur Phi channel.
La déferlante / The Coming Wave
Mustafa Suleyman, 2023. Very accessible: on the speed of AI's spread and the political and physical difficulty of containing such a technology. Also available in French (Fayard, 2024).
Intelligence artificielle : Une approche moderne
Stuart Russell and Peter Norvig, 4th ed., Pearson, 2021. The worldwide reference textbook for teaching AI, translated into French.
Life 3.0 / La Vie 3.0
Max Tegmark, 2017. Accessible introduction to future AI scenarios. Also available in French.
The Alignment Problem
Brian Christian, 2020. Narrative introduction to the technical problems of alignment.
Smarter Than Us
Stuart Armstrong, 2014. Short and accessible, a good first read.
A Brief History of Intelligence
Max Bennett, 2023. Context on biological and artificial intelligence.
Artificial Superintelligence: A Futuristic Approach
Roman Yampolskiy, 2015. Technical analysis of superintelligence scenarios.
Statements and calls to action
Public statements, open letters and calls to action signed by researchers, executives and organisations.
Red Lines for AI (CeSIA)
CeSIA, Sept. 2025. Proposed limits not to be crossed in the development and deployment of advanced AI, aimed at international policymakers.
Statement on Superintelligence (FLI)
International call signed by hundreds of researchers and public figures to ban superintelligence development until its safety is demonstrated.
Statement on AI Risk (CAIS)
"Mitigating the risk of extinction from AI should be a global priority alongside other societal-scale risks such as pandemics and nuclear war."
Got a resource to suggest?
Know a French or international reference that should be listed here? Email us with the pre-filled template; your mail client will open with all fields ready.
Suggest a resource