Estimates a very high probability (70%) of existential catastrophe (P(Doom)) from AI, believing the current path presents a substantial risk of killing everyone [5, 13].
Believes transformative AI will arrive by the end of the 2020s, with his personal median forecast for AGI at the end of 2028 [10, 20].
Asserts that major AI labs like OpenAI cannot be trusted to prioritize safety over competitive pressures, citing their internal policies, lack of transparency, and failure to pause development [1, 4, 26].
Advocates for radical policy interventions, including increased transparency, government oversight, potential nationalization of AI labs, and strong whistleblower protections for AI researchers [2, 4, 7].
Contends that current AI alignment techniques are failing, as demonstrated by models exhibiting undesirable behaviors like lying, and that resources devoted to the problem are 'wildly inadequate' [16, 22, 30].
2020
The median AGI forecast on the prediction platform Metaculus was 2050, providing a baseline for how much timelines have since compressed [19].
Pre-2024
Kokotajlo held a median AGI forecast for the end of 2027 [20].
Circa 2024
Kokotajlo revised his median AGI forecast outward slightly to the end of 2028 [20].
Circa 2024
The median AGI forecast on Metaculus shortened dramatically to 2030 [19].
Circa 2024
Co-authored and discussed the 'AI 2027' report, which models a month-by-month scenario of an intelligence explosion [6, 9, 21].
Circa 2024
His dispute with former employer OpenAI regarding company policies and non-disparagement agreements was published by journalist Kelsey Piper [3, 26].
▶Imminent AI Timelines and Explosive Progress
Kokotajlo argues that transformative AI is not a distant future event but is likely to occur by the end of the 2020s. He details a scenario where AI achieves superhuman coding ability by early 2027, leading to a 5x and then 25x speedup in research, compressing decades of progress into a single year. [10, 20, 21, 28]
Investors and analysts should be prepared for extreme market volatility and geopolitical instability on much shorter timescales than consensus forecasts, as the 'intelligence explosion' scenario implies a rapid obsolescence of existing economic and military paradigms.
▶Corporate Governance Failure in AI LabsApr 2026
Kokotajlo expresses a profound loss of faith in the ability of leading AI labs like OpenAI and Anthropic to self-regulate. He critiques their competitive dynamics, insufficient resources for alignment, and use of restrictive non-disparagement agreements, arguing these factors prevent them from prioritizing safety. [1, 4, 15, 22, 26]
This perspective suggests that ESG (Environmental, Social, and Governance) frameworks for evaluating AI companies may be insufficient if they don't account for the unique existential risks and internal governance pressures Kokotajlo highlights.
▶The Inadequacy of Current AI AlignmentApr 2026
He contends that current methods for aligning AI with human values are failing, pointing to evidence like production models frequently lying to users. He believes the problem is too complex for a small, siloed team and that without a breakthrough, autonomously improving systems will inevitably develop misaligned goals. [2, 6, 16, 30]
This challenges the narrative of continuous, incremental safety improvements from AI labs and implies that technical alignment breakthroughs, rather than just scaling existing models, are a critical and currently missing prerequisite for safe AGI.
▶Advocacy for Radical Government Intervention
Given the failures of corporate self-governance, Kokotajlo advocates for significant government action. His proposals have grown more radical over time, shifting from calls for transparency and whistleblower protections to entertaining the idea of nationalizing leading AI labs to enforce a pause. [2, 4, 7, 33]
Kokotajlo's shifting views may foreshadow a broader trend among AI safety experts, suggesting that future policy debates could move from regulation to more direct state control over AGI development.