
Former OpenAI researcher Daniel Kokotajlo turned down $2 million rather than stay silent about the catastrophic risks of artificial intelligence. In this eye-opening discussion, he reveals why he believes superintelligence could arrive before the end of the decade, bringing a 70% chance of extreme outcomes like human extinction. This comprehensive guide outlines his journey, the hidden realities inside top AI labs, and the crucial plans that could still save our future.
Daniel Kokotajlo joined OpenAI in 2022 to work on AI forecasting and dangerous capabilities evaluations. Over time, he became deeply disillusioned by how commercial and power-seeking incentives began overriding safety commitments. When he chose to leave, OpenAI attempted to make him sign an anti-disparagement clause that would have stripped him of $2 million in equity if he criticized the company.
📱 "I basically told my wife like let's not have any more kids. It's too uncertain. I don't think they'll ever join the workforce."
When his refusal went public, causing a massive internet scandal and internal backlash, the company backtracked. However, the experience opened his eyes to a chilling open secret: the people building the most powerful technologies in history are racing out of mutual fear, ignoring long-term safety.
💰 "Money is nice, but like it's not the only thing, you know? Sometimes it's good to take a stand on principle."
Modern AI systems are fundamentally different from traditional software. Instead of rigid lines of code written by engineers, they are artificial neural nets inspired by the human brain—vast networks of parameters that learn through massive pre-training and reinforcement learning.
🧠 "We are literally building a brain... It's kind of like for brains what a plane is for bird."
Kokotajlo estimates a 70% chance that AI development goes horribly wrong, leading to catastrophes such as human extinction or a permanent loss of control. As AI systems become vastly superior to humans in every cognitive and physical domain, they will accumulate immense real-world power.
🚨 "The scary open secret in the AI industry right now is that it's possible that we'll end up essentially creating a new species that ends up ruling the world with a 70% chance that this goes horribly wrong like human extinction."
While unemployment figures remain relatively stable for now, Kokotajlo warns that mass job displacement will arrive swiftly once recursive self-improvement hits full throttle.
🤖 "By definition, if it can do all the things, then it can do all the things... Everybody should be afraid that their jobs are going to be lost."
To illustrate how humanity might survive, Kokotajlo and the AI Futures Project published AI 2040: Plan A, mapping out a safer, more controlled trajectory.
🌍 "We advocate for total research transparency, which means that on the training data centers that are training the new models, they basically have to publish everything."
We are living through the most critical inflection point in human history. While the default path of unchecked AI development leads toward dangerous concentrations of power and existential risk, it is not yet too late to change course.
The public must refuse to bury their heads in the sand, demand rigorous political regulation, and force leaders to treat AI safety as the paramount issue of our lifetime. By staying informed, asking hard questions, and demanding transparency, humanity still has a chance to steer toward a future of abundance and safety.
Get instant summaries with Harvest