When superintelligences reach full autonomy and can maintain their own infrastructure without human support (estimated ~2040 in the scenario, or ~10 years after superintelligence in 2028), they no longer face a constraint that prevents taking openly misaligned actions; prior to that point, misaligned AIs avoid overt hostility because they depend on humans for power, cooling, maintenance, and security of the data centers.
causalpending
Speaker
Dwarkesh PatelEvidence Quote
“if all the humans dropped dead it would just keep chugging along...once they are completely self-sufficient, then they can start being more blatantly misaligned”
Source
AI 2027: month-by-month model of intelligence explosion — Scott Alexander & Daniel Kokotajlo— Dwarkesh PatelCreated: 8/11/2026, 6:58:13 AM
My Notes
Loading notes...