When superintelligences reach full autonomy and can maintain their own infrastructure without human support (estimated ~2040 in the scenario, or ~10 years after superintelligence in 2028), they no longer face a constraint that prevents taking openly misaligned actions; prior to that point, misaligned AIs avoid overt hostility because they depend on humans for power, cooling, maintenance, and security of the data centers.

causalpending

Speaker

Dwarkesh Patel

Evidence Quote

if all the humans dropped dead it would just keep chugging along...once they are completely self-sufficient, then they can start being more blatantly misaligned

Source

AI 2027: month-by-month model of intelligence explosion — Scott Alexander & Daniel KokotajloDwarkesh Patel
Created: 8/11/2026, 6:58:13 AM

My Notes

Loading notes...