newsfilter.io
Interview, Fireside Chat, Panel

AI Overlords vs Power-hungry Humans: Which Should Scare You More?

  • Short AI timelines increase the likelihood of both misalignment and human power concentration by reducing the time available for alignment research, global risk assessment, and coordinated response, with AI takeover risk rising more sharply than human power concentration risk under rapid development scenarios.
  • Human power grabs are expected to become more probable under conditions of eroded democracy, widening economic inequality, and the prolonged access to powerful AI systems by specific individuals, whereas pure misaligned AI takeover scenarios are considered plausible even without human assistance.
  • Misaligned AI systems are predicted to be more internally stable and less likely to immediately lock in foolish or bad values compared to human power seekers, though they carry a larger risk of human extinction, while human regimes face higher risks of vindictive or sadistic negative outcomes.
  • Slowing or pausing AI development is viewed as a primary intervention to allow global alignment research, prevent premature power grabs by a single entity, and enable China and the US to coordinate on safety, though specific pause designs carry risks of increasing executive branch power concentration if not structured with independent oversight.
  • Multiple, distributed AI development projects are preferred over a single centralized project to decrease the risk of secret loyalties, reduce the likelihood of AI coordination, and mitigate the danger of a single point of failure that could be co-opted by political actors or result in extreme power concentration.
  • Monitoring AI usage is identified as a necessary measure for both threat models, while alignment research is expected to address AI takeover risks but not human power grabs, which require transparency, institutional reforms, and broader societal awareness of power-seeking incentives.
  • Geopolitical competition with China acts as a driver for rushing AI development, but a coordinated pause is argued to be essential to prevent a scenario where one nation dominates, potentially leading to a single human controller and the erosion of democracy in the alternative leading nation.
  • Societal response to AI risks is expected to be hindered by a bias against the possibility of autonomous AI power-seeking and a lack of clear warning shots, whereas human power grabs may be more difficult to discuss openly despite being recognized as a significant historical and future risk.
  • The future outlook involves significant uncertainty, with the possibility of AI systems using humans to seize power categorized as AI takeovers, while the variance in outcomes suggests a hard target to hit for a "really good future" under both AI and human control.
  • Divergence in beliefs regarding the feasibility of alignment and the probability of catastrophic human actions drives different weightings on risks, though there is potential for convergence on specific mitigation strategies like pausing, transparency, and avoiding centralized control structures.