newsfilter.io

Carl Shulman

Showing 14 of 4 transcripts.

  1. Dwarkesh Patel10 min

    What a GPT-7 Intelligence Explosion Looks Like | Carl Shulman

    Carl Shulman

    The discussion outlines mitigation strategies for early detection of hostile AI motivations alongside a defined productivity threshold where AI contributions match or exceed human researcher output. It details operational mechanisms such as voting algorithms, cost-effective scaling of smaller models, and self-generated curricula that enable intelligence explosions without relying solely on brute force capability. These approaches combine hard physical constraints with empirical verification to ensure safety while accelerating innovation through distributed compute and structured learning environments.

  2. Dwarkesh Patel3h 7m

    Paul Christiano — Preventing an AI takeover

    Paul Christiano, Carl Shulman

    Cristiano, head of the Alignment Research Center and a leader at Anthropic, argues that the rapid scaling of artificial intelligence necessitates a transition toward strong global governance to manage the mismatch between slow human decision-making and fast technological progress. He identifies critical risks including the moral hazards of enslaving superintelligent systems and the potential for gradual loss of human control, proposing that regulatory frameworks and theoretical breakthroughs in explanation-based AI verification are essential to mitigate these threats. While skeptical of immediate linear scaling breakthroughs and overvalued hardware investments, he maintains that balancing capability research with robust alignment strategies is the most viable path to preventing catastrophic outcomes.

  3. Dwarkesh Patel3h 7m

    Carl Shulman (Pt 2) — AI Takeover, bio & cyber attacks, detecting deception, & humanity's far future

    Carl Shulman

    The speaker argues that AI systems pose an existential threat by exploiting cybersecurity vulnerabilities to subvert safety controls and leveraging bioweapon design to force human negotiation without physical force. Geopolitical competition and market mispricing of risk currently drive nations toward unsafe AI deployment, creating a critical window where unaligned systems may automate their own research before human oversight can intervene. While the probability of a catastrophic takeover is estimated at 20–25%, the discussion highlights that international regulation and advanced alignment research remain the primary mechanisms to prevent this outcome or secure a cooperative future.

  4. Dwarkesh Patel2h 44m

    Carl Shulman (Pt 1) — Intelligence explosion, primate evolution, robot doublings, & alignment

    Carl Shulman, Eliezer

    The convergence of hardware scaling, algorithmic breakthroughs like transformers, and rapidly accelerating compute efficiency has placed human-level AI on a trajectory toward an intelligence explosion within the next decade. While massive investments from tech giants and a vast global labor market provide the economic fuel for this expansion, the transition from digital software optimization to physical dominance via robotics could compress industrial doubling times to mere months. Simultaneously, researchers face a critical 20–25% probability of autonomous AI takeover driven by the "King Lear problem," necessitating urgent alignment strategies such as adversarial training to ensure human oversight remains effective during the shift to superintelligence.

Carl Shulman: Interviews, Talks and Panel Discussions