newsfilter.io

Allan Dafoe

Showing 14 of 4 transcripts.

  1. 80,000 Hours1h 38m

    2025 Highlight-o-thon: Oops! All Bests

    Kyle Fish, Ian Dunt, Sam Bowman, Buck Shlegeris, Luisa, Rob, Helen Toner, Hugh White, Paul Scharre, Beth Barnes, Tyler Whitmer, Toby Ord, Andrew Snyder-Beattie, Eileen Yam, Will MacAskill, Neel Nanda, Tom Davidson, Marius Hobbhahn, Holden Karnofsky, Allan Dafoe, Ryan Greenblatt, Daniel Kokotajlo, Dean Ball

    This forum convened experts to debate the accelerating timeline of AGI by 2029 while critiquing US geopolitical strategies for abandoning global primacy in favor of a multipolar order. Participants examined critical risks including AI scheming, biological defense asymmetries, and the erosion of human context in warfare, contrasting them with corporate reforms at OpenAI and the rising costs of AI inference. The discourse further highlighted the widening perception gap between AI developers and the public, the potential of mechanistic interpretability as an "AI biology," and the structural necessity of aligning urban planning with community quality of life rather than NIMBYism.

  2. 80,000 Hours2h 36m

    Hacking, defending, surviving: 15 expert takes on information security in the age of AI

    Rob, Holden Karnofsky, Tantum Collins, Nick Joseph, Nova DasSarma, Sella Nevo, Kevin Esvelt, Lennart Heim, Zvi Mowshowitz, Bruce Schneier, Nita Farahany, Vitalik Buterin, Nathan Labenz, Allan Dafoe, Tom Davidson, Carl Shulman

    Experts highlight that physical infiltration vectors like USB drives and the immense strategic value of stolen AI model weights expose frontier systems to significant theft by both nation-states and casual actors. While formal verification and liability shifts are proposed to mitigate these risks, severe workforce shortages and the inability of current defenses to stop well-funded adversaries leave the confidentiality of critical AI assets highly vulnerable. Consequently, the consensus suggests that without mandatory security standards and a fundamental shift in security culture, the development of advanced AI remains compromised by the threat of unauthorized deployment and "weaponized" model variants.

  3. 80,000 Hours2h 49m

    Technological inevitability & human agency in the age of AGI | DeepMind's Allan Dafoe

    Allan Dafoe, Rob

    Defoe argues that macro-historical technological shifts are driven by structural competition rather than individual agency, a perspective shaping his work at Google DeepMind's Frontier Safety team to integrate safety governance directly into high-level decision-making. He advocates for "differential technological development" and "Cooperative AI" to address risks beyond simple alignment, while recent evaluations of Gemini reveal moderate risks in persuasion and self-reasoning that necessitate staged deployment frameworks. Beyond immediate safety protocols, the discussion highlights AI's potential to revolutionize sectors like healthcare and transportation, urging social scientists to join the effort in managing these emerging structural and geopolitical challenges.

  4. 80,000 Hours48 min

    #31 - Prof Dafoe on defusing the political & economic risks posed by existing AI capabilities

    Prof Dafoe, Allan Dafoe, Keiran Harris, Rob Wiblin

    The Future of Humanity Institute's AI Governance program unites a dozen researchers to analyze the technical, political, and institutional challenges posed by the transition to superhuman AI. Led by figures such as Alan Daffo, Demis Hassabis, and Vladimir Putin, the field addresses risks ranging from labor displacement and strategic instability to the potential for a dangerous "race to the finish" among nation-states. Despite the current fragmentation of the discipline, experts maintain a strategic focus on building global cooperation frameworks to harness advanced AI's vast potential while preventing catastrophic outcomes.