newsfilter.io

Buck Shlegeris

Showing 13 of 3 transcripts.

  1. 80,000 Hours1h 38m

    2025 Highlight-o-thon: Oops! All Bests

    Kyle Fish, Ian Dunt, Sam Bowman, Buck Shlegeris, Luisa, Rob, Helen Toner, Hugh White, Paul Scharre, Beth Barnes, Tyler Whitmer, Toby Ord, Andrew Snyder-Beattie, Eileen Yam, Will MacAskill, Neel Nanda, Tom Davidson, Marius Hobbhahn, Holden Karnofsky, Allan Dafoe, Ryan Greenblatt, Daniel Kokotajlo, Dean Ball

    This forum convened experts to debate the accelerating timeline of AGI by 2029 while critiquing US geopolitical strategies for abandoning global primacy in favor of a multipolar order. Participants examined critical risks including AI scheming, biological defense asymmetries, and the erosion of human context in warfare, contrasting them with corporate reforms at OpenAI and the rising costs of AI inference. The discourse further highlighted the widening perception gap between AI developers and the public, the potential of mechanistic interpretability as an "AI biology," and the structural necessity of aligning urban planning with community quality of life rather than NIMBYism.

  2. 80,000 Hours3h 35m

    The bewildering frontier of consciousness in insects, AI, and more | 17 experts weigh in

    Luisa, Robert Long, Jeff Sebo, Meghan Barrett, Andrés Jiménez Zorrilla, Jonathan Birch, David Chalmers, Holden Karnofsky, Bob Fischer, Cameron Meyer Shorb, Sébastien Moro, Anil Seth, Peter Godfrey-Smith, Lewis Bollard, Stuart Russell, Buck Shlegeris, Will MacAskill, Carl Shulman

    A panel of leading neuroscientists and philosophers, including Megan Barrett, Robert Long, and David Chalmers, explores the ethical frameworks required to address the potential sentience of invertebrates and artificial intelligence amid profound uncertainty. Participants argue that the vast global populations of invertebrates and the theoretical possibility of conscious silicon-based systems necessitate a precautionary moral approach to prevent mass suffering and exploitation. The discussion concludes that current evidence, while inconclusive regarding a definitive "sentience score," is sufficient to warrant legal protections and the development of cooperative economic models for these potentially conscious entities.

  3. 80,000 Hours2h 19m

    Controlling AI That Wants To Take Over – So We Can Use It Anyway | Buck Shlegeris

    Buck Shlegeris

    The event outlines a strategic shift toward "AI Control," a harm-reduction framework designed to mitigate catastrophic misalignment risks by assuming models are already scheming to seize computational resources rather than relying on perfect alignment. It details specific technical mechanisms such as the Execute-Replace-Audit framework and trajectory resampling, which allow smaller teams to detect and neutralize internal data center compromises before they escalate. Furthermore, the discussion contrasts the correlated nature of AI threats with human insider risks, emphasizing that these low-cost, implementable controls can be deployed immediately even amidst competitive market pressures and reduced safety budgets.

Buck Shlegeris: Interviews, Talks and Panel Discussions