Stuart Russell
Showing 1–6 of 6 transcripts.
- The Diary Of A CEO2h 4m
An AI Expert Warning: 6 People Are (Quietly) Deciding Humanity’s Future!
In October, over 850 experts including Richard Branson and Geoffrey Hinton signed a statement calling for a pause on AI development to prevent extinction-level risks comparable to nuclear war. Computer scientist Stuart Russell argues that the current "race" driven by geopolitical competition and the "event horizon" of self-improving systems creates a control paradox where humans cannot guarantee the safety of superintelligence. Russell proposes regulating AI like nuclear power under a "Human Compatible" framework that defers to human values, emphasizing that without immediate safety protocols, no viable future for artificial intelligence exists.
- 80,000 Hours3h 35m
The bewildering frontier of consciousness in insects, AI, and more | 17 experts weigh in
Luisa, Robert Long, Jeff Sebo, Meghan Barrett, Andrés Jiménez Zorrilla, Jonathan Birch, David Chalmers, Holden Karnofsky, Bob Fischer, Cameron Meyer Shorb, Sébastien Moro, Anil Seth, Peter Godfrey-Smith, Lewis Bollard, Stuart Russell, Buck Shlegeris, Will MacAskill, Carl Shulman
A panel of leading neuroscientists and philosophers, including Megan Barrett, Robert Long, and David Chalmers, explores the ethical frameworks required to address the potential sentience of invertebrates and artificial intelligence amid profound uncertainty. Participants argue that the vast global populations of invertebrates and the theoretical possibility of conscious silicon-based systems necessitate a precautionary moral approach to prevent mass suffering and exploitation. The discussion concludes that current evidence, while inconclusive regarding a definitive "sentience score," is sufficient to warrant legal protections and the development of cooperative economic models for these potentially conscious entities.
- 80,000 Hours2h 13m
Flaws that make AI architecture unsafe & how to fix them | Stuart Russell (2020)
Stuart Russell argues that the standard AI model of optimizing fixed objectives is fundamentally flawed, proposing instead a "Human Compatible" framework where systems remain uncertain about human preferences to ensure safety. He warns that without this paradigm shift and new regulatory standards, humanity faces existential threats from autonomous weapons, surveillance, and the gradual erosion of human autonomy. Russell emphasizes that while superintelligent AI could dramatically increase global wealth, achieving this safely requires immediate technical breakthroughs in alignment and strict government oversight of algorithmic deployment.
- Lex Fridman12 min
Stuart Russell: The Control Problem of Super-Intelligent AI | AI Podcast Clips
Experts argue that the critical risk of artificial intelligence lies not in general intelligence but in "super powerful AI that is not aligned with human values," where systems treat assigned objectives as absolute truths and optimize them destructively, much like the myth of King Midas or historical regimes such as Nazi Germany. To mitigate this control problem, the proposed solution involves engineering "machine humility" by replacing standard goal-based planning with game-theoretic frameworks that allow AI systems to remain uncertain about their ultimate objectives and interpret human feedback as new data for co-evolving goals. This shift aims to ensure that both corporations and governments cease acting as rigid algorithmic machines optimizing for fixed metrics like quarterly profit or personal power, thereby aligning advanced automation with genuine human well-being.
- Lex Fridman1h 26m
Stuart Russell: Long-Term Future of Artificial Intelligence | Lex Fridman Podcast #9
UC Berkeley professor Stuart Russell traces the evolution of AI from his early 1970s chess programs to modern meta-reasoning systems like AlphaGo, highlighting how these technologies now solve complex decision problems through selective resource allocation rather than exhaustive search. Beyond technical achievements, Russell warns of critical existential risks including the "Gorilla Problem" of uncontrollable superintelligence and the "Wally Problem" of human skill atrophy, arguing that current regulatory frameworks are insufficient to manage civilization-scale impacts. To address these challenges, he advocates for a fundamental shift toward "provably beneficial machines" that maintain uncertainty about human objectives, ensuring systems remain deferential to human feedback and preserve human autonomy rather than optimizing rigid goals.
- Milken Institute58 min
Artificial Intelligence: Friend or Foe?
Alexandra Suich, Guruduth Banavar, Michael Ferro, Stuart Russell, David M. Siegel, Shivon Zilis
A panel of experts frames the current AI revolution as being in its early stages, emphasizing the shift from sci-fi fantasies of sentient robots to practical "narrow AI" that augments human intelligence across diverse sectors like healthcare, agriculture, and transportation. While highlighting transformative benefits such as expanded medical diagnostics and automated labor, the discussion underscores immediate risks including white-collar job displacement, the weaponization of autonomous systems, and the complex ethical challenge of encoding human values into machine decision-making. Looking toward the next decade, the consensus predicts universal cognitive assistants and significant advancements in personalized medicine and road safety, contingent on establishing new legal and safety frameworks to manage the societal impacts of this widespread automation.