Sveriges mest populära poddar

Dan Hendrycks on Catastrophic AI Risks

2 tim 7 min•3 november 2023

Dan Hendrycks joins the podcast again to discuss X.ai, how AI risk thinking has evolved, malicious use of AI, AI race dynamics between companies and between militaries, making AI organizations safer, and how representation engineering could help us understand AI traits like deception. You can learn more about Dan's work at https://www.safe.ai Timestamps: 00:00 X.ai - Elon Musk's new AI venture 02:41 How AI risk thinking has evolved 12:58 AI bioengeneering 19:16 AI agents 24:55 Preventing autocracy 34:11 AI race - corporations and militaries 48:04 Bulletproofing AI organizations 1:07:51 Open-source models 1:15:35 Dan's textbook on AI safety 1:22:58 Rogue AI 1:28:09 LLMs and value specification 1:33:14 AI goal drift 1:41:10 Power-seeking AI 1:52:07 AI deception 1:57:53 Representation engineering

Fler avsnitt av Future of Life Institute Podcast

How AI Companions Trap Users Through Addictive Design (with Claire Boine)

12 juni•1 tim 8 min

Why AI Chatbots Are a Rival to the Family (with Michael Toscano)

26 maj•1 tim 14 min

Why We Should Build AI Tools, Not AI Replacements (with Anthony Aguirre)

11 maj•1 tim 36 min

How to Govern AI When You Can't Predict the Future (with Charlie Bullock)

7 maj•1 tim 7 min

Why AI Is Not a Normal Technology (with Peter Wildeford)

29 apr.•1 tim 24 min

Why AI Evaluation Science Can't Keep Up (with Carina Prunkl)

17 apr.•54 min

Defense in Depth: Layered Strategies Against AI Risk (with Li-Lian Ang)

2 apr.•56 min

What AI Companies Get Wrong About Curing Cancer (with Emilia Javorsky)

20 mars•1 tim 12 min

AI vs Cancer - How AI Can, and Can't, Cure Cancer (by Emilia Javorsky)

16 mars•2 tim 43 min

How AI Hacks Your Brain's Attachment System (with Zak Stein)

5 mars•1 tim 45 min

Future of Life Institute Podcast med Future of Life Institute finns tillgänglig på flera plattformar. Informationen på denna sida kommer från offentliga podd-flöden.