Audio versions of blogs and papers from BlueDot courses.
This post tries to explain a simplified version of Paul Christiano’s mechanism introduced here, (referred to there as ‘Learning the Prior’) and explain why a mechanism like this potentially addresses some of the safety problems with naïve approaches. First we’ll go through a simple example in a familiar domain, then explain the problems with the example. Then I’ll discuss the open questions for making Imitative Generalization actually work, and the connection with the Microscope AI idea. A more detailed explanation of exactly what the training objective is (with diagrams), and the correspondence with Bayesian inference, are in the appendix.
Source:
Narrated for AI Safety Fundamentals by Perrin Walker of TYPE III AUDIO.
---
A podcast by BlueDot Impact.
Fler avsnitt av BlueDot Narrated
Visa alla avsnitt av BlueDot NarratedBlueDot Narrated med BlueDot Impact finns tillgänglig på flera plattformar. Informationen på denna sida kommer från offentliga podd-flöden.
