Paul Christiano — Preventing an AI takeover
Paul Christiano is the world’s leading AI safety researcher. My full episode with him is out! We discuss: - Does he regret inventing RLHF, and is alignment necessarily dual-use? - Why he has relatively modest timelines (40% by 2040, 15% by 2030), - What do we want post-AGI world to look like (do we…
Sources
- T2Paul Christiano — Preventing an AI takeoverPodcasts: Practical AI / TWIML / No Priors / Dwarkesh