Center for AI Policy Podcast

#10: Stephen Casper on Technical and Sociotechnical AI Safety Research


Listen Later

Stephen Casper, a computer science PhD student at MIT, joined the podcast to discuss AI interpretability, red-teaming and robustness, evaluations and audits, reinforcement learning from human feedback, Goodhart’s law, and more.

Our music is by Micah Rubin (Producer) and John Lisi (Composer).

For a transcript and relevant links, visit the Center for AI Policy Podcast Substack.

...more
View all episodesView all episodes
Download on the App Store

Center for AI Policy PodcastBy Center for AI Policy