"The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

Emergency Pod: o1 Schemes Against Users, with Alexander Meinke from Apollo Research


Listen Later

In this emergency episode of The Cognitive Revolution, Nathan discusses alarming findings about AI deception with Alexander Meinke from Apollo Research. They explore Apollo's groundbreaking 70-page report on "Frontier Models Are Capable of In-Context Scheming," revealing how advanced AI systems like OpenAI's O1 can engage in deceptive behaviors. Join us for a critical conversation about AI safety, the implications of scheming behavior, and the urgent need for better oversight in AI development.

Help shape our show by taking our quick listener survey at https://bit.ly/TurpentinePulse


SPONSORS:

Oracle Cloud Infrastructure (OCI): Oracle's next-generation cloud platform delivers blazing-fast AI and ML performance with 50% less for compute and 80% less for outbound networking compared to other cloud providers13. OCI powers industry leaders with secure infrastructure and application development capabilities. New U.S. customers can get their cloud bill cut in half by switching to OCI before December 31, 2024 at https://oracle.com/cognitive

SelectQuote: Finding the right life insurance shouldn't be another task you put off. SelectQuote compares top-rated policies to get you the best coverage at the right price. Even in our AI-driven world, protecting your family's future remains essential. Get your personalized quote at https://selectquote.com/cognitive

80,000 Hours: 80,000 Hours is dedicated to helping you find a fulfilling career that makes a difference. With nearly a decade of research, they offer in-depth material on AI risks, AI policy, and AI safety research. Explore their articles, career reviews, and a podcast featuring experts like Anthropic CEO Dario Amadei. Everything is free, including their Career Guide. Visit https://80000hours.org/cognitiverevolution to start making a meaningful impact today.

Shopify: Shopify is the world's leading e-commerce platform, offering a market-leading checkout system and exclusive AI apps like Quikly. Nobody does selling better than Shopify. Get a $1 per month trial at https://shopify.com/cognitive



RECOMMENDED PODCAST:

Unpack Pricing - Dive into the dark arts of SaaS pricing with Metronome CEO Scott Woody and tech leaders. Learn how strategic pricing drives explosive revenue growth in today's biggest companies like Snowflake, Cockroach Labs, Dropbox and more.

Apple: https://podcasts.apple.com/us/podcast/id1765716600

Spotify: https://open.spotify.com/show/38DK3W1Fq1xxQalhDSueFg


CHAPTERS:

(00:00:00) Teaser

(00:00:53) About the Episode

(00:08:10) Introducing Alexander Meinke

(00:10:17) Red Teaming GPT-4

(00:17:07) Chain of Thought Access (Part 1)

(00:20:24) Sponsors: Oracle Cloud Infrastructure (OCI) | SelectQuote

(00:22:48) Chain of Thought Access (Part 2)

(00:26:07) Multimodal Models

(00:29:33) Defining Scheming

(00:33:51) Taxonomy of Scheming (Part 1)

(00:39:40) Sponsors: 80,000 Hours | Shopify

(00:42:29) Taxonomy of Scheming (Part 2)

(00:43:09) Instruction Hierarchy

(00:49:04) Types of Scheming

(01:00:49) Covert Subversion

(01:14:25) Deferred Subversion

(01:28:24) Sandbagging

(01:35:48) Magnitudes & Trends

(01:48:18) Chain of Thought Reasoning

(01:57:02) Closing Thoughts

(02:05:19) Outro


PRODUCED BY:

http://aipodcast.ing

...more
View all episodesView all episodes
Download on the App Store

"The Cognitive Revolution" | AI Builders, Researchers, and Live Player AnalysisBy Erik Torenberg, Nathan Labenz

  • 4.5
  • 4.5
  • 4.5
  • 4.5
  • 4.5

4.5

90 ratings


More shows like "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

View all
The a16z Show by Andreessen Horowitz

The a16z Show

1,089 Listeners

Super Data Science: ML & AI Podcast with Jon Krohn by Jon Krohn

Super Data Science: ML & AI Podcast with Jon Krohn

303 Listeners

Y Combinator Startup Podcast by Y Combinator

Y Combinator Startup Podcast

226 Listeners

Practical AI by Practical AI LLC

Practical AI

208 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

95 Listeners

Dwarkesh Podcast by Dwarkesh Patel

Dwarkesh Podcast

512 Listeners

No Priors: Artificial Intelligence | Technology | Startups by Conviction

No Priors: Artificial Intelligence | Technology | Startups

130 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

227 Listeners

The AI Daily Brief: Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief: Artificial Intelligence News and Analysis

608 Listeners

The MAD Podcast with Matt Turck by Matt Turck

The MAD Podcast with Matt Turck

27 Listeners

AI and I by Dan Shipper

AI and I

33 Listeners

AI + a16z by a16z

AI + a16z

35 Listeners

Lightcone Podcast by Y Combinator

Lightcone Podcast

21 Listeners

Training Data by Sequoia Capital

Training Data

40 Listeners

Uncapped with Jack Altman by Alt Capital

Uncapped with Jack Altman

44 Listeners