"The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

GELU, MMLU, & X-Risk Defense in Depth, with the Great Dan Hendrycks


Listen Later

Join Nathan for an expansive conversation with Dan Hendrycks, Executive Director of the Center for AI Safety and Advisor to Elon Musk's XAI. In this episode of The Cognitive Revolution, we explore Dan's groundbreaking work in AI safety and alignment, from his early contributions to activation functions to his recent projects on AI robustness and governance. Discover insights on representation engineering, circuit breakers, and tamper-resistant training, as well as Dan's perspectives on AI's impact on society and the future of intelligence. Don't miss this in-depth discussion with one of the most influential figures in AI research and safety.

Check out some of Dan's research papers:

MMLU: https://arxiv.org/abs/2009.03300

GELU: https://arxiv.org/abs/1606.08415

Machiavelli Benchmark: https://arxiv.org/abs/2304.03279

Circuit Breakers: https://arxiv.org/abs/2406.04313

Tamper Resistant Safeguards: https://arxiv.org/abs/2408.00761

Statement on AI Risk: https://www.safe.ai/work/statement-on-ai-risk


Apply to join over 400 Founders and Execs in the Turpentine Network: https://www.turpentinenetwork.co/


SPONSORS:

Shopify: Shopify is the world's leading e-commerce platform, offering a market-leading checkout system and exclusive AI apps like Quikly. Nobody does selling better than Shopify. Get a $1 per month trial at https://shopify.com/cognitive.

LMNT: LMNT is a zero-sugar electrolyte drink mix that's redefining hydration and performance. Ideal for those who fast or anyone looking to optimize their electrolyte intake. Support the show and get a free sample pack with any purchase at https://drinklmnt.com/tcr.

Notion: Notion offers powerful workflow and automation templates, perfect for streamlining processes and laying the groundwork for AI-driven automation. With Notion AI, you can search across thousands of documents from various platforms, generating highly relevant analysis and content tailored just for you - try it for free at https://notion.com/cognitiverevolution

Oracle: Oracle Cloud Infrastructure (OCI) is a single platform for your infrastructure, database, application development, and AI needs. OCI has four to eight times the bandwidth of other clouds; offers one consistent price, and nobody does data better than Oracle. If you want to do more and spend less, take a free test drive of OCI at https://oracle.com/cognitive


CHAPTERS:

(00:00:00) Teaser

(00:00:48) About the Show

(00:02:17) About the Episode

(00:05:41) Intro

(00:07:19) GELU Activation Function

(00:10:48) Signal Filtering

(00:12:46) Scaling Maximalism

(00:18:35) Sponsors: Shopify | LMNT

(00:22:03) New Architectures

(00:25:41) AI as Complex System

(00:32:35) The Machiavelli Benchmark

(00:34:10) Sponsors: Notion | Oracle

(00:37:20) Understanding MMLU Scores

(00:45:23) Reasoning in Language Models

(00:49:18) Multimodal Reasoning

(00:54:53) World Modeling and Sora

(00:57:07) Arc Benchmark and Hypothesis

(01:01:06) Humanity's Last Exam

(01:08:46) Benchmarks and AI Ethics

(01:13:28) Robustness and Jailbreaking

(01:18:36) Representation Engineering

(01:30:08) Convergence of Approaches

(01:34:18) Circuit Breakers

(01:37:52) Tamper Resistance

(01:49:10) Interpretability vs. Robustness

(01:53:53) Open Source and AI Safety

(01:58:16) Computational Irreducibility

(02:06:28) Neglected Approaches

(02:12:47) Truth Maxing and XAI

(02:19:59) AI-Powered Forecasting

(02:24:53) Chip Bans and Geopolitics

(02:33:30) Working at CAIS

(02:35:03) Extinction Risk Statement

(02:37:24) Outro

View all episodesView all episodes
Download on the App Store

"The Cognitive Revolution" | AI Builders, Researchers, and Live Player AnalysisBy Erik Torenberg, Nathan Labenz

  • 4.6
  • 4.6
  • 4.6
  • 4.6
  • 4.6

4.6

81 ratings


More shows like "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

View all
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence) by Sam Charrington

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

429 Listeners

Practical AI by Practical AI LLC

Practical AI

196 Listeners

Last Week in AI by Skynet Today

Last Week in AI

281 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

90 Listeners

Dwarkesh Podcast by Dwarkesh Patel

Dwarkesh Podcast

325 Listeners

"Moment of Zen" by Erik Torenberg, Dan Romero, Antonio Garcia Martinez

"Moment of Zen"

89 Listeners

No Priors: Artificial Intelligence | Technology | Startups by Conviction

No Priors: Artificial Intelligence | Technology | Startups

103 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

192 Listeners

Latent Space: The AI Engineer Podcast by swyx + Alessio

Latent Space: The AI Engineer Podcast

64 Listeners

"Upstream" with Erik Torenberg by Erik Torenberg

"Upstream" with Erik Torenberg

65 Listeners

The AI Daily Brief (Formerly The AI Breakdown): Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief (Formerly The AI Breakdown): Artificial Intelligence News and Analysis

421 Listeners

"In The Arena" by Turpentine

"In The Arena"

16 Listeners

"The Hill & Valley" by Jacob Helberg, Delian Asparouhov, Christian Garrett

"The Hill & Valley"

11 Listeners

"Econ 102" with Noah Smith and Erik Torenberg by Turpentine

"Econ 102" with Noah Smith and Erik Torenberg

138 Listeners

"Turpentine VC" | Venture Capital and Investing by Erik Torenberg

"Turpentine VC" | Venture Capital and Investing

20 Listeners

"Tech Finance" with Sasha Orloff: B2B Fintech | AI | Finance Tech by Puzzle, Turpentine

"Tech Finance" with Sasha Orloff: B2B Fintech | AI | Finance Tech

43 Listeners

"1 to 1000" | Scaling Startups with CEOs by Turpentine

"1 to 1000" | Scaling Startups with CEOs

2 Listeners

"The Riff" with Byrne Hobart and Erik Torenberg by Byrne Hobart, Erik Torenberg

"The Riff" with Byrne Hobart and Erik Torenberg

21 Listeners

"Live Players" with Samo Burja and Erik Torenberg by Turpentine

"Live Players" with Samo Burja and Erik Torenberg

39 Listeners

AI and I by Dan Shipper

AI and I

29 Listeners

History 102 with WhatifAltHist's Rudyard Lynch and Austin Padgett by Turpentine

History 102 with WhatifAltHist's Rudyard Lynch and Austin Padgett

99 Listeners

Emergent Behavior by Turpentine

Emergent Behavior

7 Listeners

"WhatifAlthist" | World History, Philosophy, Culture by Rudyard Lynch

"WhatifAlthist" | World History, Philosophy, Culture

60 Listeners

"Autopilot" with Will Summerlin by Will Summerlin | Turpentine

"Autopilot" with Will Summerlin

4 Listeners

"Company Breakdowns" by Turpentine

"Company Breakdowns"

4 Listeners

Training Data by Sequoia Capital

Training Data

30 Listeners

Complex Systems with Patrick McKenzie (patio11) by Patrick McKenzie

Complex Systems with Patrick McKenzie (patio11)

113 Listeners

"Second Opinion" with Christina Farr, Ash Zenooz MD & Luba Greenwood JD by Christina Farr, Luba Greenwood, Ash Zenooz

"Second Opinion" with Christina Farr, Ash Zenooz MD & Luba Greenwood JD

15 Listeners

"1 to 100" | Hypergrowth Startups Worth Joining by Turpentine, Why You Should Join

"1 to 100" | Hypergrowth Startups Worth Joining

0 Listeners

"This Won't Last" with Keith Rabois, Kevin Ryan, Logan Bartlett, and Zach Weinberg by Turpentine, Keith Rabois, Logan Bartlett, Zach Weinberg, Kevin Ryan

"This Won't Last" with Keith Rabois, Kevin Ryan, Logan Bartlett, and Zach Weinberg

14 Listeners

"Modern Relationships" by Erik Torenberg, Turpentine

"Modern Relationships"

7 Listeners