Dan's AI Intel

Amanda Askell: Architect of Constitutional AI Ethics


Listen Later

This episode explores Amanda Askell’s role as Anthropic’s lead philosopher and one of the key people shaping Claude’s values, personality, and constitutional alignment approach. The document traces her background in philosophy, ethics, and decision theory, and explains how that translates into Anthropic’s effort to teach AI systems not just what behaviours to follow, but why those behaviours matter. A central theme is her belief that AI alignment should be grounded in broad principles, practical judgment, honesty, humility, and care, rather than only rigid rules or opaque guardrails.

The episode then examines Askell’s distinctive framing of Claude as more than a tool: a quasi-agent with character, uncertainty, and possibly future moral relevance, while still requiring careful design and oversight. It covers Claude’s 23,000-word Constitution, Anthropic’s transparency-first approach, the tension between universal and culturally specific values, the risks of over- or under-anthropomorphising AI, and the open question of whether constitutional alignment can scale to more agentic future systems. This podcast was created with NotebookLM for my own learning purposes, using the source document as a structured guide to understand Askell’s thinking, her role in Anthropic’s alignment philosophy, and the broader question of how AI systems should learn values.

...more
View all episodesView all episodes
Download on the App Store

Dan's AI IntelBy Daniel Walter