The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

Beyond Accuracy: Behavioral Testing of NLP Models with Sameer Singh - #406


Listen Later

Today we’re joined by Sameer Singh, an assistant professor in the department of computer science at UC Irvine. 

Sameer’s work centers on large-scale and interpretable machine learning applied to information extraction and natural language processing. We caught up with Sameer right after he was awarded the best paper award at ACL 2020 for his work on Beyond Accuracy: Behavioral Testing of NLP Models with CheckList.

In our conversation, we explore CheckLists, the task-agnostic methodology for testing NLP models introduced in the paper. We also discuss how well we understand the cause of pitfalls or failure modes in deep learning models, Sameer’s thoughts on embodied AI, and his work on the now famous LIME paper, which he co-authored alongside Carlos Guestrin. 

The complete show notes for this episode can be found at twimlai.com/go/406.

...more
View all episodesView all episodes
Download on the App Store

The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)By Sam Charrington

  • 4.7
  • 4.7
  • 4.7
  • 4.7
  • 4.7

4.7

416 ratings


More shows like The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)

View all
a16z Podcast by Andreessen Horowitz

a16z Podcast

1,061 Listeners

Data Skeptic by Kyle Polich

Data Skeptic

477 Listeners

Super Data Science: ML & AI Podcast with Jon Krohn by Jon Krohn

Super Data Science: ML & AI Podcast with Jon Krohn

294 Listeners

NVIDIA AI Podcast by NVIDIA

NVIDIA AI Podcast

340 Listeners

AI Today Podcast by AI & Data Today

AI Today Podcast

151 Listeners

Practical AI by Practical AI LLC

Practical AI

189 Listeners

Last Week in AI by Skynet Today

Last Week in AI

297 Listeners

Machine Learning Street Talk (MLST) by Machine Learning Street Talk (MLST)

Machine Learning Street Talk (MLST)

91 Listeners

Dwarkesh Podcast by Dwarkesh Patel

Dwarkesh Podcast

424 Listeners

No Priors: Artificial Intelligence | Technology | Startups by Conviction

No Priors: Artificial Intelligence | Technology | Startups

126 Listeners

This Day in AI Podcast by Michael Sharkey, Chris Sharkey

This Day in AI Podcast

201 Listeners

Latent Space: The AI Engineer Podcast by swyx + Alessio

Latent Space: The AI Engineer Podcast

69 Listeners

The AI Daily Brief (Formerly The AI Breakdown): Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief (Formerly The AI Breakdown): Artificial Intelligence News and Analysis

508 Listeners

AI + a16z by a16z

AI + a16z

32 Listeners

Training Data by Sequoia Capital

Training Data

44 Listeners