
Sign up to save your podcasts
Or


In their first episode of the New Year, Katherine Forrest and Scott Caravello unpack OpenAI’s release of GPT-5.2, covering its performance on benchmarks like Humanity’s Last Exam and on OpenAI’s own GDPval—an evaluation designed to test the model's ability to match or surpass professionals' performance on real world tasks. Our hosts also examine the model’s sharp drop in hallucinations and break down OpenAI’s discussion of the model’s resistance to prompt injections and how it stacks up under the company’s safety framework.
##
Learn More About Paul, Weiss’s Artificial Intelligence practice:
By Paul, Weiss4.8
2323 ratings
In their first episode of the New Year, Katherine Forrest and Scott Caravello unpack OpenAI’s release of GPT-5.2, covering its performance on benchmarks like Humanity’s Last Exam and on OpenAI’s own GDPval—an evaluation designed to test the model's ability to match or surpass professionals' performance on real world tasks. Our hosts also examine the model’s sharp drop in hallucinations and break down OpenAI’s discussion of the model’s resistance to prompt injections and how it stacks up under the company’s safety framework.
##
Learn More About Paul, Weiss’s Artificial Intelligence practice:

91,009 Listeners

6,849 Listeners

30,694 Listeners

2,434 Listeners

9,772 Listeners

112,956 Listeners

369,807 Listeners

10,335 Listeners

5,839 Listeners

10,227 Listeners

5,547 Listeners

16,362 Listeners

659 Listeners

1,471 Listeners

13,434 Listeners