June 24, 2025

Greg Kamradt: Benchmarking Intelligence | ARC Prize

48 minutes

What makes a good AI benchmark? Greg Kamradt joins Demetrios to break it down—from human-easy, AI-hard puzzles to wild new games that test how fast models can truly learn. They talk about hidden datasets, compute tradeoffs, and why benchmarks might be our best bet for tracking progress toward AGI. It’s nerdy, strategic, and surprisingly philosophical.

// Bio

Greg has mentored thousands of developers and founders, empowering them to build AI-centric applications. By crafting tutorial-based content, Greg aims to guide everyone from seasoned builders to ambitious indie hackers. Greg partners with companies during their product launches, feature enhancements, and funding rounds. His objective is to cultivate not just awareness, but also a practical understanding of how to optimally utilize a company's tools. He previously led Growth @ Salesforce for Sales & Service Clouds in addition to being early on at Digits, a FinTech Series-C company.

// Related Links

Website: https://gregkamradt.com/

YouTube channel: https://www.youtube.com/@DataIndependent

~~~~~~~~ ✌️Connect With Us ✌️ ~~~~~~~

Catch all episodes, blogs, newsletters, and more: https://go.mlops.community/TYExplore

MLOps Swag/Merch: [https://shop.mlops.community/]

Connect with Demetrios on LinkedIn: /dpbrinkm

Connect with Greg on LinkedIn: /gregkamradt/

Timestamps:

[00:00] Human-Easy, AI-Hard

[05:25] When the Model Shocks Everyone

[06:39] “Let’s Circle Back on That Benchmark…”

[09:50] Want Better AI? Pay the Compute Bill

[14:10] Can We Define Intelligence by How Fast You Learn?

[16:42] Still Waiting on That Algorithmic Breakthrough

[20:00] LangChain Was Just the Beginning

[24:23] Start With Humans, End With AGI

[29:01] What If Reality’s Just... What It Seems?

[32:21] AI Needs Fewer Vibes, More Predictions

[36:02] Defining Intelligence (No Pressure)

[36:41] AI Building AI? Yep, We're Going There

[40:13] Open Source vs. Prize Money Drama

[43:05] Architecting the ARC Challenge

[46:38] Agent 57 and the Atari Gauntlet

...more

View all episodes

By Demetrios

4.6

2323 ratings

June 24, 2025

Greg Kamradt: Benchmarking Intelligence | ARC Prize

48 minutes

// Bio

// Related Links

Website: https://gregkamradt.com/

YouTube channel: https://www.youtube.com/@DataIndependent

~~~~~~~~ ✌️Connect With Us ✌️ ~~~~~~~

Catch all episodes, blogs, newsletters, and more: https://go.mlops.community/TYExplore

MLOps Swag/Merch: [https://shop.mlops.community/]

Connect with Demetrios on LinkedIn: /dpbrinkm

Connect with Greg on LinkedIn: /gregkamradt/

Timestamps:

[00:00] Human-Easy, AI-Hard

[05:25] When the Model Shocks Everyone

[06:39] “Let’s Circle Back on That Benchmark…”

[09:50] Want Better AI? Pay the Compute Bill

[14:10] Can We Define Intelligence by How Fast You Learn?

[16:42] Still Waiting on That Algorithmic Breakthrough

[20:00] LangChain Was Just the Beginning

[24:23] Start With Humans, End With AGI

[29:01] What If Reality’s Just... What It Seems?

[32:21] AI Needs Fewer Vibes, More Predictions

[36:02] Defining Intelligence (No Pressure)

[36:41] AI Building AI? Yep, We're Going There

[40:13] Open Source vs. Prize Money Drama

[43:05] Architecting the ARC Challenge

[46:38] Agent 57 and the Atari Gauntlet

...more

More shows like MLOps.community

View all

This Week in Startups

1,290 Listeners

The Changelog: Software Development, Open Source

288 Listeners

The a16z Show

1,096 Listeners

Software Engineering Daily

624 Listeners

Talk Python To Me

583 Listeners

Super Data Science: ML & AI Podcast with Jon Krohn

301 Listeners

NVIDIA AI Podcast

344 Listeners

Practical AI

213 Listeners

Dwarkesh Podcast

561 Listeners

Big Technology Podcast

507 Listeners

No Priors: Artificial Intelligence | Technology | Startups

145 Listeners

Latent Space: The AI Engineer Podcast

100 Listeners

This Day in AI Podcast

227 Listeners

The AI Daily Brief: Artificial Intelligence News and Analysis

693 Listeners

AI + a16z

32 Listeners

Share Greg Kamradt: Benchmarking Intelligence | ARC Prize

Sign up to save your podcasts

Greg Kamradt: Benchmarking Intelligence | ARC Prize

Greg Kamradt: Benchmarking Intelligence | ARC Prize

More shows like MLOps.community

This Week in Startups

The Changelog: Software Development, Open Source

The a16z Show

Software Engineering Daily

Talk Python To Me

Super Data Science: ML & AI Podcast with Jon Krohn

NVIDIA AI Podcast

Practical AI

Dwarkesh Podcast

Big Technology Podcast

No Priors: Artificial Intelligence | Technology | Startups

Latent Space: The AI Engineer Podcast

This Day in AI Podcast

The AI Daily Brief: Artificial Intelligence News and Analysis

AI + a16z