Podcast Archives - Software Engineering Daily

Optimizing Agent Behavior in Production with Gideon Mendels


Listen Later

LLM -powered systems continue to move steadily into production, but this process is presenting teams with challenges that traditional software practices don’t commonly encounter. Models and agents are non-deterministic systems, which makes it difficult to test changes, reason about failures, and confidently ship updates. This has created the need for new evaluation tooling designed specifically around the properties of LLMs.

Comet is a platform with Roots and MLOps, to the rapidly evolving world of agent-based systems by treating prompts, tools, and workflows as optimizable components that can be evaluated and improved over time.

Gideon Mendels is the co -founder and CEO of Comet. He previously worked at Google on hate speech and deception detection, and he founded GroupWise, which trained and deployed NLP models processing billions of chats. In this episode, Gideon joins Kevin Ball to discuss how agent development sits between software engineering and ML, why eVals are the missing foundation for most AI teams, prompt optimization as a search problem, and the future for continuously improving agents in production.

Full Disclosure: This episode is sponsored by Comet.

Kevin Ball or KBall, is the vice president of engineering at Mento and an independent coach for engineers and engineering leaders. He co-founded and served as CTO for two companies, founded the San Diego JavaScript meetup, and organizes the AI inaction discussion group through Latent Space.

 

 

Please click here to see the transcript of this episode.

Sponsorship inquiries: [email protected]

The post Optimizing Agent Behavior in Production with Gideon Mendels appeared first on Software Engineering Daily.

...more
View all episodesView all episodes
Download on the App Store

Podcast Archives - Software Engineering DailyBy Podcast Archives - Software Engineering Daily

  • 4
  • 4
  • 4
  • 4
  • 4

4

4 ratings


More shows like Podcast Archives - Software Engineering Daily

View all
The Eastern Border by Kristaps Andrejsons

The Eastern Border

824 Listeners

Software Engineering Daily by Software Engineering Daily

Software Engineering Daily

623 Listeners

The Daily by The New York Times

The Daily

112,982 Listeners

Kubernetes Podcast from Google by Abdel Sghiouar, Kaslin Fields

Kubernetes Podcast from Google

180 Listeners

Post Reports by The Washington Post

Post Reports

5,449 Listeners

AWS Podcast by Amazon Web Services

AWS Podcast

209 Listeners

The Stack Overflow Podcast by The Stack Overflow Podcast

The Stack Overflow Podcast

64 Listeners

The 7 by The Washington Post

The 7

1,248 Listeners

The AI Daily Brief: Artificial Intelligence News and Analysis by Nathaniel Whittemore

The AI Daily Brief: Artificial Intelligence News and Analysis

659 Listeners

Rust in Production by Matthias Endler

Rust in Production

25 Listeners