
Sign up to save your podcasts
Or
In this episode of No Priors, hosts Sarah and Elad are joined by Jared Quincy Davis, former DeepMind researcher and the Founder and CEO of Foundry, a new AI cloud computing service provider. They discuss the research problems that led him to starting Foundry, the current state of GPU cloud utilization, and Foundry's approach to improving cloud economics for AI workloads. Jared also touches on his predictions for the GPU market and the thinking behind his recent paper on designing compound AI systems.
Sign up for new podcasts every week. Email feedback to [email protected]
Follow us on Twitter: @NoPriorsPod | @Saranormous | @EladGil | @jaredq_
Show Notes:
(00:00) Introduction
(02:42) Foundry background
(03:57) GPU utilization for large models
(07:29) Systems to run a large model
(09:54) Historical value proposition of the cloud
(14:45) Sharing cloud compute to increase efficiency
(19:17) Foundry’s new releases
(23:54) The current state of GPU capacity
(29:50) GPU market dynamics
(36:28) Compound systems design
(40:27) Improving open-ended tasks
4.6
9393 ratings
In this episode of No Priors, hosts Sarah and Elad are joined by Jared Quincy Davis, former DeepMind researcher and the Founder and CEO of Foundry, a new AI cloud computing service provider. They discuss the research problems that led him to starting Foundry, the current state of GPU cloud utilization, and Foundry's approach to improving cloud economics for AI workloads. Jared also touches on his predictions for the GPU market and the thinking behind his recent paper on designing compound AI systems.
Sign up for new podcasts every week. Email feedback to [email protected]
Follow us on Twitter: @NoPriorsPod | @Saranormous | @EladGil | @jaredq_
Show Notes:
(00:00) Introduction
(02:42) Foundry background
(03:57) GPU utilization for large models
(07:29) Systems to run a large model
(09:54) Historical value proposition of the cloud
(14:45) Sharing cloud compute to increase efficiency
(19:17) Foundry’s new releases
(23:54) The current state of GPU capacity
(29:50) GPU market dynamics
(36:28) Compound systems design
(40:27) Improving open-ended tasks
1,281 Listeners
1,008 Listeners
525 Listeners
121 Listeners
439 Listeners
2,329 Listeners
214 Listeners
196 Listeners
8,385 Listeners
315 Listeners
189 Listeners
70 Listeners
397 Listeners
106 Listeners
419 Listeners