
Sign up to save your podcasts
Or


Today we’re joined by Jilei Hou, a VP of Engineering at Qualcomm Technologies. In our conversation with Jilei, we focus on the emergence of generative AI, and how they've worked towards providing these models for use on edge devices. We explore how the distribution of models on devices can help amortize large models' costs while improving reliability and performance and the challenges of running machine learning workloads on devices, including model size and inference latency. Finally, Jilei we explore how these emerging technologies fit into the existing AI Model Efficiency Toolkit (AIMET) framework.
The complete show notes for this episode can be found at twimlai.com/go/633
By Sam Charrington4.7
419419 ratings
Today we’re joined by Jilei Hou, a VP of Engineering at Qualcomm Technologies. In our conversation with Jilei, we focus on the emergence of generative AI, and how they've worked towards providing these models for use on edge devices. We explore how the distribution of models on devices can help amortize large models' costs while improving reliability and performance and the challenges of running machine learning workloads on devices, including model size and inference latency. Finally, Jilei we explore how these emerging technologies fit into the existing AI Model Efficiency Toolkit (AIMET) framework.
The complete show notes for this episode can be found at twimlai.com/go/633

480 Listeners

1,090 Listeners

170 Listeners

303 Listeners

334 Listeners

208 Listeners

201 Listeners

95 Listeners

512 Listeners

130 Listeners

227 Listeners

608 Listeners

25 Listeners

35 Listeners

40 Listeners