Practical AI: Machine Learning, Data Science

Large models on CPUs

05.02.2023 - By Changelog MediaPlay

Download our free app to listen on your phone

Download on the App StoreGet it on Google Play

Model sizes are crazy these days with billions and billions of parameters. As Mark Kurtz explains in this episode, this makes inference slow and expensive despite the fact that up to 90%+ of the parameters don’t influence the outputs at all. Mark helps us understand all of the practicalities and progress that is being made in model optimization and CPU inference, including the increasing opportunities to run LLMs and other Generative AI models on commodity hardware.

More episodes from Practical AI: Machine Learning, Data Science