Gå offline med appen Player FM !
Stable Diffusion and LLMs at the Edge with Jilei Hou - #633
Manage episode 365858086 series 2355587
Today we’re joined by Jilei Hou, a VP of Engineering at Qualcomm Technologies. In our conversation with Jilei, we focus on the emergence of generative AI, and how they've worked towards providing these models for use on edge devices. We explore how the distribution of models on devices can help amortize large models' costs while improving reliability and performance and the challenges of running machine learning workloads on devices, including model size and inference latency. Finally, Jilei we explore how these emerging technologies fit into the existing AI Model Efficiency Toolkit (AIMET) framework.
The complete show notes for this episode can be found at twimlai.com/go/633
779 episoder
Stable Diffusion and LLMs at the Edge with Jilei Hou - #633
The TWIML AI Podcast (formerly This Week in Machine Learning & Artificial Intelligence)
Manage episode 365858086 series 2355587
Today we’re joined by Jilei Hou, a VP of Engineering at Qualcomm Technologies. In our conversation with Jilei, we focus on the emergence of generative AI, and how they've worked towards providing these models for use on edge devices. We explore how the distribution of models on devices can help amortize large models' costs while improving reliability and performance and the challenges of running machine learning workloads on devices, including model size and inference latency. Finally, Jilei we explore how these emerging technologies fit into the existing AI Model Efficiency Toolkit (AIMET) framework.
The complete show notes for this episode can be found at twimlai.com/go/633
779 episoder
所有剧集
×Velkommen til Player FM!
Player FM is scanning the web for high-quality podcasts for you to enjoy right now. It's the best podcast app and works on Android, iPhone, and the web. Signup to sync subscriptions across devices.