Next upHack for Humanity: San Francisco (powered by Google Gemini)
55:03

A Whirlwind Tour of ML Model Serving Strategies (Including LLMs)

There are many recipes to serve machine learning models to end users today, and even though new ways keep popping up as time passes, some questions remain: How do we pick the appropriate serving recipe from the menu we have available, and how can we execute it as fast and efficiently as possible? In this talk, we’re going to go through a whirlwind tour of the different machine learning deployment strategies available today for both traditional ML systems and Large Language Models, and we

Dmytro Spodarets
Dmytro Spodarets
Jan 29, 2024
Summary

A whirlwind tour of machine learning model serving strategies for both traditional ML systems and Large Language Models, covering how to choose the right deployment recipe and execute it as fast and efficiently as possible.