Efficiently Serving LLMs
Join our new short course, Efficiently Serving Large Language Models, to build a ground-up understanding of how to serve LLM applications from Travis Addair, CTO at Predibase. Whether you’re ready to launch your own application or just getting started building it, the topics you’ll explore in this course will deepen your foundational knowledge of how LLMs work, and help you better understand the performance trade-offs you must consider when building LLM applications that will serve large numbers of users.
You’ll walk through the most important optimizations that allow LLM vendors to efficient…
Watch on Coursera ↗
(saves to browser)
DeepCamp AI