Class Central is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

CNCF [Cloud Native Computing Foundation]

How Fast Can Your Model Composition Run in Serverless Inference?

CNCF [Cloud Native Computing Foundation] via YouTube

Overview

Google, IBM & Meta Certificates – 40% Off
One Coursera Plus subscription covers most Professional Certificates on Coursera.
Unlock All Certificates
The session presents a RAG application that combines an LLM, an embedding model, and OCR for inference in serverless Kubernetes. It explains how BentoML, Dragonfly’s peer-to-peer network, and other open-source technologies support model packaging, distribution, and rapid deployment.

Syllabus

How Fast Can Your Model Composition Run in Serverless Inference? - Fog Dong, BentoML & Wenbo Qi

Taught by

CNCF [Cloud Native Computing Foundation]

Reviews

Start your review of How Fast Can Your Model Composition Run in Serverless Inference?

Never Stop Learning.

Get personalized course recommendations, track subjects and courses with reminders, and more.

Someone learning on their laptop while sitting on the floor.