Class Central is learner-supported. When you buy through links on our site, we may earn an affiliate commission.

YouTube

Hugging Face LLMs with SageMaker - RAG with Pinecone

James Briggs via YouTube

Overview

Google, IBM & Meta Certificates – 40% Off
One Coursera Plus subscription covers most Professional Certificates on Coursera.
Unlock All Certificates
A hands-on tutorial for building a retrieval-augmented generation pipeline with open-source Hugging Face models on AWS SageMaker. It uses MiniLM embeddings and a Pinecone vector index to retrieve FAQ context for LLM responses.

Syllabus

Open Source LLMs on AWS SageMaker
Open Source RAG Pipeline
Deploying Hugging Face LLM on SageMaker
LLM Responses with Context
Why Retrieval Augmented Generation
Deploying our MiniLM Embedding Model
Creating the Context Embeddings
Downloading the SageMaker FAQs Dataset
Creating the Pinecone Vector Index
Making Queries in Pinecone
Implementing Retrieval Augmented Generation
Deleting our Running Instances

Taught by

James Briggs

Reviews

Start your review of Hugging Face LLMs with SageMaker - RAG with Pinecone

Never Stop Learning.

Get personalized course recommendations, track subjects and courses with reminders, and more.

Someone learning on their laptop while sitting on the floor.