Learn Excel and Financial Modeling the Way Finance Teams Actually Use Them
Learn Backend Development Part-Time, Online
Overview
Google, IBM & Meta Certificates – 40% Off
One Coursera Plus subscription covers most Professional Certificates on Coursera.
Unlock All Certificates
A hands-on tutorial for building a retrieval-augmented generation pipeline with open-source Hugging Face models on AWS SageMaker. It uses MiniLM embeddings and a Pinecone vector index to retrieve FAQ context for LLM responses.
Syllabus
Open Source LLMs on AWS SageMaker
Open Source RAG Pipeline
Deploying Hugging Face LLM on SageMaker
LLM Responses with Context
Why Retrieval Augmented Generation
Deploying our MiniLM Embedding Model
Creating the Context Embeddings
Downloading the SageMaker FAQs Dataset
Creating the Pinecone Vector Index
Making Queries in Pinecone
Implementing Retrieval Augmented Generation
Deleting our Running Instances
Taught by
James Briggs