Local RAG with Llama 3.1 for PDFs - Private Chat with Documents using LangChain and Streamlit
Most AI Pilots Fail to Scale. MIT Sloan Teaches You Why — and How to Fix It
Pass the PMP® Exam on Your First Try — Expert-Led Training
Overview
Google, IBM & Meta Certificates — All 10,000+ Courses at 40% Off
One annual plan covers every course and certificate on Coursera. 40% off for a limited time.
Get Full Access
Discover how to build a local Retrieval-Augmented Generation (RAG) system for efficient document processing using Large Language Models (LLMs) in this comprehensive tutorial video. Learn to extract high-quality text from PDFs, split and format documents for optimal LLM performance, create vector stores with Qdrant, implement advanced retrieval techniques, and integrate local and remote LLMs. Follow along to develop a private chat application for your documents using LangChain and Streamlit, covering everything from project structure and UI design to document ingestion, retrieval methods, and deployment on Streamlit Cloud.
Syllabus
- What is RagBase?
- Text tutorial on MLExpert.io
- How RagBase works
- Project Structure
- UI with Streamlit
- Config
- File Upload
- Document Processing Ingestion
- Retrieval Reranker & LLMChainFilter
- QA Chain
- Chat Memory/History
- Create Models
- Start RagBase Locally
- Deploy to Streamlit Cloud
- Conclusion
Taught by
Venelin Valkov