High Performance (Realtime) RAG Chains: From Basic to Advanced
Key Takeaways
This video teaches how to build high-performance real-time Retrieval Augmented Generation systems using Llama 3, GroqCloud, LangChain, and Redis
Original Description
I will build high-performance real-time Retrieval Augmented Generation (RAG) systems in this tutorial using Llama 3, GroqCloud, LangChain, and Redis.
The written tutorial, along with the code:
https://www.rabbitmetrics.com/realtime-rag-with-llama-3
Here's a tutorial on how to set up a database on Redis:
https://youtu.be/MXhfLUoIRno
The dataset used in the video
https://huggingface.co/datasets/ashraq/fashion-product-images-small
▬▬▬▬▬▬ V I D E O C H A P T E R S & T I M E S T A M P S ▬▬▬▬▬▬
0:00 Intro
1:29 Installing libraries and connecting to a databases
3:25 Simple RAG Chain
5:29 Hybrid RAG Chain
7:17 Contextualized RAG Chain
AI explanation not available for this lesson yet
This lesson is still being prepared for the AI tutor. In the meantime, explore lessons that are ready.
Browse explainer-ready lessons →
More on: RAG Basics
View skill →Related Reads
📰
📰
📰
📰
The architect frowned: “Your RAG users have to wait 8 seconds to see the first word?
Medium · Programming
Your RAG System Can’t Find “Error Code E-4471” and Vector Search Is Why
Medium · AI
Your RAG System Can’t Find “Error Code E-4471” and Vector Search Is Why
Medium · Machine Learning
Your RAG System Can’t Find “Error Code E-4471” and Vector Search Is Why
Medium · Data Science
Chapters (5)
Intro
1:29
Installing libraries and connecting to a databases
3:25
Simple RAG Chain
5:29
Hybrid RAG Chain
7:17
Contextualized RAG Chain
🎓
Tutor Explanation
DeepCamp AI