56. RAG From Scratch
A basic RAG pipeline consists of:
- load documents
- split documents into chunks
- create embeddings
- store vectors
- embed the user query
- retrieve similar chunks
- construct a prompt
- generate an answer
The central loop is:
Query
↓
Retrieve
↓
Context
↓
Generate