Stage 2 · Core
LLM Apps: Caching & RAG
Put retrieval and caching in front of a model without leaking data or serving the wrong answer.
You'll be able to: Design a semantic cache and a RAG pipeline that respect tenants and permissions.
0 of 2 done
-
1
Semantic Caching for LLMs
A semantic cache is only as safe as its key: key on meaning alone and it serves wrong answers and leaks data across tenants.
-
2
Post-Filter RAG Leaks
If you check permissions *after* retrieval, the LLM has already read the secret — pre-filter by ACL inside the search.
- Prompt Injection Through ToolsAnnounced at the end of an episode in this track. Subscribe to catch it.
- Pre-Filter RAGAnnounced at the end of an episode in this track. Subscribe to catch it.