Caching
A section of the "LLMs in production" topic. The questions check an understanding of practice, not memory of specific library APIs. After each answer comes a review: why the correct option is right, what is wrong with each incorrect one and which chapter to read.
Quiz difficulty Formats: several correct options, one correct option
01
What we check
- Exact caching
- A leak through the cache
- A semantic cache
- A client-side and a server-side cache
- Why cache at all
- Total savings from optimizations
02
Where to read
- Li, AI Agents in Depthsec. 7.6.3
- Huyen, AI Engineeringch. 10
- Lakshmanan, Hapke, Generative AI Design Patternsch. 8, Pattern 25