Data & Infrastructure
Response Caching
Client-side or gateway-side caching of complete LLM responses by prompt hash. Different from prompt caching (server-side KV cache reuse). Useful for high-repetition workloads where identical prompts appear.
Related terms
Related on this site
Where this fits
Response Caching is part of the Data & Infrastructure vocabulary used in the Generative AI Maturity Framework. See the full glossary for the complete set of 149 defined terms, or take the free maturity assessment to see where your organisation stands.