Data & Infrastructure

Response Caching

Client-side or gateway-side caching of complete LLM responses by prompt hash. Different from prompt caching (server-side KV cache reuse). Useful for high-repetition workloads where identical prompts appear.

Related terms

Framework dimensions

Next steps

Where this fits

Response Caching is part of the Data & Infrastructure vocabulary used in the Generative AI Maturity Framework. See the full glossary for the complete set of 149 defined terms, or take the free maturity assessment to see where your organisation stands.