AI Fundamentals

Contextual Compression

A RAG technique that summarizes or filters retrieved chunks before feeding them to the generation model, reducing token count while preserving relevance. Common technique for long-context workloads where naive retrieval overflows the context window.

Related terms

Next steps

Where this fits

Contextual Compression is part of the AI Fundamentals vocabulary used in the Generative AI Maturity Framework. See the full glossary for the complete set of 149 defined terms, or take the free maturity assessment to see where your organisation stands.