FrontierAI.Engineer
RAG & Grounded Generation

Contextual Compression

Contextual compression trims retrieved passages down to only the sentences or spans relevant to the query before they enter the prompt. By discarding boilerplate and off-topic text, it lowers token cost and noise while keeping the evidence the generator actually needs. The result is a higher signal-to-noise context window and often better faithfulness at reduced expense.