the.ai

Context / Methods

verified

Context Compression

If the window is expensive and much of what goes into it is redundant, spend some compute shortening the context before the model reads it. Summarise old turns, drop passages the query does not touch, replace verbatim documents with extracted facts. The context that arrives is smaller and, if the compression was any good, says the same things.

Viz primitive · budget-splitdiscarded-tokens = 30

discarded-tokens holds 33% of the budget; rest holds the remaining 67%.

Context tokens compression discards, against the ones it keeps, in tokens. Drag the compression up to watch most of the context go — and note that the questions it makes unanswerable are the specific ones, not a random sample of them.

30

Reviewed by opendroid · 2026-08-18