Context that stays slim
The window is spent on your task, not on ceremony — every reduction is deliberate, not incidental:
- oversized tool results are evicted to files — a preview stays inline, the agent pages the rest back on demand
- skills are indexed, lazily loaded & searched: what rides on every request is a grouped overview — one line per source, descriptions omitted, large sources collapsed to their first few names plus a count
- compaction summaries precompute in the background from ~60% full and fold in at ~80% without pausing the session
- the prompt prefix is byte-stable, so provider caches keep hitting — and a per-turn meter shows how full the window actually is