Research questionHow can long-horizon LLM agents preserve answer quality under tight prompt-token budgets?As interactions accumulate, supplying the full history becomes costly or exceeds context limits, while compressing memory can omit information needed for later answers. The central difficulty is retaining useful information within a small prompt budget.