[arXiv]score: 0.14
Compact-Memory LLM Agents via Online Max-Member Clustering and Atom-Aware Packing
September 7, 2026
RSM-full achieves 83% of full-context quality using only 32% of the tokens at a 4k budget. The pipeline utilizes a cosine-gated max-member merge write rule and atom-aware grouped context packing to optimize the quality-token Pareto point. On AMA-Bench, it outperforms Online K-Means baselines by 3.5 to 6.0 percentage points.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy