Abstract
We introduce Context Language Models (CLMs), language models that natively manage their own context. We implement this by treating the context as a file and allowing the model to make unrestricted updates to this file. This allows the model to learn what is most important to maintain in context, and naturally extends to multi-agent systems where multiple agent contexts coexist as files. Building CLMs zero-shot with existing models outperforms SOTA context management strategies across a variety of tasks: 11.4% higher accuracy with 21.5% fewer FLOPs on BrowseComp-Plus, 5% higher scores with 59% fewer FLOPs on 12-hour EdgeBench, and 65% greater improvement with the same compute on a 24-hour multi-repository agent-swarm task. Moreover, by shifting context management from external harness control to intrinsic model behavior, CLMs naturally enable both in-context and parametric learning of context-management strategies. We show that CLMs can be steered with natural-language instructions evolved through a standard skill-optimization loop, improving held-out accuracy by up to 35.9 points on a context-management task while reducing compute. We also introduce an online reinforcement learning method for CLMs, improving Qwen3.5-9B performance on BrowseComp-Plus by 47.6% while using 12% fewer FLOPs. Finally, we co-design Suffix Cache Reuse for CLM serving, further reducing server-side compute by 35% relative to standard SGLang at matched performance.
Community
The Bitter Lesson for context management: Giving LMs unrestricted access to their own context beats human-designed SOTA!
Introducing ๐ฉต Context Language Models (CLMs) ๐ฉต
- Natively manage their own context
- Treat context as a file
- Learn policies in CLM weights, no harness
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- Continuous Context Management (2026)
- Beyond Skill Evolution: Self-Evolving Context Management Policies for Long-Horizon Agent Harnesses (2026)
- ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL (2026)
- Grow the Harness, Not the Context: From Strategy-Free Scaffolds to Reusable Specialist Agents (2026)
- Language Models Can Control Their Own Attention (2026)
- Context as an Environment: Programmatic Context Management for Long-Horizon Agents (2026)
- Dyad: Extending Large Language Models with Native Typed Decision-Making (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2609.37725 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper