few thoughts, I've been sent this paper many times today, and it's cool!
1. it's a v clever idea, and I'd say even more extreme than RLMs on the spectrum of ReAct-style vs. pure context offloading (so no they're not the same)
2. im hopeful to see harness designs that use this principle, and they somewhat go hand-in-hand with model + RLM progress as well
3. the KV issues scare me admittedly, and the proposed fix is a bit hacky despite it seeming to work well for their results. that being said, I suspect different model shapes won't have this issue, so it's fixable! another + for this direction
‼️The Bitter Lesson for context management: Giving LMs unrestricted control over their context beats human-designed SOTA!
Introducing 🩵Context Language Models (CLMs)🩵
- Natively manage their own context
- Treat context as a file
- Learn policies in CLM weights, no harness