We introduce Context Language Models (CLMs), language models that natively manage their own context. We implement this by treating the context as a file and allowing the model to make unrestricted updates to this file.…
10 comments
Do you want your agent solving its own memory crisis, or do you want it solving the actual task? It can probably do both at the same time, but I suspect there is a non trivial cost associated with this.
A separate hypervisor agent that manages the main agent's context would be much better in my experience. You can run it on a different schedule and the main agent has to spend zero tokens thinking about it. This also makes it a lot easier to control when caches will be missed.
Not exactly like what this paper is suggesting, but similar in the sense it lets the model decide what and how to persist across turns.
I recreated this in Pi, with a max token limit on how long the note can be, to pressure the model to be concise. Ends up being cheaper than summary compaction too.
Do you have a public repo for your approach?
Read the full thread on Hacker News →
Related stories
- Context Language Modelsarxiv.orgHacker News · 3 points · 1 day ago
- Hacker News · 377 points · 16 days ago
- Hacker News · 1 points · 3 days ago
- DEV Community · 8 points · 21 days ago
- Lobsters · 33 points · 7 days ago
- Hacker News · 1 points · 2 days ago