Context Language Models

(arxiv.org)

56 points | by emersonmacro 5 hours ago ago

11 comments

  • bob1029 an hour ago ago

    I would be concerned with context management consuming limited attention resources.

    Do you want your agent solving its own memory crisis, or do you want it solving the actual task? It can probably do both at the same time, but I suspect there is a non trivial cost associated with this.

    A separate hypervisor agent that manages the main agent's context would be much better in my experience. You can run it on a different schedule and the main agent has to spend zero tokens thinking about it. This also makes it a lot easier to control when caches will be missed.

  • svachalek an hour ago ago

    Wow. Context management is one of the big remaining hassles with modern LLMs so this could be big. The obvious complication is cache busting so it's also exciting they investigated solutions for that.

  • Bolwin an hour ago ago

    The biggest discovery might actually be that they ignored regular caching rules and kept invalid cache suffixes and it didn't hurt performance

  • visarga an hour ago ago

    Can't we do this trick today with any model? Just send the file as next context. Of course you pay the price for cache misses, depending how deep you make changes, while CLM just ignores the recomputation.

    • nsingh2 an hour ago ago

      One approximation of this is the experimental context management Codex has been moving towards (not released yet). Rather than relying on summary compaction, the model maintains notes as it works and as it approaches the context limit. A new session is just a fresh context with those notes attached, and a pointer back to the previous session.

      Not exactly like what this paper is suggesting, but similar in the sense it lets the model decide what and how to persist across turns.

      I recreated this in Pi, with a max token limit on how long the note can be, to pressure the model to be concise. Ends up being cheaper than summary compaction too.

      • visarga an hour ago ago

        That is similar to what I am thinking... not just edit the context as a file or string, but have a way to evict blocks and replace them with summary notes and also be able to retrieve them on demand.

        Do you have a public repo for your approach?

      • ijidak 31 minutes ago ago

        Memento