跳到正文
elvis· @omarsar0 · X·· 3 小时前AI 评分60
AI 导读

Meta 等机构发布论文提出 Context Language Models(CLM),将上下文作为可读写的文件让模型用 Bash 自主编辑,决定保留、重写或删除内容。

正文

Banger paper from Meta and colleagues.

(bookmark it)

They discuss benefits of giving your agent write access to its own context.

In other words, they investigate how effective it is to allow a language model to natively manage its context.

Context Language Models keep the context as a file the model edits with Bash, so the model decides what to keep, rewrite or remove.

It's a bit different from RLM, for those who are wondering but it pulls an interesting theme.

RLMs place a large input in an external variable that the model can read and process recursively, but they do not let the model directly edit its own live interaction context. CLMs instead expose the live context as a read-write file, including information accumulated during execution.

Applied zero-shot, this gets 11.4% higher accuracy with 21.5% fewer FLOPs on BrowseComp-Plus than existing context-management strategies.

On a 24-hour task where a swarm of agents works across six repositories, it gets 65% more improvement for the same compute.

The strategy can also be trained. Online RL improves Qwen3.5-9B on BrowseComp-Plus by 47.6% while using 12% fewer FLOPs.

Edits in the middle of the context break prefix caching, so the authors add Suffix Cache Reuse, which cuts server compute by 35% against standard SGLang.

Paper: https://academy.dair.ai/papers/context-language-models-2609.37725

来源:elvis · x.com