训练式智能体上下文管理论文发布
Get the story
2026-10-05,arXiv Machine Learning Theory 发布《Trained Agentic Context Management》论文(一手来源),提出训练式智能体上下文管理方法。研究在最简工具框架下微调 Qwen3.6-35B-A3B,工具包括以任意提示词调用自身、读取输入上下文指定区间的 token。在 OOLONG-synth 基准上,当文档长度超过 40K token 时,模型仅用 8,000 token 上下文即达到 GPT-5.4 使用 1M token 上下文的同等强度。论文共 17 页、6 幅图、4 张表,并给出代码链接。
Generated from reports · updated 17 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Machine Learning Theory精选Trained Agentic Context Management 研究用自调用工具训练 Qwen3.6-35B-A3B
论文提出训练式智能体上下文管理方法,在最简工具框架下微调 Qwen3.6-35B-A3B,工具包括以任意提示词调用自身以及读取输入上下文指定区间的 token。在 OOLONG-synth 基准上,当文档长度超过 40K token 时,模型仅用 8,000 token 上下文就达到 GPT-5.4 使用 1M token 上下文的同等强度。论文共 17 页、6 幅图、4 张表,并给出代码链接。
Heat trend
Current heat 6·Comparable peak 10(Oct 5)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.