WakeKV:反应式可逆 KV 驻留策略论文发布
Get the story
2026年10月5日,arXiv Computation and Language 发布 WakeKV 论文(一手报道)。WakeKV 提出一种反应式 KV 驻留策略,针对中途改变读取行为的注意力头,将其移入可恢复的 CPU 储层,而非冻结或永久驱逐其状态。研究在 1.5B–8B 三个模型上,于针检索、长链式推理与多轮回忆三种场景测试,测得多数头在生成中至少改变一次读取行为;在匹配内存或预算下,WakeKV 在五个模型-场景组合上优于 SnapKV、uniform R-KV 与 ReasonAlloc 三个基线。目前未见后续报道或同行评议信息。
Generated from reports · updated 16 hr ago
Timeline
Follow the coverage from different angles.
- arXiv · Computation and LanguageWakeKV:为中途改变读取行为注意力头设计的反应式可逆 KV 驻留方案
WakeKV 提出一种反应式 KV 驻留策略,把降温的注意力头移入可恢复的 CPU 储层,而非冻结或永久驱逐其状态。研究在 1.5B-8B 三个模型、针检索、长链式推理与多轮回忆三种场景下测得多数头在生成中至少改变一次读取行为;在匹配内存或预算下,WakeKV 在五个模型-场景组合上优于 SnapKV、uniform R-KV 与 ReasonAlloc 三个基线。
Heat trend
Current heat 7·Comparable peak 10(Oct 5)·Comparable change over 24 hours –
The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.