Rohan Paul· @rohanpaul_ai · X·2026-09-05 11:13· 2026-09-05AI 评分38AI 导读Google 与 KAIST 的论文提出 Declarative Attention(DA),模型通过 <global>、<focus>、<local> 三种模式在思维链中声明需要关注的上下文,推理引擎据此跳过大部分 KV cache 读取,无需额外的外置打分器。整理与数据来源:AIHOT来源:Rohan Paul · x.com#Google#推理#论文/研究#部署/工程