跳到正文
原文
Rohan Paul· @rohanpaul_ai · X·· 2026-09-05AI 评分38
AI 导读

Google 与 KAIST 的论文提出 Declarative Attention(DA),模型通过 <global>、<focus>、<local> 三种模式在思维链中声明需要关注的上下文,推理引擎据此跳过大部分 KV cache 读取,无需额外的外置打分器。

整理与数据来源:AIHOT

来源:Rohan Paul · x.com