Anthropic:Transformer Circuits(可解释性研究)·· 13 小时前AI 评分32
Anthropic 可解释性研究更新:用特征语言重访 Transformer 数学框架,并将可解释性应用于生物学
Circuits Updates — July 2025 A collection of small updates: revisiting A Mathematical Framework and applications of interpretability to biology.
AI 导读
Anthropic 可解释性团队发布 7 月研究更新,用特征语言重访《A Mathematical Framework for Transformer Circuits》,重新解释 copy head、previous token head 和 induction head 的 OV 与 QK 电路。
来源:Anthropic:Transformer Circuits(可解释性研究) · transformer-circuits.pub