跳到正文
原文
arXiv:cs.AI(全量分类)· Jiameng Zhang, Hongqiu Wu·· 5 小时前AI 评分47

推理链到底传递了什么?角色分工问答中的消息干预研究

What Do Rationales Communicate? A Message-Intervention Study in Role-Specialized QA

AI 导读

一项消息干预诊断研究固定证据与候选答案、只改变从推理器传给验证器的 rationale,在 400 条 MuSiQue、HotpotQA 和 2WikiMultiHopQA 样本上以 DeepSeek 分别担任生成器与验证器测试。

正文

View PDF HTML (experimental)

Abstract:Role-specialized QA pipelines increasingly pass rationales from a reasoner to a verifier, but it is unclear what this message actually buys: better answers, stronger support assessment, or a new failure surface. We introduce a message-intervention diagnostic that fixes the evidence and candidate answer while varying only the rationale passed across the reasoner-to-verifier boundary. On 400 MuSiQue, HotpotQA, and 2WikiMultiHopQA examples with DeepSeek as generator and verifier, faithful rationales add almost no answer accuracy over no rationale, while corrupted rationales strongly alter support judgments. Under a blind verifier prompt, harmless paraphrases shift support by only 0--2.5%, whereas corrupted rationales shift support by 10--22%; an explicit rationale-checking prompt amplifies the same pattern to 34--55%. Final answers move less (2--30%), and only 2.9--35.3% of corrupted support flips co-occur with answer changes. Human audits show why this matters: 16/42 valid corruptions are corruption-overtrust cases, and blind humans reject or mark unclear 9/10 audited corrupted rationales that the model accepts. Cross-model and task-boundary checks show when the channel is active, amplified, inert, or folded into the task label. Rationale sharing should be evaluated as a verification-message mechanism, not merely as a route to higher answer accuracy.
Comments: 14 pages, 6 figures, 9 tables. Preprint
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
Cite as: arXiv:2610.00018 [cs.AI]
  (or arXiv:2610.00018v1 [cs.AI] for this version)
  https://doi.org/10.48550/arXiv.2610.00018

arXiv-issued DOI via DataCite

Submission history

From: Jiameng Zhang [view email]
[v1] Mon, 27 Jul 2026 15:52:20 UTC (1,405 KB)

来源:arXiv:cs.AI(全量分类) · arxiv.org