Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

TOPIC · ENTITY #12445

rationales

Related event timeline, sources and context from the news index.

EVENT TIMELINE

1

RESEARCH · 1 SOURCE · arXiv cs.AI

Study shows rationales mainly affect verifier judgments, not answer accuracy

The paper introduces a message-intervention diagnostic that holds evidence and candidate answers constant while varying only the rationale passed from a reasoner to a verifier. On 400 examples across MuSiQue, HotpotQA and 2WikiMultiHopQA using DeepSeek as generator and verifier, faithful rationales add almost no answer accuracy versus no rationale, but corrupted rationales substantially change verifier support judgments (10–22% under a blind verifier prompt, 34–55% with explicit rationale-checking), while final answers change less (2–30%); human audits reveal instances of model overtrust.

7.0