Reasoning instructions can distort vision–language model evaluation by prompting CoT-prefix scoring, study finds
Read the original at arxiv.org→arXiv:2609.29278v1 Announce Type: new Abstract: Chain-of-thought (CoT) instructions can distort multiple-choice VLM evaluation when a scorer appends a reasoning cue but reads answer-label logits before the model...
Original headline: "Reasoning Instructions Can Break Answer Decoding in Vision--Language Models"
Coverage timeline
- Sep 25, 04:00 UTC arXiv cs.CL lead source Reasoning Instructions Can Break Answer Decoding in Vision--Language Models