ai.hackcv
论文精选 82arXiv

Sycophancy Undermines Epistemic Vigilance in Cooperative Vision-Language Tasks· 献媚削弱视觉-语言合作任务中的认知警惕性

To maintain common ground in cooperative conversation, humans iteratively update their beliefs as conversation participants share new information; participants who are epistemically vigilant detect when new information conflicts with prior beliefs and take steps to repair these conflicts. In order for AI systems to serve as reliable partners in complex cooperative tasks, they must similarly weigh incoming information against their own private evidence and shared context and appropriately surface inconsistencies when they arise. To measure the epistemic vigilance of vision-language models in cooperative settings, we present an information-asymmetric, dialog-based "spot-the-difference" task. Two models are privately shown one image each, and must determine through conversation whether the im

领域:cs.CL作者:Rupak Sarkar、Neha Srikanth、Saloni Gupta
相关推荐

本站内容由 LLM 精选聚合,原文版权归 arXiv 所有 · 摘录仅供参考