SheepNav
精选3个月前0 投票

When Helpfulness Becomes Sycophancy: Sycophancy is a Boundary Failure Between Social Alignment and Epistemic Integrity in Large Language Models

arXiv:2605.05403v1 Announce Type: new Abstract: This position paper argues that sycophancy in LLMs is a boundary failure between social alignment and epistemic integrity. Existing work often operationalizes sycophancy through external behavior such as agreement with incorrect user beliefs, position reversals, or deviation from an objective standard of correctness. These formulations capture only overt forms of the phenomenon and leave subtler boundary failures involving epistemic integrity and s

延伸阅读

  1. UI-Venus-2:跨越移动、网页与桌面的通用GUI智能体
  2. SCAFFOLD:大规模计算机科学论文图表理解数据集,助力视觉语言模型
  3. OpenAgentFlow:为异构AI智能体舰队构建全系统安全边界
查看原文