MemSyco-Bench: Benchmarking Sycophancy in Agent Memory

작성자

카테고리:

← 피드로
arXiv cs.AI · Zhishang Xiang, Zerui Chen, Yunbo Tang, Zhimin Wei, Ruqin Ning, Yujie Lin, Qinggang Zhang, Jinsong Su · 2026-07-03 AI

[Submitted on 1 Jul 2026 (v1), last revised 2 Jul 2026 (this version, v2)]

View PDF HTML (experimental)

Abstract:Memory has emerged as a cornerstone of modern LLM-based agents, supporting their evolution from single-turn assistants to long-term collaborators. However, memory is not always beneficial: retrieved memories often induce a critical issue of sycophancy, causing agents to over-align with the user at the cost of factual accuracy or objective reasoning. Despite this emerging risk, existing memory benchmarks primarily evaluate whether memories are correctly stored, retrieved, or updated, while overlooking how retrieved memories influence downstream reasoning and decision-making. To bridge this gap, we propose MemSyco-Bench, a comprehensive benchmark for evaluating memory-induced sycophancy in agent systems. MemSyco-Bench measures when memory should influence a decision and how valid memory should be used. Specifically, it covers five tasks that assess whether agents can reject memory as factual evidence, respect its applicable scope, resolve conflicts between memory and objective evidence, track memory updates, and use valid memory for personalization. All related resources are collected for the community at this https URL.

Submission history

From: Qinggang Zhang [view email]
[v1] Wed, 1 Jul 2026 15:30:33 UTC (10,562 KB)
[v2] Thu, 2 Jul 2026 15:40:30 UTC (10,475 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2607.01071

코멘트

답글 남기기

이메일 주소는 공개되지 않습니다. 필수 필드는 *로 표시됩니다