跳到正文
原文
arXiv cs.AI· Junyu Guo, Yuchen Fang, Shangding Gu, Costas Spanos, James Demmel, Javad Lavaei·· 8 小时前AI 评分45

上下文变化时:理解 LLM 的信息更新失败

When Context Changes: Understanding Update Failures in LLMs

AI 导读

研究者提出 Controlled In-Context Memory(CICM)基准,用于评测 LLM 在对话与智能体日志中追踪并使用更新信息的能力,发现即便前沿推理模型也会在上下文变更后答出旧值,即 stale binding。

来源:arXiv cs.AI · arxiv.org