<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>上下文鲁棒性 on MessageDaily</title><link>https://inkeast.github.io/MessageDaily/tags/%E4%B8%8A%E4%B8%8B%E6%96%87%E9%B2%81%E6%A3%92%E6%80%A7/</link><description>Recent content in 上下文鲁棒性 on MessageDaily</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Sat, 08 Aug 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://inkeast.github.io/MessageDaily/tags/%E4%B8%8A%E4%B8%8B%E6%96%87%E9%B2%81%E6%A3%92%E6%80%A7/index.xml" rel="self" type="application/rss+xml"/><item><title>When History Lies: Evaluating and Improving Tool Use under Misleading Multi-Turn Histories 精读</title><link>https://inkeast.github.io/MessageDaily/posts/2026-08-08-when-history-lies-paper-reading/</link><pubDate>Sat, 08 Aug 2026 00:00:00 +0000</pubDate><guid>https://inkeast.github.io/MessageDaily/posts/2026-08-08-when-history-lies-paper-reading/</guid><description>深度精读论文《When History Lies》——首次形式化『历史诱导的策略劫持』现象：结构有效、语义合理的历史轨迹仍可劫持 Agent 已有的正确策略，在 Qwen3-1.7B 上翻转 32.1% 的正确决策。论文提出 ContextPollute-Bench（同步三视图 Original/Polluted/Oracle State，十一类干扰算子）与 Oracle-OPD 方法（基于 Oracle State 的教师通过反向 KL 在策略蒸馏迁移到仅观察污染历史的学生），将 1.7B 模型的 BTA 从 47.2% 提升至 87.0%，8B 教师蒸馏后达 91.9%，并迁移到干净历史、未见工具与噪声多跳 QA。</description></item></channel></rss>