<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>LLM训练 on MessageDaily</title><link>https://inkeast.github.io/MessageDaily/tags/llm%E8%AE%AD%E7%BB%83/</link><description>Recent content in LLM训练 on MessageDaily</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Wed, 09 Sep 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://inkeast.github.io/MessageDaily/tags/llm%E8%AE%AD%E7%BB%83/index.xml" rel="self" type="application/rss+xml"/><item><title>Split-LLM 隐私失效审计：返回梯度的零模式完美暴露真实数据 精读</title><link>https://inkeast.github.io/MessageDaily/posts/2026-09-09-split-llm-gradient-privacy-failure-paper-reading/</link><pubDate>Wed, 09 Sep 2026 00:00:00 +0000</pubDate><guid>https://inkeast.github.io/MessageDaily/posts/2026-09-09-split-llm-gradient-privacy-failure-paper-reading/</guid><description>独立研究者对两节点 split-LLM 训练的预注册审计：隐私损失忽略诱饵行→其返回梯度恰好为零→零模式逐帧完美暴露真实数据（9 种子 4096/4096 全中）。所有运行通过前向隐私检查与质量检查，加上返回梯度后同检查失败。逐行裁剪+加噪以 0.01 nats 代价封堵，并诚实声明五类未测攻击面。</description></item><item><title>Locked at the Entrance 精读：RLVR 的多样性坍缩发生在推理的门口</title><link>https://inkeast.github.io/MessageDaily/posts/2026-09-06-locked-at-the-entrance-rlvr-paper-reading/</link><pubDate>Sun, 06 Sep 2026 00:00:00 +0000</pubDate><guid>https://inkeast.github.io/MessageDaily/posts/2026-09-06-locked-at-the-entrance-rlvr-paper-reading/</guid><description>RLVR 提升 pass@1 却坍缩解空间已是共识，但坍缩发生在哪一步始终未知。上海大学×伯明翰大学在 Countdown 任务上穷举解空间、按首操作数+算子划分入口族，把求解分解为 access×execution：PPO 覆盖 0.337→0.111，首算术操作前的似然偏移是下游的 11–16 倍，塞一个入口前缀就能让低覆盖族完成率 0.018→0.212——能力还在，只是不再进门。后层参数插值恢复 37% 覆盖且 pass@1 零损失。</description></item></channel></rss>