tracked author

Yujiang Li

SAO and CompactionRL co-first author and Tsinghua computer science PhD student from 2025. Yuxiao Dong's student list and the MLE-RL OpenReview record connect his work to reasoning, coding agents, and reinforcement learning; both papers connect this work to his Z.AI internship.

2 archived notes X: not-found

Related Notes

按论文归档时间排序,展示该作者在本站已经出现的材料。

归档

CompactionRL: Reinforcement Learning with Context Compaction for Long Horizon Agents

CompactionRL 用独立 critic 和按后续 token 数折扣的跨段优势训练上下文压缩轨迹;两个模型在启用压缩的 coding 评测中提高 Pass@1,但 SUPO 已覆盖摘要—执行联合训练与全 token 归一化,且论文未报告直接对照或总计算。

待审阅 2607.05378-compactionrl-context-compaction-agent-rl Agent MemoryAgent RLCredit Assignment