tracked author
Qiying Yu (禹棋赢)
DAPO共同第一作者。个人主页将其列为清华大学智能产业研究院博士生,并将多模态基础模型与大语言模型强化学习列为研究方向;DAPO arXiv v2 的贡献声明将其列为项目负责人。
1 archived notes Homepage
Representative Papers
来自作者已核验个人主页的重点论文;本站单篇归档见下方 Related Notes。
- 01 CapsFusion: Rethinking Image-Text Data at Scale CVPR 2024 · 2024
- 02 Generative Multimodal Models are In-Context Learners CVPR 2024 · 2024
- 03 Generative Pretraining in Multimodality ICLR 2024 · 2024
- 04 Multimodal Molecular Pretraining via Modality Blending ICLR 2024 · 2024
- 05 Multimodal Federated Learning via Contrastive Representation Ensemble ICLR 2023 · 2023
- 06 Adversarial Contrastive Learning via Asymmetric InfoNCE ECCV 2022 · 2022
- 07 EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters arXiv · 2024
Related Notes
按论文归档时间排序,展示该作者在本站已经出现的材料。
归档 更新
DAPO: An Open Source LLM Reinforcement Learning System at Scale
在 Qwen2.5 32B 与 AIME 2024 设置中,用解耦裁剪、动态采样、token 级损失和超长奖励整形将朴素 GRPO 的 avg@32 从 30 提高到 50,并开源代码、数据与模型。