tracked author
Yu Wu (吴俣)
Math-Shepherd coauthor and DeepSeek LLM Alignment Team lead. Homepage states that his team pioneered GRPO and DeepSeek-R1-Zero; Google Scholar identifies him as Yu Wu (吴俣), DeepSeek AI. X search result matches the same name and DeepSeek technical-staff role.