Chujie Zheng
I am a member of technical staff of the Qwen Team, working on RL scaling. My responsibilities span the following:
- Leading the OPD team for post-training consolidation
- Led RL research for large-scale RL training
- Initiated and maintaining the RL/OPD infrastructure
Recent Projects
- Qwen3.8: A New Bar for Coding and Cowork
[blog] - Qwen3.7: The Agent Frontier
[blog] - Qwen3.5: Towards Native Multimodal Agents
[blog] [model] - Stabilizing Reinforcement Learning with LLMs: Formulation and Practices
Chujie Zheng, Kai Dang, Bowen Yu, Mingze Li, Huiqiang Jiang, Junrong Lin, Yuqiong Liu, Hao Lin, Chencan Wu, Feng Hu, An Yang, Jingren Zhou, Junyang Lin
[paper] - Group Sequence Policy Optimization
Chujie Zheng, Shixuan Liu, Mingze Li, Xiong-Hui Chen, Bowen Yu, Chang Gao, Kai Dang, Yuqiong Liu, Rui Men, An Yang, Jingren Zhou, Junyang Lin
[blog] [paper] - Qwen3: Think Deeper, Act Faster
[blog] [model] - QwQ-32B: Embracing the Power of Reinforcement Learning
[blog] [model] - QwQ: Reflect Deeply on the Boundaries of the Unknown
[blog] [model]
Education
- Aug 2020 – Jun 2025. Ph.D in Computer Science and Technology, Tsinghua University. Advisor: Minlie Huang
- Nov 2023 – Jun 2024. Visiting Researcher. University of California, Los Angeles. Host: Nanyun Peng
- Aug 2016 – Jul 2020. B.Sc. in Mathematics and Physics, Tsinghua University
Selected Awards and Honors
- CIPS Doctoral Dissertation Incentive Program (中国中文信息学会博士学位论文激励计划, Top 10 in China), 2025
- ACL SAC Award, 2025
- Spotlight Recipient of the 2025 WAIC Yunfan Award (云帆奖·明日之星), 2025
- National Scholarship (Top 2/100), Ministry of Education of China, 2019
