Chujie Zheng

I am a member of technical staff of the Qwen Team, working on RL scaling. My responsibilities span the following:

  • Leading the OPD team for post-training consolidation
  • Led RL research for large-scale RL training
  • Initiated and maintaining the RL/OPD infrastructure

Recent Projects

Full paper list

  • Qwen3.8: A New Bar for Coding and Cowork
    [blog]
  • Qwen3.7: The Agent Frontier
    [blog]
  • Qwen3.5: Towards Native Multimodal Agents
    [blog] [model]
  • Stabilizing Reinforcement Learning with LLMs: Formulation and Practices
    Chujie Zheng, Kai Dang, Bowen Yu, Mingze Li, Huiqiang Jiang, Junrong Lin, Yuqiong Liu, Hao Lin, Chencan Wu, Feng Hu, An Yang, Jingren Zhou, Junyang Lin
    [paper]
  • Group Sequence Policy Optimization
    Chujie Zheng, Shixuan Liu, Mingze Li, Xiong-Hui Chen, Bowen Yu, Chang Gao, Kai Dang, Yuqiong Liu, Rui Men, An Yang, Jingren Zhou, Junyang Lin
    [blog] [paper]
  • Qwen3: Think Deeper, Act Faster
    [blog] [model]
  • QwQ-32B: Embracing the Power of Reinforcement Learning
    [blog] [model]
  • QwQ: Reflect Deeply on the Boundaries of the Unknown
    [blog] [model]

Education

  • Aug 2020 – Jun 2025. Ph.D in Computer Science and Technology, Tsinghua University. Advisor: Minlie Huang
  • Nov 2023 – Jun 2024. Visiting Researcher. University of California, Los Angeles. Host: Nanyun Peng
  • Aug 2016 – Jul 2020. B.Sc. in Mathematics and Physics, Tsinghua University

Selected Awards and Honors

  • CIPS Doctoral Dissertation Incentive Program (中国中文信息学会博士学位论文激励计划, Top 10 in China), 2025
  • ACL SAC Award, 2025
  • Spotlight Recipient of the 2025 WAIC Yunfan Award (云帆奖·明日之星), 2025
  • National Scholarship (Top 2/100), Ministry of Education of China, 2019