
Jingren ZhouResearcher
Alibaba
Senior Vice President, Chief Scientist
via Alibaba
Researcher
Researcher, Qwen Team at Alibaba
Zheng helps build Qwen's reasoning models and proposed Group Sequence Policy Optimization, a reinforcement-learning method for stabilizing large-scale training. His work spans Qwen3, QwQ, and tools for detecting errors in step-by-step mathematical reasoning.
See something inaccurate or outdated?
AlibabaWork and education history is primarily focused on AI-relevant roles and may not be comprehensive.
Last editorial review: August 9, 2026