
Zhihong ShaoResearcher
DeepSeek
Research Scientist
via DeepSeek
Researcher
Researcher at DeepSeek
Wang is a core contributor to DeepSeek's open language models. His work spans instruction tuning, mathematical reasoning and reinforcement learning, including DeepSeek-R1, which showed that strong step-by-step reasoning can emerge through reward-driven training, and the efficient V3 model family.
See something inaccurate or outdated?
DeepSeekWork and education history is primarily focused on AI-relevant roles and may not be comprehensive.
Last editorial review: August 10, 2026