All people

Researcher

Zhihong Shao

ZS

Research Scientist at DeepSeek

Shao is a researcher at DeepSeek known for DeepSeekMath and for co-developing Group Relative Policy Optimization (GRPO), a reinforcement-learning method central to recent reasoning models. He is a research scientist at DeepSeek.

See something inaccurate or outdated?

Background

Current role
Research Scientist
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Last editorial review: August 13, 2026

Citation Trend

2 snapshots
Citation history from Aug 25, '25 to Jun 15, '2634,22022,63511,050Aug 25, '25Jun 15, '26
34,220citations as of Jun 15, '26+23,170 since Aug 25, '25
For today’s citation count and citations per paper, visit Google Scholar.