All people

Researcher

Yunhao Tang

YT

Member of Technical Staff at Anthropic

Tang develops the reinforcement-learning methods used to improve frontier language models. He helped scale online RL for Llama 3.3 and 4, contributed to Gemini post-training and agent tool use, and now works on pretraining at Anthropic after a short period researching reasoning at Mistral.

See something inaccurate or outdated?

Background

Current role
Member of Technical Staff 2025-present
Previously
2025-2025Research Scientist,
2024-2025Research Scientist,
2021-2024Research Scientist,
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Last editorial review: August 9, 2026

Citation Trend

1 snapshots
Citation snapshot on Aug 9, '2619,077Aug 9, '26
19,077citations as of Aug 9, '26
For today’s citation count and citations per paper, visit Google Scholar.