
Tom BrownResearcher
Anthropic
Co-founder
via Anthropic
Researcher
Member of Technical Staff at Anthropic
Tang develops the reinforcement-learning methods used to improve frontier language models. He helped scale online RL for Llama 3.3 and 4, contributed to Gemini post-training and agent tool use, and now works on pretraining at Anthropic after a short period researching reasoning at Mistral.
See something inaccurate or outdated?
Anthropic
Mistral AI
Meta
Google DeepMindWork and education history is primarily focused on AI-relevant roles and may not be comprehensive.
Last editorial review: August 9, 2026