All people

Researcher

Sholto Douglas

Sholto Douglas portrait

Member of Technical Staff at Anthropic

Douglas contributed to Gemini and Gemma at Google DeepMind and now leads work at Anthropic on scaling reinforcement learning for Claude. This work uses feedback from tasks and evaluations to improve a language model's reasoning after its initial training.

See something inaccurate or outdated?

Background

Current role
Member of Technical Staff 2025-present
Previously
2022-2025Research Engineer,
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Last editorial review: August 8, 2026

Citation Trend

1 snapshots
Citation snapshot on Jun 14, '2621,591Jun 14, '26
21,591citations as of Jun 14, '26
For today’s citation count and citations per paper, visit Google Scholar.