
Tom BrownResearcher
Anthropic
Co-founder
via Anthropic
Researcher

Member of Technical Staff at Anthropic
Douglas contributed to Gemini and Gemma at Google DeepMind and now leads work at Anthropic on scaling reinforcement learning for Claude. This work uses feedback from tasks and evaluations to improve a language model's reasoning after its initial training.
See something inaccurate or outdated?
Anthropic
Google DeepMindWork and education history is primarily focused on AI-relevant roles and may not be comprehensive.
Last editorial review: August 8, 2026