All people

Researcher · Executive

Max Schwarzer

Max Schwarzer portrait

Member of Technical Staff at Anthropic

Schwarzer helped lead the post-training of OpenAI's o1 and o3 reasoning models before joining Anthropic. His earlier research showed how reinforcement-learning agents can learn efficiently from limited experience, exposed fragile evaluation practices and found that periodically resetting parts of an agent can improve learning.

See something inaccurate or outdated?

Background

Current role
Member of Technical Staff Mar 2026-present
Previously
Nov 2023-Mar 2026Vice President of Research; Post-training Lead,
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Last editorial review: August 10, 2026

Citation Trend

2 snapshots
Citation history from Aug 25, '25 to Jun 15, '267,2515,0292,807Aug 25, '25Jun 15, '26
7,251citations as of Jun 15, '26+4,444 since Aug 25, '25
For today’s citation count and citations per paper, visit Google Scholar.
Max Schwarzer: Researcher profile · Turing Tree