Researcher · Founder
Andi Peng
AP
Current role not listed
Peng worked on reinforcement learning and post-training for Claude 3.5 through 4.5 at Anthropic. Her research asks how AI agents can learn nuanced human preferences from feedback, combining technical alignment work with experience in government technology policy.
See something inaccurate or outdated?
Humans&
Anthropic