All people

Researcher · Executive

Evan Hubinger

Evan Hubinger portrait

Member of Technical Staff, Manager at Anthropic

Hubinger helped define the problem of inner alignment: whether a trained model develops its own internal objective rather than the one intended by its designers. At Anthropic he leads alignment stress-testing, probing models for deceptive behavior, hidden goals and sabotage risks before deployment.

See something inaccurate or outdated?

Background

Current role
Member of Technical Staff, Manager Jan 2023-present
Previously
Nov 2019-Jan 2023Research Fellow,
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Last editorial review: August 10, 2026

Citation Trend

2 snapshots
Citation history from Aug 25, '25 to Jun 15, '265,7454,0252,305Aug 25, '25Jun 15, '26
5,745citations as of Jun 15, '26+3,440 since Aug 25, '25
For today’s citation count and citations per paper, visit Google Scholar.