
Sandipan KunduResearcher
Anthropic
Member of Technical Staff
via Anthropic
Researcher · Executive

Member of Technical Staff, Manager at Anthropic
Hubinger helped define the problem of inner alignment: whether a trained model develops its own internal objective rather than the one intended by its designers. At Anthropic he leads alignment stress-testing, probing models for deceptive behavior, hidden goals and sabotage risks before deployment.
See something inaccurate or outdated?
Anthropic
Machine Intelligence Research InstituteWork and education history is primarily focused on AI-relevant roles and may not be comprehensive.
Last editorial review: August 10, 2026