All people

Researcher · Executive

Nora Belrose

NB

Head of Interpretability at EleutherAI

Belrose studies how language models represent concepts internally and whether those representations can be interpreted or controlled. She co-founded the independent research collective EleutherAI and led work on probing model internals, scalable oversight and open tools for alignment research.

See something inaccurate or outdated?

Background

Current role
Head of Interpretability
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Last editorial review: August 10, 2026

Citation Trend

2 snapshots
Citation history from Aug 25, '25 to Jun 15, '261,334999663Aug 25, '25Jun 15, '26
1,334citations as of Jun 15, '26+671 since Aug 25, '25
For today’s citation count and citations per paper, visit Google Scholar.