All people

Researcher

Mantas Mazeika

MM

Research Scientist at Center for AI Safety

Mazeika is a lead author of HarmBench (automated red-teaming evaluation) and the WMDP benchmark for measuring and unlearning hazardous knowledge, and co-authored work on tamper-resistant safeguards for open-weight LLMs. He is a research scientist at the Center for AI Safety.

See something inaccurate or outdated?

Background

Current role
Research Scientist
Previously
About this data

Work and education history is primarily focused on AI-relevant roles and may not be comprehensive.

Citation Trend

2 snapshots
Citation history from Aug 25, '25 to Jun 15, '2626,32821,35016,371Aug 25, '25Jun 15, '26
26,328citations as of Jun 15, '26+9,957 since Aug 25, '25
For today’s citation count and citations per paper, visit Google Scholar.