
Dan HendrycksResearcher
Center for AI Safety
Director
via Center for AI Safety
Researcher
Research Scientist at Center for AI Safety
Mazeika is a lead author of HarmBench (automated red-teaming evaluation) and the WMDP benchmark for measuring and unlearning hazardous knowledge, and co-authored work on tamper-resistant safeguards for open-weight LLMs. He is a research scientist at the Center for AI Safety.
See something inaccurate or outdated?
Center for AI SafetyWork and education history is primarily focused on AI-relevant roles and may not be comprehensive.