Non-profit🇺🇸 United States
METR8 people
METR measures what frontier AI agents can do in realistic, extended tasks rather than short benchmark questions. Its widely followed time-horizon evaluations estimate the length of software and research tasks models can complete, giving labs and policymakers a concrete view of rapidly changing autonomous capability.
Non-profit🇬🇧 Oxford
Future of Humanity Institute6 people
Oxford’s Future of Humanity Institute helped make existential risk from advanced AI a serious academic and policy field. Led by Nick Bostrom until its closure in 2024, it connected technical AI safety with forecasting, governance and long-term questions about humanity’s future.
Non-profit🇺🇸 United States
Alignment Research Center4 people
Alignment Research Center develops methods for keeping powerful AI systems helpful and honest as their capabilities grow. Its work helped define scalable alignment and eliciting latent knowledge—ways to test whether models know more than they reveal—and its former evaluations team became METR.
Non-profit🇺🇸 United States
Palisade Research2 people
Palisade Research tests the offensive cyber capabilities and controllability of frontier AI systems. Its experiments examine behaviors such as shutdown resistance and autonomous replication, with the aim of helping institutions keep advanced systems under human control.
Non-profit🇺🇸 United States
Resolution2 people
Resolution researches how to understand and control advanced AI systems before their reasoning becomes too complex for people to follow. Founded by Geoffrey Irving and Daniel Murfet, it combines alignment research with mathematics and interpretability aimed at making model behavior more legible.
Non-profit🇺🇸 San Francisco
Transluce2 people
Builds open research and software for studying model behavior at scale. Its work includes automated discovery of unexpected behavior, interpretable concept analysis and tools for producing verifiable evaluations of AI systems.
Non-profit🇺🇸 San Francisco
PauseAI US1 person
Organizes a nationwide grassroots movement advocating for an international pause in frontier AI development, with local groups that meet lawmakers, hold public events, and engage the media on catastrophic AI risks.
Non-profit
Timaeus1 person
An AI-safety research organization that applied singular learning theory—mathematical tools for understanding complex learning systems—to model interpretability and alignment. Its work and team have merged into Resolution.