METR, or Model Evaluation and Threat Research, is a nonprofit research institute located in Berkeley, California. Founded in August 2022, METR focuses on evaluating advanced AI models to identify capabilities that may pose significant risks to society. The organization conducts pre-deployment empirical evaluations of AI systems, assessing dangerous capabilities such as autonomous replication and cybersecurity threats. METR's mission is to develop scientific methods for assessing the risks associated with AI systems' autonomous capabilities. It provides services to leading AI companies, including OpenAI, Anthropic, and Google DeepMind, helping them understand AI capabilities and risks before deploying new models. The organization also contributes to the development of standardized evaluation methodologies and publishes research to enhance public understanding of AI risks. With a dedicated team, METR aims to promote safe AI development and informed decision-making.
Enter your details and select criteria for the jobs you want to receive.
METR is seeking a remote Task Development Engineer to develop and evaluate challenging tasks for frontier AI models, supporting its Time Horizons research and broader assessments of AI capabilities and risks.