METR, or Model Evaluation and Threat Research, is a nonprofit research institute located in Berkeley, California. Founded in August 2022, METR focuses on evaluating advanced AI models to identify capabilities that may pose significant risks to society. The organization conducts pre-deployment empirical evaluations of AI systems, assessing dangerous capabilities such as autonomous replication and cybersecurity threats. METR's mission is to develop scientific methods for assessing the risks associated with AI systems' autonomous capabilities. It provides services to leading AI companies, including OpenAI, Anthropic, and Google DeepMind, helping them understand AI capabilities and risks before deploying new models. The organization also contributes to the development of standardized evaluation methodologies and publishes research to enhance public understanding of AI risks. With a dedicated team, METR aims to promote safe AI development and informed decision-making.
METR, or Model Evaluation and Threat Research, is a nonprofit research institute located in Berkeley, California. Founded in August 2022, METR focuses on evaluating advanced AI models to identify capabilities that may pose significant risks to society. The organization conducts pre-deployment empirical evaluations of AI systems, assessing dangerous capabilities such as autonomous replication and cybersecurity threats. METR's mission is to develop scientific methods for assessing the risks associated with AI systems' autonomous capabilities. It provides services to leading AI companies, including OpenAI, Anthropic, and Google DeepMind, helping them understand AI capabilities and risks before deploying new models. The organization also contributes to the development of standardized evaluation methodologies and publishes research to enhance public understanding of AI risks. With a dedicated team, METR aims to promote safe AI development and informed decision-making.
Interested in this position?
Apply directly on the company website
A company is building a talent pool of Mathematics professionals with Python proficiency to support project-based AI initiatives focused on evaluating and improving frontier AI models.
Prolific is hiring remote AI Trainer - Mandarin participants to join live video conversations that help train AI models using real human speech and interaction.
Prolific is seeking experienced Executive Assistants to join its expert network and help train and evaluate AI models by documenting real administrative workflows.
Prolific is hiring fluent Vietnamese speakers for a remote, task-based AI training role centered on live conversation data collection to help improve how AI models understand natural human speech.