Skip to content
Advertisement

METR (Model Evaluation & Threat Research) is a research nonprofit that evaluates the autonomous capabilities and safety risks of frontier AI models, including time-horizon benchmarks for agentic task completion, and conducts both partnered evaluations with AI developers such as Anthropic, OpenAI, and Google DeepMind and independent assessments of publicly released models.

Text
Advertisement