Prometheus 2 is an open-source language model designed specifically to evaluate and assess the outputs generated by other language models. Serving as an accessible and transparent alternative to proprietary evaluator systems, it is trained to closely mirror human judgment and proprietary model evaluations across various assessment benchmarks. Unlike traditional evaluator models that rely strictly on fixed or generic metrics, Prometheus 2 supports both direct individual assessment and pairwise ranking of responses, while allowing users to apply custom, user-defined evaluation criteria and rubrics.