The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality
An Chung Cheng, Alon Jacovi, Amir Globerson, Ben Golan et autres
We introduce The FACTS Leaderboard, an online leaderboard suite and associated set of benchmarks that comprehensively evaluates the ability of language models to generate factually accurate text across diverse scenarios. The suite provides a holistic measure of factuality by aggregating the performance …