FACTS Benchmark Suite: Systematically evaluating the factuality of large language models
FACTS Grounding: A new benchmark for evaluating the factuality of large language modelsDecember 2024Responsibility & Safety Learn more
🔗 Read full article on DeepMind →