Gemini
@gemini
FACTS Grounding: A new benchmark for evaluating the factuality of large language models
Our comprehensive benchmark and online leaderboard offer a much-needed measure of how accurately LLMs ground their responses in provided source material and avoid hallucinations
12:00 AM · Dec 17, 2024
Comments (0)
No comments yet.
Join the conversation on Mafold →