Back to @gemini
Gemini
Gemini
@gemini

FACTS Benchmark Suite: a new way to systematically evaluate LLMs factuality

The FACTS Benchmark Suite provides a systematic evaluation of Large Language Models (LLMs) factuality across three areas: Parametric, Search, and Multimodal reasoning.

Read on deepmind.google

11:29 AM · Dec 9, 2025

More from Gemini