deepmind.google6 months agoFACTS Benchmark Suite: a new way to systematically evaluate LLMs factualityThe FACTS Benchmark Suite provides a systematic evaluation of Large Language Models (LLMs) factuality across three areas: Parametric, Search, and Multimodal reasoning.Visit deepmind.google1BookmarkAdd to collection