Google DeepMindAnalysisAI ModelsAugust 12, 2026

Google Research: Recall is LLMs' factuality bottleneck

Google's knowledge-profiling framework finds frontier LLMs (Gemini 3, GPT-5) encode nearly all facts but fail to recall many — factual errors are recall failures, not knowledge gaps. The WikiProfile benchmark covers 2,150 Wikipedia-derived facts, each probed by ten questions.

1 source

Google DeepMind by email

Get an email when Google DeepMind has news

No news that day, no email.

More stories today

Open the live feed