New papers target LLM hallucinations with hidden-state probes and critique

Seven arXiv papers posted Aug 6-12, 2026 propose hallucination-detection methods: hidden-state probes (PEP), agentic critique, and reflection-based abstention (REIN). One study finds linear probes catch corrupted context near-perfectly yet fail at failure prediction.
How this story unfolded
5 days · 7 reports · 1 community post · from Aug 7
- Aug 7
- Aug 11
- Aug 12
Actionable Hallucination Detection: Translating Latent Uncertainty into Agentic Critiquearxiv.org
What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the modelarxiv.org
Can Gemma and Qwen models catch hallucinations by looking at their own logprobs?
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Sequoia Capital invests in AI-native video platform Preview
- US Launches Effort to Speed Trade in AI Goods Between Allies
- DeepMind launches SL2T sign language-to-text model
- Liquid AI releases LFM2.5-VL-3B vision-language model for edge
- Grok and Meta's release discussed on ETN podcast episode