AnalysisAI ModelsAugust 20, 2026

AQuA's self-improvement updates research state, not agent LM

AQuA's preprint uses "recursive self-improvement" for a bounded research loop, but the research-agent LM and evaluator stay fixed within each part. The paper separates the LM from the research state it updates.

1 source

More stories today

Open the live feed
AQuA's self-improvement updates research state, not agent LM