AQuA's self-improvement updates research state, not agent model weights

Post clarifies AQuA's preprint: the research-agent LM and evaluator stay fixed within each part of the bounded loop, so 'self-improvement' only updates the research state. A local port would freeze the fixed LM and evaluator objects.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Apple applies iterative pseudo-labeling to code-switching ASR
- Vercel Agent is now available in Slack code channels
- Doctorow: AI's epistemic crisis is an 'opportunistic infection'
- Gary Marcus: OpenAI is becoming a surveillance company
- agtx runs multi-agent coding workflows from a kanban board