Google DeepMindAnalysisDevelopersSeptember 11, 2026

Google's autofinetune runs autonomous LLM post-training on TPUs

Google built autofinetune, an autonomous research loop that applies the autoresearch paradigm to LLM post-training (SFT and GRPO reinforcement learning) using Tunix, Gemma, and Cloud TPUs. An agent edits run.py, runs training, keeps winning commits or reverts regressions, and logs results to results.tsv.

1 source

More stories today

Open the live feed