What is autonomous post-training?
The "autofinetune" project shared on the Google Developers Blog proposes a research loop that runs large language model (LLM) post-training without human intervention. It covers Supervised Fine-Tuning (SFT) and Reinforcement Learning via GRPO.
How it works
A developer defines boundary conditions and evaluation metrics in a single Markdown specification. An AI agent then takes over:
- Iteratively edits training scripts,
- Launches experiments,
- Automatically commits verified hyperparameter optimizations to Git.
This removes manual tuning cycles.

