Agents1 min read
Autonomous LLM post-training with Tunix on TPUs
The "autofinetune" project introduces an autonomous research loop that fully automates LLM post-training workflows, including Supervised Fine-Tuning (SFT) and Reinforcement Learning via GRPO. By defining boundary conditions and evaluatio...
From Google AI developers blog
