Automated model post-training
Scaling Automated Post-Training
August 2026
Locus plans and steers many post-training experiments in parallel over multi-day horizons. It leads PostTrainBench, and with additional compute surpasses the official Qwen3-1.7B-Instruct checkpoint across the PostTrainBench+ suite. Locus has also post-trained a model deployed in production.
Selected results
- 44.7 PostTrainBench SOTA (verified)
- 51.6 on PostTrainBench+, above Qwen3-1.7B-Instruct
- 4th among accounts entered in all live prize-money Kaggle competitions