Locus tops PostTrainBench, saying automated post-training beats human-tuned Qwen3

dair_ai · x · 2026-08-04

Locus, an automated research system from Intology, reports a new SOTA on PostTrainBench. In the quoted results, it post-trains Qwen3 base models that outperform the human-post-trained Qwen3 1.7B Instruct release.

Key points:

The broader claim is that automated systems can now meaningfully improve models during post-training, not just assist with scaffolding around them.

Related event: Automated System Locus Sets New Post-Training Records(4 posts)→

Original post →

More from Models

Models channel →