Lab built its AI model with agents: humans took over just 0.7% of stuck cases

alex_verem · x · 2026-09-20

The team behind the open model Atria Dawn published a self-study of their own build process, analyzing 769 task records from 56 people plus agent logs:

The authors warn that each human decision now depends on more AI work than one person can verify — people risk becoming reviewers who can only say yes — and admit many researchers ran agents with permission checks off because approvals slowed work down. The model topped 5 of 16 reported benchmarks, but the build-process record may be the more valuable half of the paper.

Original post →

More from AGI Musings

AGI Musings channel →