ApprenticeBench: Agents Continually Learn Real Jobs, Surpassing Human Pros
ysu_nlp · x · 2026-09-12
NeoCognition introduced ApprenticeBench, a benchmark combining computer use with continual learning on real jobs. The team reports Fable 5.1 and GPT-6 Astra can continually learn on the job and surpass human professionals — a "decisive step change" in AI job readiness. Notably, no forward-deployed engineers are involved: agents deploy themselves into the role, testing the full zero-to-productive loop.
More from coding & agent
- Sentry CEO Dropped Claude Months Ago, Slams Overcomplicated Coding Harnesses After Hitting Codex Limits — zeeg · 2026-09-12
- LangChain explains agent harnesses in under 90 seconds — LangChain · 2026-09-12
- Sentry CEO: Smarter Models Now Overshoot Simple Tasks — I Want Fewer Mistakes, Not More Initiative — zeeg · 2026-09-12
- Fable 5.1 agent demo shows self-onboarding and lifelong learning at work — ysu_nlp · 2026-09-12
- Vyact: Open-Source Desktop App Builds Reviewable Workflows Around Local LLMs — vyact · 2026-09-12
- Runway launches MCP to generate images and videos inside ChatGPT, Claude and Cursor — runwayml · 2026-09-12