Why LLMs haven't automated more: RLHF rewards likable answers, not task completion

venturetwins · x · 2026-09-29

On the a16z podcast, @CompleteSkeptic argues LLMs are already smart enough to automate more human tasks — the bottleneck is that current models aren't built to function as reliable components inside software. RLHF rewards responses humans like, which doesn't necessarily lead to getting tasks done.

Original post →

More from AGI Musings

AGI Musings channel →