General Agents Won't Win, Specialized Harnesses Might Be Stronger

NirantK · x · 2026-07-13

This reply dives into the "agent wars," arguing that **code and programming languages hold the advantage in agentic workflows**. - Referencing Lilian Weng's harness evaluation, the author argues that many papers claiming "agents don't work" rely on older generation models lacking failure-detection capabilities. - They conclude that programming languages inherently provide better **deterministic context engineering**, meaning real-world effective systems rely more heavily on code-based orchestration than pure natural language agents. - They further predict that **general harnesses** (like Claude Code, Codex) might not necessarily win, and the future could belong to **specialized harnesses**, as "generality" introduces disadvantages in efficiency and control.

Related event: Specialized Coding Agents May Outperform General Agents(2 posts)→

Original post →

More from coding & agent

coding & agent channel →