Prompt optimization may not need search trees: NPO paper shows teacher quality beats elaborate search

rohanpaul_ai · x · 2026-09-02

A new paper tests NPO, which keeps a single prompt lineage and has a teacher model revise it from recent rollout traces and rewards, against GEPA's multi-candidate Pareto selection. NPO matched or beat GEPA on IFBench and HotpotQA with slightly fewer rollouts (3,500 vs 3,593; 6,800 vs 6,871). The gap widened with stronger teachers—DeepSeek-V4-Flash and especially GPT-5.5—suggesting strong teacher reasoning plus rich feedback can replace optimizer-side search. Optimized prompts also transferred to other student models, especially within the same family.

Related event: NPO Paper: Prompt Optimization Needs Good Teacher Feedback, Not Search Trees(2 posts)→

Original post →

More from coding & agent

coding & agent channel →