Sai tops OSWorld 2.0 benchmark, beats GPT-5.6 and Opus 5 at lower cost

taoyds · x · 2026-08-28

Simular AI's computer-use agent Sai has achieved a 73% success rate on the OSWorld 2.0 benchmark, outperforming GPT-5.6 Sol and Claude Opus 5 at roughly two-thirds of the cost per task.

Key Highlights:

Related event: Simular's Sai Agent Tops OSWorld 2.0 at Two-Thirds the Cost(4 posts)→

Original post →

More from coding & agent

coding & agent channel →