“Posterity will find it ludicrous”: Sai agent hits 73% on OSWorld 2.0 performing routine (but necessary) work
The New Stack (AI)
Read full postSimular's Sai agent achieved a 73% success rate on the OSWorld 2.0 benchmark, outperforming GPT-5.6 Sol and Opus 5 while operating at about two-thirds the cost. Sai focuses on practical, routine professional tasks using a neurosymbolic approach to balance cost and performance.



