Runta finds 17x cost gap across nine AI agent harnesses testing same model
Runta ran an evaluation called FrontierHarness Eval, testing nine different AI agent harnesses against the same underlying model to compare their performance and efficiency. The results showed the cost per successful task completion varied by as much as 17 times depending on which harness was used, despite the model being identical. Runta is now offering $100 in credits to developers who want to test their own harness on its platform.