Illustration of a human difficulty meter pointing a different way from an agent cost meter

If It Feels Easy to a Human, That Still Won’t Predict the Agent’s Bill

Human time-to-fix ratings barely track agent token spend—and frontier models also systematically underestimate their own costs.