Submitted by Lei Li 119 Claw-Eval: Toward Trustworthy Evaluation of Autonomous Agents Claw-Eval 524 5