Evaluate, benchmark, and A/B-test AI coding agents (Claude Code, Codex, Gemini/Antigravity) with sandboxed, reproducible YAML task suites.
$git clone https://github.com/uipath/coder_evalInstalls into the current project.
Install coder_eval by running `git clone https://github.com/uipath/coder_eval`, then use it for the current task and follow its documentation at https://github.com/uipath/coder_eval.