coder-eval 0.10.2

Source: Pypi.org· coder-eval@uipath.com· August 18, 2026

Evaluate, benchmark, and A/B-test AI coding agents (Claude Code, Codex, Gemini/Antigravity) with sandboxed, reproducible YAML task suites.

This story was reported by Pypi.org. Read the full original article:
Read on Pypi.org

More in AI Research

View all