Benchmark framework that measures cost, quality, and duration of coding agents across any AI Coding Assistant CLI, any model, and any use case — with pluggable verification and real-repo support.
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).