An Alibaba open-source multi-language benchmark for evaluating LLMs in repository-level automatic code review, featuring an AI-assisted and expert-verified dataset.
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).