CodeScore: Evaluating Code Generation by Learning Code Execution
Published in ACM Transactions on Software Engineering and Methodology • Feb 23, 2025
Authors:,,
Yihong Dong
Jiazheng Ding
Xue Jiang
Abstract
A proper code evaluation metric (CEM) profoundly impacts the evolution of code generation, which is an important research field in NLP and software engineering. Prevailing match-based CEMs (e.g., BLEU, Accuracy, and CodeBLEU) suffer from two significant drawbacks. 1. They primarily measure the surfa...
Finding related papers...
Discussions
(0)No comments yet
Be the first to share your thoughts!