NobleBlocks
Public

CodeScore: Evaluating Code Generation by Learning Code Execution

Published in ACM Transactions on Software Engineering and Methodology • Feb 23, 2025
Authors:
Yihong Dong
,
Jiazheng Ding
,
Xue Jiang

Abstract

A proper code evaluation metric (CEM) profoundly impacts the evolution of code generation, which is an important research field in NLP and software engineering. Prevailing match-based CEMs (e.g., BLEU, Accuracy, and CodeBLEU) suffer from two significant drawbacks. 1. They primarily measure the surfa...

Finding related papers...

Discussions

(0)

No comments yet

Be the first to share your thoughts!