RM-EVAL Reference-Free Reward Model for Grammatical Error Correction
September 21, 2026
RM-EVAL is a reward model trained on human preference data to evaluate Grammatical Error Correction (GEC) without gold references. It functions as both a meta-evaluator and a learning signal for Reward-Guided Text Generation (RGTG) to improve model decoding.
HOW THIS AFFECTS YOU
●
builderYou can use reward-driven decoding to improve GEC output quality without changing the base model.
●
researcherYou can move beyond rigid reference-based metrics like M2 to human-aligned evaluation.