A new policy training method predicts the recoverability of a draft answer when using retrieved evidence. On 25,870 open-domain questions, this approach outperformed standard draft-correctness scorers by 0.23 to 0.68 accuracy points on average.
HOW THIS AFFECTS YOU
●
builderYou can implement more efficient RAG pipelines by deciding when to revise versus when to return existing answers.
●
researcherYou can evaluate the specific impact of revisions rather than just the correctness of the initial draft.