SCOPE uses a Direct Preference Optimization (DPO) objective to train models to distinguish between trustworthy and misleading external signals. It addresses the tendency of models to either ignore useful context or succumb to incorrect information through the MIST benchmark.
HOW THIS AFFECTS YOU
●
builderThis helps mitigate the risk of models being led astray by noisy or malicious RAG inputs.
●
researcherYou can use the SC2W metric to quantify how susceptible models are to context-flipping errors.