●researcherThe study shows that narrative wrappers move harmful request representations significantly farther from refusal directions in latent space.
●policyYou should account for linguistic registers and narrative context when evaluating model safety and alignment robustness.