Identifying Assistant Bias in User Simulators via Role Vectors
September 2, 2026
LLM user simulators suffer from assistant bias, where models overly cooperate instead of mimicking real user frustration. This research identifies a specific user role vector within model activations that can be used to steer simulations toward more realistic, non-cooperative behaviors.
HOW THIS AFFECTS YOU
●
builderYou can improve the validity of your agent evaluations by steering simulators away from cooperative bias.
●
researcherYou can use activation-based role vectors to better understand and mitigate inherent model biases.