MediateRec Benchmark for Personal-Agent Recommendation Systems
October 7, 2026
The MediateRec benchmark evaluates personal LLM agents that mediate platform recommendations using cross-platform user history. It measures the ability of agents to balance platform-scale population evidence against private, user-authorized historical data.
HOW THIS AFFECTS YOU
●
builderYou can use this benchmark to test how personal agents override or refine third-party recommendation rankings.
●
founderThis identifies a shift toward user-governed personalization where personal agents mediate service interactions.