TOUR Benchmark for Trajectory-Level Unlearning in Offline RL
July 24, 2026
TOUR introduces a benchmark to evaluate trajectory-level data deletion in offline reinforcement learning. It utilizes matched non-member controls and multi-attack privacy auditing to distinguish between successful data removal, residual memorization, and policy collapse during D4RL and AntMaze experiments.
HOW THIS AFFECTS YOU
●
researcherYou can use this benchmark to evaluate if your unlearning methods actually remove data or simply break the agent's policy.