[arXiv]score: 0.24
Threat-guided Policy-aware Scene Perturbation for Safe Autonomous Driving with Online Reinforcement Learning
August 12, 2026
TPSP improves online reinforcement learning for autonomous driving by utilizing a policy-aware scene encoder to generate perturbations targeting specific policy weaknesses. This approach synchronizes scene generation with the evolving agent, addressing the long-tail distribution of safety-critical scenarios that conventional sampling methods often miss.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy