Advantage-level Aggregation RL for X-point Divertor Control
August 24, 2026
Advantage Aggregation (AdvA) solves multi-objective reinforcement learning for X-point target divertor control in tokamak simulations. This method prevents reward scalarization from collapsing temporal credit, enabling precise control of secondary X-points in free-boundary environments calibrated to EXL-50U discharges.
HOW THIS AFFECTS YOU
●
researcherYou can apply AdvA to multi-objective RL problems where scalarized rewards fail to capture objective-specific temporal dynamics.