Two-Phase Preference Elicitation for Linear Utility Policy Optimization
September 1, 2026
A two-phase algorithm optimizes policy by eliciting preferences directly over outcomes rather than decision-level comparisons. The first phase uses cutting planes to shrink attribute weight search spaces, followed by a second phase that provably converges to the user's utility function.
HOW THIS AFFECTS YOU
●
researcherYou can apply this two-phase approach to optimize complex resource allocation policies with high-dimensional utility functions.