Structural Conditions for Moral Competence in LLM Agents
September 7, 2026
LLM agents lack the prerequisites for coherent alignment due to failures in verdict stability, monotonicity, decisiveness, and Pareto viability. Testing nine frontier models shows they fail to maintain consistent policy mappings when morally relevant features are preserved or changed in simulated dilemmas.
HOW THIS AFFECTS YOU
●
researcherYou can use these four structural conditions as a non-normative framework to evaluate agentic consistency.
●
policyThis highlights a fundamental gap between model performance and predictable, rule-based behavior in safety alignment.