[arXiv]score: 0.24
Direct Optimization of Generators for Search in Automated Theorem Proving
September 23, 2026
Compute-Aligned Training (CAT) extends policy-guided search alignment to automated theorem proving by deriving tractable, trace-supported losses that account for off-trace states. The method introduces a search-agnostic uniform-allocation loss to manage computational budgets, applying scalar weights to per-tactic cross-entropy to optimize LLMs for tree search rather than single-attempt generation.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy