Automata-Based Reward Machines for Signal Temporal Logic
August 17, 2026
This method introduces an automata-based approach to provide an efficient memory mechanism for reinforcement learning under Signal Temporal Logic (STL) specifications. It avoids the intractable state space expansion typically caused by history-dependent STL robustness scores in long-horizon tasks.
HOW THIS AFFECTS YOU
●
researcherThis offers a more scalable way to implement formal temporal specifications in RL control policies.