PetriBench Evaluates LLM Reasoning via Dynamic Petri Net State Spaces
September 18, 2026
PetriBench provides a scalable benchmark for evaluating reasoning over concurrent and distributed systems using Petri net formalisms. The framework tests models across four task families of increasing structural complexity, measuring how test-time compute interacts with different reasoning horizons.
HOW THIS AFFECTS YOU
●
researcherYou can use this to move beyond static knowledge benchmarks toward testing reasoning in dynamic environments.