●builderThis provides a testing framework for ensuring your multimodal agents obey complex, text-based operational constraints.
●researcherYou can use this benchmark to isolate and measure the intersection of visual perception and logical reasoning in MLLMs.