●builderYou can use this benchmark to test the fine-grained medical grounding of your vision-language models.
●researcherYou can study the gap between semantic understanding and pixel-level localization in medical VLMs.
●healthYou can assess whether medical AI provides sufficient visual evidence for its diagnostic claims.