MMTClinic Benchmark for Multimodal Clinical Time-Series Reasoning
September 7, 2026
MMTClinic provides a multilingual benchmark for evaluating LLMs on clinical time-series data, combining text, medical images, and multivariate physiological signals. The dataset includes 30,000 QA pairs across English, Hindi, Bengali, Marathi, and Tamil.
HOW THIS AFFECTS YOU
●
researcherYou can use this to evaluate how well LLMs reason across different modalities and languages in a medical context.
●
healthThis provides a standardized way to measure the reliability of clinical AI in diverse linguistic settings.