ViSTA Adapts Vision-Language Models for Clinical Time-Series
September 28, 2026
ViSTA uses a compact, 0.516M parameter adapter to incorporate irregular clinical numerical measurements into pretrained vision-language models. The method achieves high AUC scores for acute kidney injury and mortality prediction on the MIMIC-IV dataset without changing pretrained weights.
HOW THIS AFFECTS YOU
●
builderYou can efficiently extend vision-language models to handle complex, irregular medical time-series data.
●
healthThis enables more accurate clinical risk prediction using multimodal LLM architectures.