Comparative Explainability Audit of DeBERTa-v3 for Medical Classification
October 2, 2026
This study evaluates five attribution methods—SHAP, LIME, occlusion, Input x Gradient, and Attention x Gradient—on DeBERTa-v3 for zero-shot medical abstract classification. Results show high predictive accuracy in clinical domains but highlight significant disagreement between different explainability techniques.
HOW THIS AFFECTS YOU
●
researcherUse these findings to evaluate the reliability of XAI methods in transformer-based medical models.
●
healthBe cautious when using automated explanations for clinical decisions due to low inter-method agreement.