MLS-Neurons Enable Cross-Dimensional Safety Alignment in LVLMs
July 31, 2026
The MLS-Neurons framework identifies shared safety neurons across language and visual modalities to defend against compound attacks in large vision-language models. It uses functional saliency and activation strength to extract neurons responsive to both textual and visual risks.
HOW THIS AFFECTS YOU
●
researcherYou can use neuron-level alignment to reduce the high fine-tuning costs of multimodal safety training.
●
policyThis improves the robustness of vision-language models against sophisticated, multi-modal malicious inputs.