The alignment illusion in multimodal large language models: when internal similarity hides lost visual content | arXiv News