MMDiff: a method to find and control visual features inside multimodal large language models | arXiv News