arXiv AI By Ruoxi Sun, Quantong Qiu, Juntao Li, Zecheng Tang, Yihang Lou, Min Zhang

Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads

Read the original on arXiv AI →

arXiv:2606. 05843v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) demonstrate remarkable proficiency on complex vision-language tasks, the mechanisms by which they extract query-relevant visual features from complex, noisy contexts remain opaque.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.