arXiv AI By Zhongye Liu, Yaopei Zeng, Yurui Chang, Lu Lin

Forced Deferral: Manipulating Routing Decisions in Multimodal LLM Cascades

Read the original on arXiv AI →

arXiv:2606. 15308v1 Announce Type: new Abstract: While multimodal large language models (MLLMs) have shown strong visual reasoning abilities, serving a large model for every query is computationally expensive.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.