arXiv AI By Shuo Wang, Shunyang Huang, Jinghui Yuan, Zhixiang Shen, Zhao Kang

Cooperation of Experts: Fusing Heterogeneous Information with Large Margin

Read the original on arXiv AI →

arXiv:2505. 20853v3 Announce Type: replace-cross Abstract: Fusing heterogeneous information remains a persistent challenge in modern data analysis.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Computer Vision
2d ago

Harnessing Domain Specialists in Multimodal Mixture-of-Experts for Efficient Adaptation

The paper investigates whether the sparsity of Mixture-of-Experts (MoE) models leads to intrinsic semantic organization across modalities and domains. It shows that experts naturally specialize semantically even without explicit modular training. The authors propose ExpertLens, a data‑free method that decodes router weights to identify domain‑specialized experts, enabling selective fine‑tuning that matches or exceeds full fine‑tuning while updating only 21.7–47.0% of parameters and achieving a 4.0× speedup, outperforming LoRA in both performance and efficiency.

By Damiano Marsili, Raphi Kang, Aditya Mehta, Pietro Perona, Georgia Gkioxari
arXiv Machine Learning
1d ago

Coupling Perception and Reasoning in Federated Multimodal Graph Foundation Models

The paper introduces FedCORE, a federated adaptation framework for multimodal graph foundation models that jointly optimizes perception (Encoder) and reasoning (GNN) modules via a shared low‑dimensional latent state. Unlike prior methods that freeze the Encoder, FedCORE allows both components to adapt together, addressing the dependency between multimodal evidence extraction and graph‑based relational reasoning. Experiments show that FedCORE significantly narrows the Encoder–GNN pairing gap, achieving an 80.7% reduction compared to independent joint adaptation.

By Zekai Chen, Xun Wu, Hailin Zhang, Xunkai Li, Yu Liu, Kairui Yang, Muyan Huang, Xuaner Chen, Rong-Hua Li, Guoren Wang
arXiv AI
Sep 4

CauseCollab: Causal Unified and Modality-Agnostic Network for Heterogeneous Collaborative Perception

CauseCollab is a causal unified and modality‑agnostic network designed to improve collaborative perception across heterogeneous sensor modalities. It disentangles semantic factors from modality‑specific confounders using causal metric learning and employs a context‑guided Unified Converter to maintain cross‑modal semantic consistency. The approach requires only minimal adapter training when adding new modalities and achieves state‑of‑the‑art results on the OPV2V and DAIR‑V2X datasets, especially in scenarios with large modality gaps.

By Weize Li, Yang Li, Quan Yuan, Xiaoyuan Fu, Guiyang Luo, Jinglin Li