Research2026-04-28
PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging
Source: Arxiv CS.AI
arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementary cross-modal alignment capabilities. Model merging provides a cost-effective mechanism for...
arxivpapersmultimodal