M-Loss: Quantifying Model Merging Compatibility with Limited Unlabeled Data
Training of large-scale models is both computationally intensive and often constrained by the availability of labeled data. Model merging offers a compelling alternative by directly integrating the weights of multiple source models without requiring additional data or extensive training. However, co