← Search

Zhengqi Xu

3 accepted papers

2024

Training-Free Pretrained Model Merging

CVPR 2024poster

Recently model merging techniques have surfaced as a solution to combine multiple single-talent models into a single multi-talent model. However previous endeavors in this field have either necessitated additional training or fine-tuning processes or require that the models possess the same pre-trai…

2023

Lookaround Optimizer: $k$ steps around, 1 step average

NeurIPS 2023poster

Weight Average (WA) is an active research topic due to its simplicity in ensembling deep networks and the effectiveness in promoting generalization. Existing weight average approaches, however, are often carried out along only one training trajectory in a post-hoc manner (i.e., the weights are avera…