Chameleon: A Flexible Data-mixing Framework for Language Model Pretraining and Finetuning
Training data mixtures greatly impact the generalization performance of large language models. Existing domain reweighting methods often rely on costly weight computations and require retraining when new data is introduced. To this end, we introduce a flexible and efficient data mixing framework, Ch…