No Free Lunch from Random Feature Ensembles: Scaling Laws and Near-Optimality Conditions
Given a fixed budget for total model size, one must choose between training a single large model or combining the predictions of multiple smaller models. We investigate this trade-off for ensembles of random-feature ridge regression models in both the overparameterized and underparameterized regime…