← Search

DOHA HWANG

1 accepted papers

2023

Can We Scale Transformers to Predict Parameters of Diverse ImageNet Models?

ICML 2023poster

Pretraining a neural network on a large dataset is becoming a cornerstone in machine learning that is within the reach of only a few communities with large-resources. We aim at an ambitious goal of democratizing pretraining. Towards that goal, we train and release a single neural network that can pr…