← Search

Junsoo Oh

4 accepted papers

2025

From Linear to Nonlinear: Provable Weak-to-Strong Generalization through Feature Learning

NeurIPS 2025poster

Weak-to-strong generalization refers to the phenomenon where a stronger model trained under supervision from a weaker one can outperform its teacher. While prior studies aim to explain this effect, most theoretical insights are limited to abstract frameworks or linear/random feature models. In this…

Cited by 0SourceScholar
2024

DASH: Warm-Starting Neural Network Training in Stationary Settings without Loss of Plasticity

NeurIPS 2024poster

Warm-starting neural network training by initializing networks with previously learned weights is appealing, as practical neural networks are often deployed under a continuous influx of new data. However, it often leads to *loss of plasticity*, where the network loses its ability to learn new inform…

Cited by 2SourcePDFScholar