← Search

Wenpeng Zhang

6 accepted papers

2022

Group-based Interleaved Pipeline Parallelism for Large-scale DNN Training

ICLR 2022poster

The recent trend of using large-scale deep neural networks (DNN) to boost performance has propelled the development of the parallel pipelining technique for efficient DNN training, which has resulted in the development of several prominent pipelines such as GPipe, PipeDream, and PipeDream-2BW. Howev…

2022

Imbalance-Aware Uplift Modeling for Observational Data

AAAI 2022technical

Uplift modeling aims to model the incremental impact of a treatment on an individual outcome, which has attracted great interests of researchers and practitioners from different communities. Existing uplift modeling methods rely on either the data collected from randomized controlled trials (RCTs) o…

Cited by 6SourcePDFScholar
2022

On the Convergence of Stochastic Multi-Objective Gradient Manipulation and Beyond

NeurIPS 2022accept

The conflicting gradients problem is one of the major bottlenecks for the effective training of machine learning models that deal with multiple objectives. To resolve this problem, various gradient manipulation techniques, such as PCGrad, MGDA, and CAGrad, have been developed, which directly alter t…

Cited by 53SourcePDFScholar
2017

Projection-free Distributed Online Learning in Networks

ICML 2017poster

The conditional gradient algorithm has regained a surge of research interest in recent years due to its high efficiency in handling large-scale machine learning problems. However, none of existing studies has explored it in the distributed online learning setting, where locally light computation is…

Cited by 91SourcePDFScholar