2018
Environment Upgrade Reinforcement Learning for Non-Differentiable Multi-Stage Pipelines
CVPR 2018poster
Recent advances in multi-stage algorithms have shown great promise, but two important problems still remain. First of all, at inference time, information can't feed back from downstream to upstream. Second, at training time, end-to-end training is not possible if the overall pipeline involves non-di…