← Search

Zhe Yu Liu

3 accepted papers

2020

GDN: A Coarse-To-Fine (C2F) Representation for End-To-End 6-DoF Grasp Detection

CoRL 2020

We proposed an end-to-end grasp detection network, Grasp Detection Network (GDN), cooperated with a novel coarse-to-fine (C2F) grasp representation design to detect diverse and accurate 6-DoF grasps based on point clouds. Compared to previous two-stage approaches which sample and evaluate multiple g

Cited by 0SourcePDFScholar
2020

Video Question Generation via Semantic Rich Cross-Modal Self-Attention Networks Learning

ICASSP 2020accepted

We introduce a novel task, Video Question Generation (Video QG). A Video QG model automatically generates questions given a video clip and its corresponding dialogues. Video QG requires a range of skills - sentence comprehension, temporal relation, the interplay between vision and language, and the…

Cited by 0SourceScholar
2019

Free-Form Video Inpainting With 3D Gated Convolution and Temporal PatchGAN

ICCV 2019poster

Free-form video inpainting is a very challenging task that could be widely used for video editing such as text removal. Existing patch-based methods could not handle non-repetitive structures such as faces, while directly applying image-based inpainting models to videos will result in temporal incon…

Cited by 270PDFcodeScholar