← Search

Yuhe Jin

5 accepted papers

2024

ViVid-1-to-3: Novel View Synthesis with Video Diffusion Models

CVPR 2024highlight

Generating novel views of an object from a single image is a challenging task. It requires an understanding of the underlying 3D structure of the object from an image and rendering high-quality spatially consistent new views. While recent methods for view synthesis based on diffusion have shown grea…

Cited by 35SourcePDFScholar
2021

MIST: Multiple Instance Spatial Transformer

CVPR 2021poster

We propose a deep network that can be trained to tackle image reconstruction and classification problems that involve detection of multiple object instances, without any supervision regarding their whereabouts. The network learns to extract the most significant top-K patches, and feeds these patches…

Cited by 15PDFcodeScholar