← Search

Zhenbo Yu

7 accepted papers

2025

LaTexBlend: Scaling Multi-concept Customized Generation with Latent Textual Blending

CVPR 2025highlight

Customized text-to-image generation renders user-specified concepts into novel contexts based on textual prompts. Scaling the number of concepts in customized generation meets a broader demand for user creation, whereas existing methods face challenges with generation quality and computational effic…

Cited by 1SourcePDFScholar
2023

Fast Fluid Simulation via Dynamic Multi-Scale Gridding

AAAI 2023technical

Recent works on learning-based frameworks for Lagrangian (i.e., particle-based) fluid simulation, though bypassing iterative pressure projection via efficient convolution operators, are still time-consuming due to excessive amount of particles. To address this challenge, we propose a dynamic multi-s…

Cited by 4SourcePDFScholar
2022

Object Wake-Up: 3D Object Rigging from a Single Image

ECCV 2022poster

"Given a single chair image, could we wake it up by reconstructing its 3D shape and skeleton, as well as animating its plausible articulations and motions, similar to that of human modeling? It is a new problem that not only goes beyond image-based object reconstruction but also involves articulated…

Cited by 7SourcePDFScholar
2021

Skeleton2Mesh: Kinematics Prior Injected Unsupervised Human Mesh Recovery

ICCV 2021poster

In this paper, we decouple unsupervised human mesh recovery into the well-studied problems of unsupervised 3D pose estimation, and human mesh recovery from estimated 3D skeletons, focusing on the latter task. The challenges of the latter task are two folds: (1) pose failure (i.e., pose mismatching -…

Cited by 29PDFcodeScholar
2021

Towards Alleviating the Modeling Ambiguity of Unsupervised Monocular 3D Human Pose Estimation

ICCV 2021poster

In this work, we study the ambiguity problem in the task of unsupervised 3D human pose estimation from 2D counterpart. On one hand, without explicit annotation, the scale of 3D pose is difficult to be accurately captured (scale ambiguity). On the other hand, one 2D pose might correspond to multiple…

Cited by 49PDFScholar
2020

Deep Kinematics Analysis for Monocular 3D Human Pose Estimation

CVPR 2020poster

For monocular 3D pose estimation conditioned on 2D detection, noisy/unreliable input is a key obstacle in this task. Simple structure constraints attempting to tackle this problem, e.g., symmetry loss and joint angle limit, could only provide marginal improvements and are commonly treated as auxilia…

Cited by 233PDFScholar