← Search

Shuang Hao

4 accepted papers

2026

A Temporal and Content Co-Awareness Latent Diffusion for Controllable Hand Image Generation

CVPR 2026

Controllable hand image generation aims to synthesize geometrically accurate images with consistent appearance. Recently, diffusion models have been widely applied for hand image synthesis. However, through input-level fusion or feature-level modulation, existing methods inject control signals with

Cited by 0SourcecodeScholar
2026

OwlCap: Harmonizing Motion-Detail for Video Captioning via HMD-270K and Caption Set Equivalence Reward

AAAI 2026technical

Video captioning aims to generate comprehensive and coherent descriptions of the video content, contributing to the advancement of both video understanding and generation. However, existing methods often suffer from motion-detail imbalance, as models tend to overemphasize one aspect while neglecting

Cited by 0SourcePDFScholar
2024

CoLA: Conditional Dropout and Language-driven Robust Dual-modal Salient Object Detection

ECCV 2024poster

"The depth/thermal information is beneficial for detecting salient object with conventional RGB images. However, in dual-modal salient object detection (SOD) model, the robustness against noisy inputs and modality missing is crucial but rarely studied. To tackle this problem, we introduce Conditiona…

2023

Development of a Lightweight Underwater Manipulator for Delicate Structural Repair Operations

RA-L 2023

In recent years, underwater robots have been increasingly used in the maintenance of hydraulic structures. Underwater manipulators are essential devices that are used to carry out such maintenance tasks. For delicate repair operations such as fixing tiny cracks, most existing underwater manipulators

Cited by 20SourceScholar