← Search

Binkai Ou

2 accepted papers

2025

Cascaded Diffusion Models for Virtual Try-On: Improving Control and Resolution

AAAI 2025technical

Previous virtual try-on methods have employed ControlNet architecture in exemplar-based inpainting diffusion models to guide the generation of try-on images, preserving the garment's features and enhancing the realism of the generated images. While these methods have maintained the identity of the g…

Cited by 0SourcePDFScholar
2025

Observation-Graph Interaction and Key-Detail Guidance for Vision and Language Navigation

IROS 2025

Vision and Language Navigation (VLN) requires an agent to navigate through environments following natural language instructions. However, existing methods often struggle with effectively integrating visual observations and instruction details during navigation, leading to suboptimal path planning an

Cited by 2SourceScholar