← Search

Sangkyung Kwak

6 accepted papers

2026

DreamControl: Human-Inspired Whole-Body Humanoid Control for Scene Interaction Via Guided Diffusion

ICRA 2026poster

We introduce DreamControl, a novel methodology for learning autonomous whole-body humanoid skills. DreamControl leverages the strengths of diffusion models and Reinforcement Learning (RL): our core innovation is the use of a diffusion prior trained on human motion data, which subsequently guides an …

2025

Controllable Human Image Generation with Personalized Multi-Garments

CVPR 2025poster

We present BootControl, a novel framework based on text-to-image diffusion models for controllable human image generation with multiple reference garments.Here, the main bottleneck is data acquisition for training: collecting a large-scale dataset of high-quality reference garment images per human s…

Cited by 0SourcePDFScholar
2025

Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think

ICLR 2025oral

Recent studies have shown that the denoising process in (generative) diffusion models can induce meaningful (discriminative) representations inside the model, though the quality of these representations still lags behind those learned through recent self-supervised learning methods. We argue that on…

2025

StarFT: Robust Fine-tuning of Zero-shot Models via Spuriosity Alignment

IJCAI 2025

Learning robust representations from data often requires scale, which has led to the success of recent zero-shot models such as CLIP. However, the obtained robustness can easily be deteriorated when these models are fine-tuned on other downstream tasks (e.g., of smaller scales). Previous works often

2024

Direct Consistency Optimization for Robust Customization of Text-to-Image Diffusion models

NeurIPS 2024poster

Text-to-image (T2I) diffusion models, when fine-tuned on a few personal images, can generate visuals with a high degree of consistency. However, such fine-tuned models are not robust; they often fail to compose with concepts of pretrained model or other fine-tuned models. To address this, we propose…

Cited by 1SourcePDFScholar
2024

Improving Diffusion Models for Authentic Virtual Try-on in the Wild

ECCV 2024poster

"This paper considers image-based virtual try-on, which renders an image of a person wearing a curated garment, given a pair of images depicting the person and the garment, respectively. Previous works adapt existing exemplar-based inpainting diffusion models for virtual try-on to improve the natura…