← Search

Byungjun Kim

13 accepted papers

2026

Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer

ICLR 2026poster

We present Durian, the first method for generating portrait animation videos with cross-identity attribute transfer from one or more reference images to a target portrait. Training such models typically requires attribute pairs of the same individual, which are rarely available at scale. To address…

Cited by 0SourcecodeScholar
2026

Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision

CVPR 2026

We present Vanast, a unified framework that generates garment-transferred human animation videos directly from a single human image, garment images, and a pose guidance video. Conventional two-stage pipelines treat image-based virtual try-on and pose-driven animation as separate processes, which oft

Cited by 0SourcecodeScholar
2025

HairCUP: Hair Compositional Universal Prior for 3D Gaussian Avatars

ICCV 2025poster

We present a universal prior model for 3D head avatars with explicit hair compositionality. Existing approaches to build generalizable priors for 3D head avatars often adopt a holistic modeling approach, treating the face and hair as an inseparable entity. This overlooks the inherent compositionalit…

Cited by 0SourcePDFScholar
2025

Leveraging Large Language Models for Active Merchant Non-player Characters

IJCAI 2025

We highlight two significant issues leading to the passivity of current merchant non-player characters (NPCs): pricing and communication. While immersive interactions with active NPCs have been a focus, price negotiations between merchant NPCs and players remain underexplored. First, passive pricing

2025

Real-time Adversarial Attack to Deep Learning-based Wi-Fi Human Activity Recognition

ICASSP 2025accepted

This study investigates adversarial attacks on deep learning (DL)-enabled Wi-Fi sensing systems using channel state information (CSI) for privacy. This paper presents a technique to disturb the signal used for channel estimation transmitted from the user device when the classifier is located at the…

Cited by 0SourceScholar
2024

GALA: Generating Animatable Layered Assets from a Single Scan

CVPR 2024poster

We present GALA a framework that takes as input a single-layer clothed 3D human mesh and decomposes it into complete multi-layered 3D assets. The outputs can then be combined with other assets to create novel clothed human avatars with any pose. Existing reconstruction approaches often treat clothed…

Cited by 7SourcePDFScholar
2023

Chupa: Carving 3D Clothed Humans from Skinned Shape Priors using 2D Diffusion Probabilistic Models

ICCV 2023oral

We propose a 3D generation pipeline that uses diffusion models to generate realistic human digital avatars. Due to the wide variety of human identities, poses, and stochastic details, the generation of 3D human meshes has been a challenging problem. To address this, we decompose the problem into 2D…

Cited by 25PDFcodeScholar
2022

SLiDE: Self-Supervised LiDAR De-Snowing through Reconstruction Difficulty

ECCV 2022poster

"LiDAR is widely used to capture accurate 3D outdoor scene structures. However, LiDAR produces many undesirable noise points in snowy weather, which hamper analyzing meaningful 3D scene structures. Semantic segmentation with snow labels would be a straightforward solution for removing them, but it r…

Cited by 17SourcePDFScholar