← Search

Soonmin Hwang

10 accepted papers

2025

Boosting Cross-Spectral Unsupervised Domain Adaptation for Thermal Semantic Segmentation

ICRA 2025

In autonomous driving, thermal image semantic segmentation has emerged as a critical research area, owing to its ability to provide robust scene understanding under adverse visual conditions. In particular, unsupervised domain adaptation (UDA) for thermal image segmentation can be an efficient solut

Cited by 0SourceScholar
2025

DEEPTalk: Dynamic Emotion Embedding for Probabilistic Speech-Driven 3D Face Animation

AAAI 2025technical

Speech-driven 3D facial animation has garnered lots of attention thanks to its broad range of applications. Despite recent advancements in achieving realistic lip motion, current methods fail to capture the nuanced emotional undertones conveyed through speech and produce monotonous facial motion. Th…

2025

MOSAIC: Generating Consistent, Privacy-Preserving Scenes from Multiple Depth Views in Multi-Room Environments

ICCV 2025poster

We introduce a diffusion-based approach for generating privacy-preserving digital twins of multi-room indoor environments from depth images only. Central to our approach is a novel Multi-view Overlapped Scene Alignment with Implicit Consistency (MOSAIC) model that explicitly considers cross-view dep…

Cited by 0SourcePDFScholar
2025

RoCaRS: Robust Camera-Radar BEV Segmentation for Sensor Failure Scenarios

IROS 2025

While camera–radar fusion has led to notable progress in autonomous driving, many existing approaches overlook the risk of sensor failures, which can critically compromise system safety. To address this limitation, we propose RoCaRS, a robust camera–radar fusion model designed for bird’s-eye view (B

Cited by 0SourceScholar
2023

T2FPV: Dataset and Method for Correcting First-Person View Errors in Pedestrian Trajectory Prediction

IROS 2023poster

Predicting pedestrian motion is essential for developing socially-aware robots that interact in a crowded environment. While the natural visual perspective for a social interaction setting is an egocentric view, the majority of existing work in trajectory prediction therein has been investigated pur…

Cited by 5SourcecodeScholar
2022

TransDSSL: Transformer Based Depth Estimation via Self-Supervised Learning

RA-L 2022

Recently, transformers have been widely adopted for various computer vision tasks and show promising results due to their ability to encode long-range spatial dependencies in an image effectively. However, very few studies on adopting transformers in self-supervised depth estimation have been conduc

Cited by 35SourcecodeScholar
2015

Multispectral Pedestrian Detection: Benchmark Dataset and Baseline

CVPR 2015poster

With the increasing interest in pedestrian detection, pedestrian datasets have also been the subject of research in the past decades. However, most existing datasets focus on a color channel, while a thermal channel is helpful for detection even in a dark environment. With this in mind, we propose a…

Cited by 1247SourcePDFScholar