← Search

Sungmin Eum

9 accepted papers

2026

UAV4D: Dynamic Neural Rendering of Human-Centric UAV Imagery Using Gaussian Splatting

AAAI 2026technical

Despite significant advancements in dynamic neural rendering, existing methods fail to address the unique challenges posed by UAV-captured scenarios, particularly those involving monocular camera setups, top-down perspective, and multiple small, moving humans, which are not adequately represented in

Cited by 0SourcePDFScholar
2025

AutoComPose: Automatic Generation of Pose Transition Descriptions for Composed Pose Retrieval Using Multimodal LLMs

ICCV 2025poster

Composed pose retrieval (CPR) enables users to search for human poses by specifying a reference pose and a transition description, but progress in this field is hindered by the scarcity and inconsistency of annotated pose transitions. Existing CPR datasets rely on costly human annotations or heurist…

Cited by 0SourcePDFScholar
2024

Two Teachers Are Better Than One: Leveraging Depth In Training Only For Unsupervised Obstacle Segmentation

IROS 2024poster

We present a novel unsupervised obstacle segmentation architecture that follows a novel Relation Distillation (RD) paradigm. Our architecture design was inspired by a self-supervised teacher-student approach that relies on the Semantic Distillation originally devised for representation learning. Whi…

Cited by 0SourceScholar
2022

Negative Samples Are at Large: Leveraging Hard-Distance Elastic Loss for Re-identification

ECCV 2022poster

"We present a Momentum Re-identification (MoReID) framework that can leverage a very large number of negative samples in training for general re-identification task. The design of this framework is inspired by Momentum Contrast (MoCo), which uses a dictionary to store current and past batches to bui…

Cited by 9SourcePDFScholar
2020

S-DOD-CNN: Doubly Injecting Spatially-Preserved Object Information for Event Recognition

ICASSP 2020accepted

We present a novel event recognition approach called Spatially-preserved Doubly-injected Object Detection CNN (S-DOD-CNN), which incorporates the spatially preserved object detection information in both a direct and an indirect way. Indirect injection is carried out by simply sharing the weights bet…

Cited by 0SourceScholar
2019

A RUGD Dataset for Autonomous Navigation and Visual Perception in Unstructured Outdoor Environments

IROS 2019poster

Research in autonomous driving has benefited from a number of visual datasets collected from mobile platforms, leading to improved visual perception, greater scene understanding, and ultimately higher intelligence. However, this set of existing data collectively represents only highly structured, ur…

Cited by 208SourceScholar
2019

Object and Text-guided Semantics for CNN-based Activity Recognition

ICASSP 2019accepted

Many previous methods have demonstrated the importance of considering semantically relevant objects for carrying out video-based human activity recognition, yet none of the methods have harvested the power of large text corpora to relate the objects and the activities to be transferred into learning…

Cited by 0SourceScholar
2018

Exploitation of Semantic Keywords for Malicious Event Classification

ICASSP 2018accepted

Learning an event classifier is challenging when the scenes are semantically different but visually similar. However, as humans, we typically handle such tasks painlessly by adding our background semantic knowledge. Motivated by this observation, we aim to provide an empirical study about how additi…

Cited by 0SourceScholar