← Search

Seok Bong Yoo

16 accepted papers

2026

DeepProtect: Proactive Face-Swapping Defense using Identity Blending and Attribute Distortion

CVPR 2026

Face-swapping deepfakes allow realistic identity transfer, which can serve creative purposes but increases the risk of identity abuse. A proactive defense aims to prevent deepfake creation by obstructing identity feature extraction from input images, essential for identity-driven face-swapping. Exis

Cited by 0SourcecodeScholar
2026

Latent-RAG: Identity Retrieval-Guided Latent Augmentation for Privacy-Preserving Person Re-Identification

ICRA 2026poster

Person re-identification (re-ID) is crucial for security applications, including autonomous robots that monitor individuals via continuous image acquisition. Such data are transmitted to a database; however, if stored without adequate protection, they can be intercepted, posing privacy risks. In res…

Cited by 0codeScholar
2026

MonoEM: Object-Level Monocular 3D Object Detection Based on Equirectangular Map under Inclement Weather

ICRA 2026poster

Monocular 3D object detection has received growing recognition in contemporary research due to its reduced hardware complexity and lower deployment cost compared to multi-sensor-based approaches. Prior research has primarily addressed ideal environmental settings, neglecting the influence of diverse…

Cited by 0Scholar
2026

MonoKey: Keypoint-Based Monocular 3D Object Detection Using Prior Guidance for Occlusion Robustness

ICRA 2026poster

Monocular 3D object detection has gained attention for its cost-efficiency and simpler setup compared to multi-sensor systems. In this task, accurate depth estimation is crucial for precise object localization, yet extracting sufficient depth cues from a single image remains inherently challenging. …

Cited by 0Scholar
2026

MonoPure: Multi-Component Purification via Disentangled, Projective Representations for Monocular 3D Object Detection

IJCAI 2026

Monocular 3D object detection is a cost-efficient alternative to multisensor systems, yet it remains fragile to multi-component adversarial attacks that perturb the image and tamper with camera calibration. Compounded distortions degrade 3D reasoning by disrupting the correspondence between the 3D g

Cited by 0Scholar
2026

Robust, Generalizable Proactive Face-swapping Defense via Semantic Gradient Divergence

IJCAI 2026

The rapid progress of identity-feature-based face-swapping technology has raised concerns about impersonation and privacy violations. Although proactive defenses aim to block identity extraction at the source, existing methods suffer from perceptible visual artifacts, poor generalization across dive

Cited by 0Scholar
2025

Data Poisoning Attack Defense and Evolutionary Domain Adaptation for Federated Medical Image Segmentation

IJCAI 2025

Federated learning has significant demonstrated potential in medical image segmentation to protect data privacy by retaining local data. However, its application is still hindered by two critical challenges: 1) the retained data poisoning attacks that severely compromise the accuracy of the global s

2025

Ego-$A{\mathbf{3}}$: Adaptive Fusion-Based Disentangled Transformer for Egocentric Action Anticipation

ICRA 2025

Recently, egocentric action anticipation for wearable robotics cameras has gained considerable attention due to its capability to analyze nouns and verbs from a firstperson view. However, this field encounters challenges due to various uncertainties, such as action-irrelevant information and semanti

Cited by 0SourcecodeScholar
2025

Equirectangular Point Reconstruction for Domain Adaptive Multimodal 3D Object Detection in Adverse Weather Conditions

AAAI 2025technical

A multimodal fusion technique using LiDAR-camera has been developed for precise 3D object detection in autonomous driving and provides acceptable detection performance in ideal conditions with clear weather. However, the existing multimodal methods are still vulnerable to adverse weather conditions,…

2025

OPRNet: Object-Centric Point Reconstruction Network for Multimodal 3D Object Detection in Adverse Weathers

ICRA 2025

The development of a multimodal fusion technique utilizing LiDAR-camera data has enabled precise 3D object detection for self-driving vehicles, particularly in ideal conditions with clear weather. Nevertheless, adverse weathers such as fog, snow, and rain remain a challenge for existing multimodal m

Cited by 0SourcecodeScholar
2025

X-FLoRA: Cross-modal Federated Learning with Modality-expert LoRA for Medical VQA

EMNLP 2025

Medical visual question answering (VQA) and federated learning (FL) have emerged as vital approaches for enabling privacy-preserving, collaborative learning across clinical institutions. However, both these approaches face significant challenges in cross-modal FL scenarios, where each client possess

Cited by 0SourcePDFScholar
2024

Child FER: Domain-Agnostic Facial Expression Recognition in Children Using a Secondary Image Diffusion Model

ICASSP 2024accepted

Facial expression recognition (FER) models often face challenges when generalizing across domains, such as different datasets and age groups. Despite the significance of this problem, FER in children (child FER) research remains relatively understudied, and such studies exhibit vulnerability to cros…

Cited by 0SourceScholar
2024

IntensPure: Attack Intensity-aware Secondary Domain Adaptive Diffusion for Adversarial Purification

IJCAI 2024poster

Adversarial attacks pose a severe threat to the accuracy of person re-identification (re-ID) systems, a critical security technology. Adversarial purification methods are promising approaches for defending against comprehensive attacks, including unseen ones. However, re-ID testing identities (IDs)…

2024

Occluded Part-aware Graph Convolutional Networks for Skeleton-based Action Recognition

ICRA 2024poster

Recognizing human action is one of the most critical factors in the visual perception of robots. Specifically, skeletonbased action recognition has been actively researched to enhance recognition performance at a lower cost. However, action recognition in occlusion situations, where body parts are n…

Cited by 4SourcecodeScholar
2023

Fluxformer: Flow-Guided Duplex Attention Transformer via Spatio-Temporal Clustering for Action Recognition

RA-L 2023

Vision transformers have demonstrated impressive performance in various robotics and automation applications, such as classification automation and action recognition. However, the drawback of transformers is their quadratic increase in computing resources with larger inputs and dependence on consid

Cited by 10SourceScholar
2023

Latent-OFER: Detect, Mask, and Reconstruct with Latent Vectors for Occluded Facial Expression Recognition

ICCV 2023poster

Most research on facial expression recognition (FER) is conducted in highly controlled environments, but its performance is often unacceptable when applied to real-world situations. This is because when unexpected objects occlude the face, the FER network faces difficulties extracting facial feature…

Cited by 39PDFcodeScholar