← Search

Boeun Kim

5 accepted papers

2025

3D Prior Is All You Need: Cross-Task Few-shot 2D Gaze Estimation

CVPR 2025poster

3D and 2D gaze estimation share the fundamental objective of capturing eye movements but are traditionally treated as two distinct research domains. In this paper, we introduce a novel cross-task few-shot 2D gaze estimation approach, aiming to adapt a pre-trained 3D gaze estimation network for 2D ga…

Cited by 0SourcePDFScholar
2025

High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation

ICCV 2025poster

Modeling high-resolution spatiotemporal representations, including both global dynamic contexts (e.g., holistic human motion tendencies) and local motion details (e.g., high-frequency changes of keypoints), is essential for video-based human pose estimation (VHPE). Current state-of-the-art methods t…

Cited by 0SourcePDFScholar
2025

PersonaBooth: Personalized Text-to-Motion Generation

CVPR 2025poster

This paper introduces Motion Personalization, a new task that generates personalized motions aligned with text descriptions using several basic motions containing Persona. To support this novel task, we introduce a new large-scale motion dataset called PerMo (PersonaMotion), which captures the uniqu…

Cited by 1SourcePDFScholar
2024

MoST: Motion Style Transformer Between Diverse Action Contents

CVPR 2024poster

While existing motion style transfer methods are effective between two motions with identical content their performance significantly diminishes when transferring style between motions with different contents. This challenge lies in the lack of clear separation between content and style of a motion.…

2022

Global-Local Motion Transformer for Unsupervised Skeleton-Based Action Learning

ECCV 2022poster

"We propose a new transformer model for the task of unsupervised learning of skeleton motion sequences. The existing transformer model utilized for unsupervised skeleton-based action learning is learned the instantaneous velocity of each joint from adjacent frames without global motion information.…