← Search

Jungho Kim

9 accepted papers

2026

STONE Dataset: A Scalable Multi-Modal Surround-View 3D Traversability Dataset for Off-Road Robot Navigation

ICRA 2026poster

Reliable off-road navigation requires accurate estimation of traversable regions and robust perception under diverse terrain and sensing conditions. However, existing datasets lack both scalability and multi-modality, which limits progress in 3D traversability prediction. In this work, we introduce …

2026

SafeDrive: Fine-Grained Safety Reasoning for End-to-End Driving in a Sparse World

CVPR 2026

The end-to-end (E2E) paradigm, which maps sensor inputs directly to driving decisions, has recently attracted significant attention due to its unified modeling capability and scalability. However, ensuring safety in this unified framework remains one of the most critical challenges. In this work, we

Cited by 4SourceScholar
2025

PersonaBooth: Personalized Text-to-Motion Generation

CVPR 2025poster

This paper introduces Motion Personalization, a new task that generates personalized motions aligned with text descriptions using several basic motions containing Persona. To support this novel task, we introduce a new large-scale motion dataset called PerMo (PersonaMotion), which captures the uniqu…

Cited by 1SourcePDFScholar
2025

ProtoOcc: Accurate, Efficient 3D Occupancy Prediction Using Dual Branch Encoder-Prototype Query Decoder

AAAI 2025technical

In this paper, we introduce ProtoOcc, a novel 3D occupancy prediction model designed to predict the occupancy states and semantic classes of 3D voxels via a deep semantic understanding of scenes. ProtoOcc consists of two main components: the Dual Branch Encoder (DBE) and the Prototype Query Decoder…

2024

Mask2Map: Vectorized HD Map Construction Using Bird's Eye View Segmentation Masks

ECCV 2024oral

"In this paper, we introduce Mask2Map, a novel end-to-end online HD map construction method designed for autonomous driving applications. Our approach focuses on predicting the class and ordered point set of map instances within a scene, represented in the bird’s eye view (BEV). Mask2Map consists of…

2024

MoST: Motion Style Transformer Between Diverse Action Contents

CVPR 2024poster

While existing motion style transfer methods are effective between two motions with identical content their performance significantly diminishes when transferring style between motions with different contents. This challenge lies in the lack of clear separation between content and style of a motion.…

2023

R-Pred: Two-Stage Motion Prediction Via Tube-Query Attention-Based Trajectory Refinement

ICCV 2023poster

Predicting the future motion of dynamic agents is of paramount importance to ensuring safety and assessing risks in motion planning for autonomous robots. In this study, we propose a two-stage motion prediction method, called R-Pred, designed to effectively utilize both scene and interaction context…

Cited by 21PDFScholar
2023

SiT Dataset: Socially Interactive Pedestrian Trajectory Dataset for Social Navigation Robots

NeurIPS 2023poster

To ensure secure and dependable mobility in environments shared by humans and robots, social navigation robots should possess the capability to accurately perceive and predict the trajectories of nearby pedestrians. In this paper, we present a novel dataset of pedestrian trajectories, referred to as…

2022

Global-Local Motion Transformer for Unsupervised Skeleton-Based Action Learning

ECCV 2022poster

"We propose a new transformer model for the task of unsupervised learning of skeleton motion sequences. The existing transformer model utilized for unsupervised skeleton-based action learning is learned the instantaneous velocity of each joint from adjacent frames without global motion information.…