← Search

Il Yong Chun

5 accepted papers

2026

SToRM: Supervised Token Reduction for Multi-Modal LLMs Toward Efficient End-To-End Autonomous Driving

ICRA 2026poster

In autonomous driving, end-to-end(E2E) driving systems that predict control commands directly from sensor data achieved significant advancements. For safe autonomous driving in unexpected scenarios, one may additionally rely on human interventions such as natural language instructions.Using a multi-…

2025

DX2CT: Diffusion Model for 3D CT Reconstruction from Bi or Mono-planar 2D X-ray(s)

ICASSP 2025accepted

Computational tomography (CT) provides high-resolution medical imaging, but it can expose patients to high radiation. X-ray scanners have low radiation exposure, but their resolutions are low. This paper proposes a new conditional diffusion model, DX2CT, that reconstructs three-dimensional (3D) CT v…

Cited by 0SourceScholar
2025

MAMS: Model-Agnostic Module Selection Framework for Video Captioning

AAAI 2025technical

Multi-modal transformers are rapidly gaining attention in video captioning tasks. Existing multi-modal video captioning methods extract a fixed number of frames, but this has critical challenges. If a limited number of frames are extracted, important frames with essential information for caption gen…

2020

Light-Field Reconstruction and Depth Estimation from Focal Stack Images Using Convolutional Neural Networks

ICASSP 2020accepted

Light-field (LF) reconstruction from focal stack images has diverse applications including face recognition, autonomous driving, and 3D reconstruction in virtual reality. It is a large-scale ill-conditioned inverse problem and typically requires regularized iterative algorithms to solve, which can b…

Cited by 0SourceScholar