← Search

Yoonwoo Jeong

6 accepted papers

2026

Vision-aligned Latent Reasoning for Multi-Modal Large Language Model

ICML 2026poster

Despite recent advancements in Multi-modal Large Language Models (MLLMs) on diverse understanding tasks, these models struggle to solve problems which require extensive multi-step reasoning. This is primarily due to the progressive dilution of visual information during long-context generation, which…

Cited by 0SourceScholar
2024

NVS-Adapter: Plug-and-Play Novel View Synthesis from a Single Image

ECCV 2024poster

"Recent advancements in Novel View Synthesis (NVS) from a single image have produced impressive results by leveraging the generation capabilities of pre-trained Text-to-Image (T2I) models. However, previous NVS approaches require extra optimization to use other plug-and-play image generation modules…

2023

Stable and Consistent Prediction of 3D Characteristic Orientation via Invariant Residual Learning

ICML 2023poster

Learning to predict reliable characteristic orientations of 3D point clouds is an important yet challenging problem, as different point clouds of the same class may have largely varying appearances. In this work, we introduce a novel method to decouple the shape geometry and semantics of the input p…

Cited by 3SourcePDFScholar
2022

PeRFception: Perception using Radiance Fields

NeurIPS 2022accept

The recent progress in implicit 3D representation, i.e., Neural Radiance Fields (NeRFs), has made accurate and photorealistic 3D reconstruction possible in a differentiable manner. This new representation can effectively convey the information of hundreds of high-resolution images in one compact for…

2021

Self-Calibrating Neural Radiance Fields

ICCV 2021poster

In this work, we propose a camera self-calibration algorithm for generic cameras with arbitrary non-linear distortions. We jointly learn the geometry of the scene and the accurate camera parameters without any calibration objects. Our camera model consists of a pinhole model, a fourth order radial d…

Cited by 268PDFcodeScholar