← Search

Suhyeon Lee

12 accepted papers

2025

Reangle-A-Video: 4D Video Generation as Video-to-Video Translation

ICCV 2025poster

We introduce Reangle-A-Video, a unified framework for generating synchronized multi-view videos from a single input video. Unlike mainstream approaches that train multi-view video diffusion models on large-scale 4D datasets, our method reframes the multi-view video generation task as video-to-videos…

2024

Decomposed Diffusion Sampler for Accelerating Large-Scale Inverse Problems

ICLR 2024poster

Krylov subspace, which is generated by multiplying a given vector by the matrix of a linear transformation and its successive powers, has been extensively studied in classical optimization literature to design algorithms that converge quickly for large linear inverse problems. For example, the conj…

2024

LLM-CXR: Instruction-Finetuned LLM for CXR Image Understanding and Generation

ICLR 2024poster

Following the impressive development of LLMs, vision-language alignment in LLMs is actively being researched to enable multimodal reasoning and visual input/output. This direction of research is particularly relevant to medical imaging because accurate medical image analysis and generation consist o…

2023

Improving 3D Imaging with Pre-Trained Perpendicular 2D Diffusion Models

ICCV 2023poster

Diffusion models have become a popular approach for image generation and reconstruction due to their numerous advantages. However, most diffusion-based inverse problem-solving methods only deal with 2D images, and even recently published 3D methods do not fully exploit the 3D distribution prior. To…

Cited by 48PDFcodeScholar
2023

Revisiting Self-Similarity: Structural Embedding for Image Retrieval

CVPR 2023poster

Despite advances in global image representation, existing image retrieval approaches rarely consider geometric structure during the global retrieval stage. In this work, we revisit the conventional self-similarity descriptor from a convolutional perspective, to encode both the visual and structural…

2023

SHUNIT: Style Harmonization for Unpaired Image-to-Image Translation

AAAI 2023technical

We propose a novel solution for unpaired image-to-image (I2I) translation. To translate complex images with a wide range of objects to a different domain, recent approaches often use the object annotations to perform per-class source-to-target style mapping. However, there remains a point for us to…

2022

WildNet: Learning Domain Generalized Semantic Segmentation From the Wild

CVPR 2022poster

We present a new domain generalized semantic segmentation network named WildNet, which learns domain-generalized features by leveraging a variety of contents and styles from the wild. In domain generalization, the low generalization ability for unseen target domains is clearly due to overfitting to…

Cited by 112PDFcodeScholar
2021

Hierarchical Memory Matching Network for Video Object Segmentation

ICCV 2021poster

We present Hierarchical Memory Matching Network (HMMN) for semi-supervised video object segmentation. Based on a recent memory-based method [33], we propose two advanced memory read modules that enable us to perform memory reading in multiple scales while exploiting temporal smoothness. We first pro…

Cited by 149PDFcodeScholar
2021

Unsupervised Domain Adaptation for Semantic Segmentation by Content Transfer

AAAI 2021technical

In this paper, we tackle the unsupervised domain adaptation (UDA) for semantic segmentation, which aims to segment the unlabeled real data using labeled synthetic data. The main problem of UDA for semantic segmentation relies on reducing the domain gap between the real image and synthetic image. To…

Cited by 52SourcePDFScholar