← Search

Jihye Park

5 accepted papers

2026

Text-Aware Image Restoration with Diffusion Models

ICLR 2026poster

While diffusion models have achieved remarkable success in natural image restoration, they often fail to faithfully recover textual regions, frequently producing plausible yet incorrect text-like patterns, a phenomenon we term text-image hallucination. To address this limitation, we propose Text-Awa…

Cited by 0SourcecodeScholar
2025

MoDiTalker: Motion-Disentangled Diffusion Model for High-Fidelity Talking Head Generation

AAAI 2025technical

Conventional GAN-based models for talking head generation often suffer from limited quality and unstable training. Recent approaches based on diffusion models have attempted to address these limitations and improve fidelity. However, they still face challenges, such as intensive sampling times and d…

2024

Hybrid Video Diffusion Models with 2D Triplane and 3D Wavelet Representation

ECCV 2024poster

"Generating high-quality videos that synthesize desired realistic content is a challenging task due to their intricate high dimensionality and complexity. Several recent diffusion-based methods have shown comparable performance by compressing videos to a lower-dimensional latent space, using traditi…

Cited by 10SourcePDFScholar
2023

LANIT: Language-Driven Image-to-Image Translation for Unlabeled Data

CVPR 2023poster

Existing techniques for image-to-image translation commonly have suffered from two critical problems: heavy reliance on per-sample domain annotation and/or inability to handle multiple attributes per image. Recent truly-unsupervised methods adopt clustering approaches to easily provide per-sample on…

2022

InstaFormer: Instance-Aware Image-to-Image Translation With Transformer

CVPR 2022poster

We present a novel Transformer-based network architecture for instance-aware image-to-image translation, dubbed InstaFormer, to effectively integrate global- and instance-level information. By considering extracted content features from an image as tokens, our networks discover global consensus of c…

Cited by 65PDFcodeScholar