← Search

Hyelin Nam

7 accepted papers

2026

Hybrid Semantic-Complementary Transmission for High-Fidelity Image Reconstruction

ICASSP 2026poster

Recent advances in semantic communication (SC) have introduced neural network (NN)-based transceivers that convey semantic representation (SR) of signals such as images. However, these NNs are trained over diverse image distributions and thus often fail to reconstruct fine-grained image-specific det…

Cited by 0SourcePDFScholar
2025

CFG++: Manifold-constrained Classifier Free Guidance for Diffusion Models

ICLR 2025poster

Classifier-free guidance (CFG) is a fundamental tool in modern diffusion models for text-guided generation. Although effective, CFG has notable drawbacks. For instance, DDIM with CFG lacks invertibility, complicating image editing; furthermore, high guidance scales, essential for high-quality output…

2025

Optical-Flow Guided Prompt Optimization for Coherent Video Generation

CVPR 2025poster

While text-to-video diffusion models have made significant strides, many still face challenges in generating videos with temporal consistency. Within diffusion frameworks, guidance techniques have proven effective in enhancing output quality during inference; however, applying these methods to video…

2025

SteerX: Creating Any Camera-Free 3D and 4D Scenes with Geometric Steering

ICCV 2025poster

Recent progress in 3D/4D scene generation emphasizes the importance of physical alignment throughout video generation and scene reconstruction. However, existing methods improve the alignment separately at each stage, making it difficult to manage subtle misalignments arising from another stage. Her…

2025

VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Joint Modeling

ICCV 2025poster

We propose VideoRFSplat, a direct text-to-3D model leveraging a video generation model to generate realistic 3D Gaussian Splatting (3DGS) for unbounded real-world scenes. To generate diverse camera poses and unbounded spatial extent of real-world scenes, while ensuring generalization to arbitrary te…

2024

Contrastive Denoising Score for Text-guided Latent Diffusion Image Editing

CVPR 2024poster

With the remarkable advent of text-to-image diffusion models image editing methods have become more diverse and continue to evolve. A promising recent approach in this realm is Delta Denoising Score (DDS) - an image editing technique based on Score Distillation Sampling (SDS) framework that leverage…

Cited by 25SourcePDFScholar
2024

Language-Oriented Communication with Semantic Coding and Knowledge Distillation for Text-to-Image Generation

ICASSP 2024accepted

By integrating recent advances in large language models (LLMs) and generative models into the emerging semantic communication (SC) paradigm, in this article we put forward to a novel framework of language-oriented semantic communication (LSC). In LSC, machines communicate using human language messag…

Cited by 0SourceScholar