← Search

Juil Koo

8 accepted papers

2026

BézierFlow: Learning Bézier Stochastic Interpolant Schedulers for Few-Step Generation

ICLR 2026poster

We introduce BézierFlow, a lightweight training approach for few-step generation with pretrained diffusion and flow models. BézierFlow achieves a 2–3× performance improvement for sampling with $\leq$ 10 NFEs while requiring only 15 minutes of training. Recent lightweight training approaches have sho…

Cited by 0SourcecodeScholar
2026

Token Warping Helps MLLMs Look from Nearby Viewpoints

CVPR 2026

Can warping tokens, rather than pixels, help multimodal large language models (MLLMs) understand how a scene appears from a nearby viewpoint? While MLLMs perform well on visual reasoning, they remain fragile to viewpoint changes, as pixel-wise warping is highly sensitive to small depth errors and of

Cited by 0SourcecodeScholar
2025

VideoHandles: Editing 3D Object Compositions in Videos Using Video Generative Priors

CVPR 2025poster

Generative methods for image and video editing use generative models as priors to perform edits despite incomplete information, such as changing the composition of 3D objects shown in a single image. Recent methods have shown promising composition editing results in the image setting, but in the vid…

Cited by 1SourcePDFScholar
2024

Neural Pose Representation Learning for Generating and Transferring Non-Rigid Object Poses

NeurIPS 2024poster

We propose a novel method for learning representations of poses for 3D deformable objects, which specializes in 1) disentangling pose information from the object's identity, 2) facilitating the learning of pose variations, and 3) transferring pose information to other object identities. Based on the…

Cited by 0SourcePDFScholar
2024

SyncTweedies: A General Generative Framework Based on Synchronized Diffusions

NeurIPS 2024poster

We introduce a general diffusion synchronization framework for generating diverse visual content, including ambiguous images, panorama images, 3D mesh textures, and 3D Gaussian splats textures, using a pretrained image diffusion model. We first present an analysis of various scenarios for synchroniz…

2023

SALAD: Part-Level Latent Diffusion for 3D Shape Generation and Manipulation

ICCV 2023poster

We present a cascaded diffusion model based on a part-level implicit 3D representation. Our model achieves state-of-the-art generation quality and also enables part-level shape editing and manipulation without any additional training in conditional setup. Diffusion models have demonstrated impressiv…

Cited by 49PDFScholar
2022

PartGlot: Learning Shape Part Segmentation From Language Reference Games

CVPR 2022oral

We introduce PartGlot, a neural framework and associated architectures for learning semantic part segmentation of 3D shape geometry, based solely on part referential language. We exploit the fact that linguistic descriptions of a shape can provide priors on the shape's parts -- as natural language h…

Cited by 33PDFcodeScholar