← Search

Jungwoo Kim

5 accepted papers

2026

Multi-view Pyramid Transformer: Look Coarser to See Broader

CVPR 2026

We propose Multi-view Pyramid Transformer (MVP), a scalable multi-view transformer architecture that directly reconstructs large 3D scenes from tens to hundreds of images in a single forward pass. Drawing on the idea of "looking broader to see the whole, looking finer to see the details," MVP is bui

Cited by 0SourcecodeScholar
2024

Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance

ECCV 2024poster

"Recent studies have demonstrated that diffusion models can generate high-quality samples, but their quality heavily depends on sampling guidance techniques, such as classifier guidance (CG) and classifier-free guidance (CFG). These techniques are often not applicable in unconditional generation or…