← Search

Junyeong Park

5 accepted papers

2026

World in a Frame: Understanding Culture Mixing as a New Challenge for Vision-Language Models

CVPR 2026

In a globalized world, cultural elements from diverse origins frequently appear together within a single visual scene. We refer to these as culture mixing scenarios, yet how Large Vision-Language Models (LVLMs) perceive them remains underexplored. We investigate culture mixing as a critical challeng

Cited by 0SourceScholar
2025

Diffusion Models Through a Global Lens: Are They Culturally Inclusive?

ACL 2025long

Text-to-image diffusion models have recently enabled the creation of visually compelling, detailed images from textual prompts. However, their ability to accurately represent various cultural nuances remains an open question. In our work, we introduce CULTDIFF benchmark, evaluating whether state-of-…

2025

MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory

ICLR 2025poster

Significant advances have been made in developing general-purpose embodied AI in environments like Minecraft through the adoption of LLM-augmented hierarchical approaches. While these approaches, which combine high-level planners with low-level controllers, show promise, low-level controllers freque…

Cited by 2SourcePDFScholar
2023

Imagine the Unseen World: A Benchmark for Systematic Generalization in Visual World Models

NeurIPS 2023poster

Systematic compositionality, or the ability to adapt to novel situations by creating a mental model of the world using reusable pieces of knowledge, remains a significant challenge in machine learning. While there has been considerable progress in the language domain, efforts towards systematic visu…

Cited by 3SourcePDFScholar