← Search

Minji Bae

2 accepted papers

2026

Adaptive Capacity Allocation for Vision Language Action Fine-Tuning

ICRA 2026poster

Vision language action models (VLAs) are increasingly used for Physical AI, but deploying a pre-trained VLA model to unseen environments, embodiments, or tasks still requires adaptation. Parameter-efficient fine-tuning (PEFT), especially LoRA, is common for VLA policies, yet the exposed capacity kno…

2025

Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models

ICLR 2025poster

Text-to-image generative models like DALL-E and Stable Diffusion have revolutionized visual content creation across various applications, including advertising, personalized media, and design prototyping. However, crafting effective textual prompts to guide these models remains challenging, often re…