← Search

Hong-Han Shuai*

2 accepted papers

2024

The Fabrication of Reality and Fantasy: Scene Generation with LLM-Assisted Prompt Interpretation

ECCV 2024poster

"In spite of recent advancements in text-to-image generation, limitations persist in handling complex and imaginative prompts due to the restricted diversity and complexity of training data. This work explores how diffusion models can generate images from prompts requiring artistic creativity or spe…

2024

TrajPrompt: Aligning Color Trajectory with Vision-Language Representations

ECCV 2024poster

"Cross-modal learning shows promising potential to overcome the limitations of single-modality tasks. However, without proper design for representation alignment between different data sources, the external modality cannot fully exhibit its value. For example, recent trajectory prediction approaches…