← Search

Hengyu Li

3 accepted papers

2025

Boosting Text-To-Image Generation via Multilingual Prompting in Large Multimodal Models

ICASSP 2025accepted

Previous work on augmenting large multimodal models (LMMs) for text-to-image (T2I) generation has focused on enriching the input space of in-context learning (ICL). This includes providing a few demonstrations and optimizing image descriptions to be more detailed and logical. However, as demand for…

Cited by 0SourceScholar
2024

Considering Temporal Connection between Turns for Conversational Speech Synthesis

ICASSP 2024accepted

Conversational speech synthesis aims to synthesize speech of an individual speaker based on history conversation. However, most studies in conversational speech synthesis only focus on the synthesis performance of the current speaker’s turn and neglect the temporal relationship between turns of inte…

Cited by 0SourceScholar
2016

Studying of rectilinear locomotion for a two-segment system with anisotropic dry friction model

IROS 2016poster

This paper contributes to the understanding of the fundamental properties of rectilinearly locomotion of an one-dimensional system travelling on the horizontal plane, where dry Coulomb friction acting between it and surface. We discuss an approximate steady-state motion on a simplified two-segment s…

Cited by 0SourceScholar