← Search

Yu-Wen Chen

4 accepted papers

2025

Read to Hear: A Zero-Shot Pronunciation Assessment Using Textual Descriptions and LLMs

EMNLP 2025

Automatic pronunciation assessment is typically performed by acoustic models trained on audio-score pairs. Although effective, these systems provide only numerical scores, without the information needed to help learners understand their errors. Meanwhile, large language models (LLMs) have proven eff

2022

Investigation of Factorized Optical Flows as Mid-Level Representations

IROS 2022poster

In this paper, we introduce a new concept of incorporating factorized flow maps as mid-level representations, for bridging the perception and the control modules in modular learning based robotic frameworks. To investigate the advantages of factorized flow maps and examine their interplay with the o…

Cited by 2SourceScholar
2022

S2F2: Single-Stage Flow Forecasting for Future Multiple Trajectories Prediction

ECCV 2022poster

"In this work, we present a single-stage framework, named S2F2, for forecasting multiple human trajectories from raw video images by predicting future optical flows. S2F2 differs from the previous two-stage approaches in that it performs detection, Re-ID, and forecasting of multiple pedestrians at t…

Cited by 4SourcePDFScholar