← Search

Fei Zuo

1 accepted papers

2025

InImageTrans: Multimodal LLM-based Text Image Machine Translation

ACL 2025finding

Multimodal large language models (MLLMs) have shown remarkable capabilities across various downstream tasks. However, when MLLMs are transferred to the text image machine translation (TiMT) task, preliminary experiments reveal that MLLMs suffer from serious repetition and omission hallucinations. To…