← Search

Huiyuan Lai

11 accepted papers

2025

Multi-perspective Alignment for Increasing Naturalness in Neural Machine Translation

ACL 2025long

Neural machine translation (NMT) systems amplify lexical biases present in their training data, leading to artificially impoverished language in output translations. These language-level characteristics render automatic translations different from text originally written in a language and human tran…

Cited by 0SourcePDFScholar
2025

POMP: Pathology-omics Multimodal Pre-training Framework for Cancer Survival Prediction

IJCAI 2025

Cancer survival prediction is an important direction in precision medicine, aiming to help clinicians tailor treatment regimens for patients. With the rapid development of high-throughput sequencing and computational pathology technologies, survival prediction has shifted from clinical features to j

2024

Fine-tuning with HED-IT: The impact of human post-editing for dialogical language models

ACL 2024findings

Automatic methods for generating and gathering linguistic data have proven effective for fine-tuning Language Models (LMs) in languages less resourced than English. Still, while there has been emphasis on data quantity, less attention has been given to its quality. In this work, we investigate the i…

2024

mCoT: Multilingual Instruction Tuning for Reasoning Consistency in Language Models

ACL 2024long

Large language models (LLMs) with Chain-of-thought (CoT) have recently emerged as a powerful technique for eliciting reasoning to improve various downstream tasks. As most research mainly focuses on English, with few explorations in a multilingual context, the question of how reliable this reasoning…

2023

Pre-Trained Language-Meaning Models for Multilingual Parsing and Generation

ACL 2023findings

Pre-trained language models (PLMs) have achieved great success in NLP and have recently been used for tasks in computational semantics. However, these tasks do not fully benefit from PLMs since meaning representations are not explicitly included. We introduce multilingual pre-trained language-meanin…

2023

Responsibility Perspective Transfer for Italian Femicide News

ACL 2023findings

Different ways of linguistically expressing the same real-world event can lead to different perceptions of what happened. Previous work has shown that different descriptions of gender-based violence (GBV) influence the reader’s perception of who is to blame for the violence, possibly reinforcing ste…

2022

Multilingual Pre-training with Language and Task Adaptation for Multilingual Text Style Transfer

ACL 2022short

We exploit the pre-trained seq2seq model mBART for multilingual text style transfer. Using machine translated data as well as gold aligned English sentences yields state-of-the-art results in the three target languages we consider. Besides, in view of the general scarcity of parallel data, we propos…

2021

Generic resources are what you need: Style transfer tasks without task-specific parallel training data

EMNLP 2021main

Style transfer aims to rewrite a source text in a different target style while preserving its content. We propose a novel approach to this task that leverages generic resources, and without using any task-specific parallel (source–target) data outperforms existing unsupervised approaches on the two…

2021

Thank you BART! Rewarding Pre-Trained Models Improves Formality Style Transfer

ACL 2021short

Scarcity of parallel data causes formality style transfer models to have scarce success in preserving content. We show that fine-tuning pre-trained language (GPT-2) and sequence-to-sequence (BART) models boosts content preservation, and that this is possible even with limited amounts of parallel dat…