← Search

Shuoyang Ding

8 accepted papers

2026

Test-Time Alignment for Large Language Models via Textual Model Predictive Control

ICLR 2026poster

Aligning Large Language Models (LLMs) with human preferences through finetuning is resource-intensive, motivating lightweight alternatives at test time. We address test-time alignment through the lens of sequential decision making, a perspective that reveals two fundamental challenges. When actions…

Cited by 0SourceScholar
2025

Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models

ACL 2025finding

Dense embeddings are fundamental to modern machine learning systems, powering Retrieval-Augmented Generation (RAG), information retrieval, and representation learning. While instruction-conditioning has become the dominant approach for embedding specialization, its direct application to low-capacity…

2025

EMMeTT: Efficient Multimodal Machine Translation Training

ICASSP 2025accepted

A rising interest in the modality extension of foundation language models warrants discussion on the most effective, and efficient, multimodal training approach. This work focuses on neural machine translation (NMT) and proposes a joint multimodal training regime of Speech-LLM to include automatic s…

Cited by 0SourceScholar
2025

Extending Automatic Machine Translation Evaluation to Book-Length Documents

EMNLP 2025

Despite Large Language Models (LLMs) demonstrating superior translation performance and long-context capabilities, evaluation methodologies remain constrained to sentence-level assessment due to dataset limitations, token number restrictions in metrics, and rigid sentence boundary requirements. We i

2024

Fine-Tuned Machine Translation Metrics Struggle in Unseen Domains

ACL 2024short

We introduce a new, extensive multidimensional quality metrics (MQM) annotated dataset covering 11 language pairs in the biomedical domain. We use this dataset to investigate whether machine translation (MT) metrics which are fine-tuned on human-generated MT quality judgements are robust to domain s…

2021

Levenshtein Training for Word-level Quality Estimation

EMNLP 2021main

We propose a novel scheme to use the Levenshtein Transformer to perform the task of word-level quality estimation. A Levenshtein Transformer is a natural fit for this task: trained to perform decoding in an iterative manner, a Levenshtein Transformer can learn to post-edit without explicit supervisi…

2019

Improving End-to-end Speech Recognition with Pronunciation-assisted Sub-word Modeling

ICASSP 2019accepted

Most end-to-end speech recognition systems model text directly as a sequence of characters or sub-words. Current approaches to sub-word extraction only consider character sequence frequencies, which at times produce inferior sub-word segmentation that might lead to erroneous speech recognition outpu…

Cited by 0SourceScholar