← Search

Tu Anh Dinh

3 accepted papers

2025

Are Generative Models Underconfident? Better Quality Estimation with Boosted Model Probability

EMNLP 2025

Quality Estimation (QE) is estimating quality of the model output during inference when the ground truth is not available. Deriving output quality from the models’ output probability is the most trivial and low-effort way. However, we show that the output probability of text-generation models can ap

Cited by 0SourcePDFScholar
2024

SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading

EMNLP 2024main

With the rapid development of Large Language Models (LLMs), it is crucial to have benchmarks which can evaluate the ability of LLMs on different domains. One common use of LLMs is performing tasks on scientific topics, such as writing algorithms, querying databases or giving mathematical proofs. Ins…

2022

Tackling Data Scarcity in Speech Translation Using Zero-Shot Multilingual Machine Translation Techniques

ICASSP 2022accepted

Recently, end-to-end speech translation (ST) has gained significant attention as it avoids error propagation. However, the approach suffers from data scarcity. It heavily depends on direct ST data and is less efficient in making use of speech transcription and text translation data, which is often m…

Cited by 0SourceScholar