← Search

Hsin-Tai Wu

5 accepted papers

2025

Does RAG Introduce Unfairness in LLMs? Evaluating Fairness in Retrieval-Augmented Generation Systems

COLING 2025main

Retrieval-Augmented Generation (RAG) has recently gained significant attention for its enhanced ability to integrate external knowledge sources into open-domain question answering (QA) tasks. However, it remains unclear how these models address fairness concerns, particularly with respect to sensiti…

2025

Does Reasoning Introduce Bias? A Study of Social Bias Evaluation and Mitigation in LLM Reasoning

EMNLP 2025

Recent advances in large language models (LLMs) have enabled automatic generation of chain-of-thought (CoT) reasoning, leading to strong performance on tasks such as math and code. However, when reasoning steps reflect social stereotypes (e.g., those related to gender, race or age), they can reinfor

Cited by 0SourcePDFScholar
2025

Evaluating Fairness in Large Vision-Language Models Across Diverse Demographic Attributes and Prompts

EMNLP 2025

Large vision-language models (LVLMs) have recently achieved significant progress, demonstrating strong capabilities in open-world visual understanding. However, it is not yet clear how LVLMs address demographic biases in real life, especially the disparities across attributes such as gender, skin to

2024

Do Large Language Models Rank Fairly? An Empirical Study on the Fairness of LLMs as Rankers

NAACL 2024long

The integration of Large Language Models (LLMs) in information retrieval has raised a critical reevaluation of fairness in the text-ranking models. LLMs, such as GPT models and Llama2, have shown effectiveness in natural language understanding tasks, and prior works such as RankGPT have demonstrated…

Cited by 8SourcePDFScholar
2021

Karaoke Key Recommendation Via Personalized Competence-Based Rating Prediction

ICASSP 2021accepted

Karaoke machines have become a popular choice for many people’s daily entertainment. In this paper, we address a novel task of recommending a suitable key for a user to sing a given song to meet his or her vocal competence, by proposing the Personalized Competence-based Rating Prediction (PCRP) mode…

Cited by 0SourceScholar