← Search

Biaoyan Fang

7 accepted papers

2025

Can VLMs Actually See and Read? A Survey on Modality Collapse in Vision-Language Models

ACL 2025finding

Vision-language models (VLMs) integrate textual and visual information, enabling the model to process visual inputs and leverage visual information to generate predictions. Such models are demanding for tasks such as visual question answering, image captioning, and visual grounding. However, some re…

Cited by 0SourcePDFScholar
2025

The More, The Better? A Critical Study of Multimodal Context in Radiology Report Summarization

EMNLP 2025

The Impression section of a radiology report summarizes critical findings of a radiology report and thus plays a crucial role in communication between radiologists and physicians. Research on radiology report summarization mostly focuses on generating the Impression section by summarizing informatio

Cited by 0SourcePDFScholar
2024

Born Differently Makes a Difference: Counterfactual Study of Bias in Biography Generation from a Data-to-Text Perspective

ACL 2024short

How do personal attributes affect biography generation? Addressing this question requires an identical pair of biographies where only the personal attributes of interest are different. However, it is rare in the real world. To address this, we propose a counterfactual methodology from a data-to-text…

Cited by 1SourcePDFScholar
2024

Understanding Faithfulness and Reasoning of Large Language Models on Plain Biomedical Summaries

EMNLP 2024finding

Generating plain biomedical summaries with Large Language Models (LLMs) can enhance the accessibility of biomedical knowledge to the public. However, how faithful the generated summaries are remains an open yet critical question. To address this, we propose FaReBio, a benchmark dataset with expert-a…

2023

More than Votes? Voting and Language based Partisanship in the US Supreme Court

EMNLP 2023short findings

Understanding the prevalence and dynamics of justice partisanship and ideology in the US Supreme Court is critical in studying jurisdiction. Most research quantifies partisanship based on voting behavior, and oral arguments in the courtroom --- the last essential procedure before the final case outc…

Cited by 0SourceScholar
2022

What does it take to bake a cake? The RecipeRef corpus and anaphora resolution in procedural text

ACL 2022findings

Procedural text contains rich anaphoric phenomena, yet has not received much attention in NLP. To fill this gap, we investigate the textual properties of two types of procedural text, recipes and chemical patents, and generalize an anaphora annotation framework developed for the chemical domain for…