← Search

Xingyi Song

10 accepted papers

2026

ESG-Bench: Benchmarking Long-Context ESG Reports for Hallucination Mitigation

AAAI 2026technical

As corporate responsibility increasingly incorporates environmental, social, and governance (ESG) criteria, ESG reporting is becoming a legal requirement in many regions and a key channel for documenting sustainability practices and assessing firms’ long-term and ethical performance. However, the le

Cited by 0SourcePDFScholar
2025

Efficient Annotator Reliability Assessment and Sample Weighting for Knowledge-Based Misinformation Detection on Social Media

NAACL 2025findings

Misinformation spreads rapidly on social media, confusing the truth and targeting potentially vulnerable people. To effectively mitigate the negative impact of misinformation, it must first be accurately detected before applying a mitigation strategy, such as X’s community notes, which is currently…

2024

Confidence Regulation Neurons in Language Models

NeurIPS 2024poster

Despite their widespread use, the mechanisms by which large language models (LLMs) represent and regulate uncertainty in next-token predictions remain largely unexplored. This study investigates two critical components believed to influence this uncertainty: the recently discovered entropy neurons a…

2024

Enhancing Data Quality through Simple De-duplication: Navigating Responsible Computational Social Science Research

EMNLP 2024main

Research in natural language processing (NLP) for Computational Social Science (CSS) heavily relies on data from social media platforms. This data plays a crucial role in the development of models for analysing socio-linguistic phenomena within online communities. In this work, we conduct an in-dept…

2024

Examining Temporalities on Stance Detection towards COVID-19 Vaccination

COLING 2024main

Previous studies have highlighted the importance of vaccination as an effective strategy to control the transmission of the COVID-19 virus. It is crucial for policymakers to have a comprehensive understanding of the public’s stance towards vaccination on a large scale. However, attitudes towards COV…

Cited by 7SourcePDFScholar
2024

Examining the Limitations of Computational Rumor Detection Models Trained on Static Datasets

COLING 2024main

A crucial aspect of a rumor detection model is its ability to generalize, particularly its ability to detect emerging, previously unknown rumors. Past research has indicated that content-based (i.e., using solely source post as input) rumor detection models tend to perform less effectively on unseen…

2024

Identifying and Aligning Medical Claims Made on Social Media with Medical Evidence

COLING 2024main

Evidence-based medicine is the practise of making medical decisions that adhere to the latest, and best known evidence at that time. Currently, the best evidence is often found in the form of documents, such as randomized control trials, meta-analyses and systematic reviews. This research focuses on…

Cited by 2SourcePDFScholar
2024

Large Language Models Offer an Alternative to the Traditional Approach of Topic Modelling

COLING 2024main

Topic modelling, as a well-established unsupervised technique, has found extensive use in automatically detecting significant topics within a corpus of documents. However, classic topic modelling approaches (e.g., LDA) have certain drawbacks, such as the lack of semantic understanding and the presen…

2024

Navigating Prompt Complexity for Zero-Shot Classification: A Study of Large Language Models in Computational Social Science

COLING 2024main

Instruction-tuned Large Language Models (LLMs) have exhibited impressive language understanding and the capacity to generate responses that follow specific prompts. However, due to the computational demands associated with training these models, their applications often adopt a zero-shot setting. In…

Cited by 29SourcePDFScholar
2023

Don't waste a single annotation: improving single-label classifiers through soft labels

EMNLP 2023short findings

In this paper, we address the limitations of the common data annotation and training methods for objective single-label classification tasks. Typically, when annotating such tasks annotators are only asked to provide a single label for each sample and annotator disagreement is discarded when a final…

Cited by 0SourceScholar