← Search

Shaoxiong Ji

7 accepted papers

2025

How Many Languages Make Good Multilingual Instruction Tuning? A Case Study on BLOOM

COLING 2025main

Instruction tuning a large language model with multiple languages can prepare it for multilingual downstream tasks. Nonetheless, it is yet to be determined whether having a handful of languages is sufficient, or whether the benefits increase with the inclusion of more. By fine-tuning large multiling…

2024

A Comparison of Language Modeling and Translation as Multilingual Pretraining Objectives

EMNLP 2024main

Pretrained language models (PLMs) display impressive performances and have captured the attention of the NLP community.Establishing best practices in pretraining has, therefore, become a major focus of NLP research, especially since insights gained from monolingual English models may not necessarily…

2024

A New Massive Multilingual Dataset for High-Performance Language Technologies

COLING 2024main

We present the HPLT (High Performance Language Technologies) language resources, a new massive multilingual dataset including both monolingual and bilingual corpora extracted from CommonCrawl and previously unused web crawls from the Internet Archive. We describe our methods for data acquisition, ma…

2024

Can Machine Translation Bridge Multilingual Pretraining and Cross-lingual Transfer Learning?

COLING 2024main

Multilingual pretraining and fine-tuning have remarkably succeeded in various natural language processing tasks. Transferring representations from one language to another is especially crucial for cross-lingual learning. One can expect machine translation objectives to be well suited to fostering su…

Cited by 1SourcePDFScholar
2024

Knowledge-augmented Graph Neural Networks with Concept-aware Attention for Adverse Drug Event Detection

COLING 2024main

Adverse drug events (ADEs) are an important aspect of drug safety. Various texts such as biomedical literature, drug reviews, and user posts on social media and medical forums contain a wealth of information about ADEs. Recent studies have applied word embedding and deep learning-based natural langu…

Cited by 6SourcePDFScholar
2023

Towards Interpretable Mental Health Analysis with Large Language Models

EMNLP 2023long main

The latest large language models (LLMs) such as ChatGPT, exhibit strong capabilities in automated mental health analysis. However, existing relevant studies bear several limitations, including inadequate evaluations, lack of prompting strategies, and ignorance of exploring LLMs for explainability. T…

Cited by 0SourcecodeScholar