2023
ViSoBERT: A Pre-Trained Language Model for Vietnamese Social Media Text Processing
EMNLP 2023long main
English and Chinese, known as resource-rich languages, have witnessed the strong development of transformer-based language models for natural language processing tasks. Although Vietnam has approximately 100M people speaking Vietnamese, several pre-trained models, e.g., PhoBERT, ViBERT, and vELECTRA…