← Search

CHIA-JEN YEH

2 accepted papers

2024

CFEVER: A Chinese Fact Extraction and VERification Dataset

AAAI 2024technical

We present CFEVER, a Chinese dataset designed for Fact Extraction and VERification. CFEVER comprises 30,012 manually created claims based on content in Chinese Wikipedia. Each claim in CFEVER is labeled as “Supports”, “Refutes”, or “Not Enough Info” to depict its degree of factualness. Similar to th…

2023

Improving Multi-Criteria Chinese Word Segmentation through Learning Sentence Representation

EMNLP 2023short findings

Recent Chinese word segmentation (CWS) models have shown competitive performance with pre-trained language models' knowledge. However, these models tend to learn the segmentation knowledge through in-vocabulary words rather than understanding the meaning of the entire context. To address this issue,…

Cited by 0SourceScholar