← Search

Hoang-Quoc Nguyen-Son

4 accepted papers

2024

SimLLM: Detecting Sentences Generated by Large Language Models Using Similarity between the Generation and its Re-generation

EMNLP 2024main

Large language models have emerged as a significant phenomenon due to their ability to produce natural text across various applications. However, the proliferation of generated text raises concerns regarding its potential misuse in fraudulent activities such as academic dishonesty, spam disseminatio…

2023

VoteTRANS: Detecting Adversarial Text without Training by Voting on Hard Labels of Transformations

ACL 2023findings

Adversarial attacks reveal serious flaws in deep learning models. More dangerously, these attacks preserve the original meaning and escape human recognition. Existing methods for detecting these attacks need to be trained using original/adversarial data. In this paper, we propose detection without t…

2022

CheckHARD: Checking Hard Labels for Adversarial Text Detection, Prediction Correction, and Perturbed Word Suggestion

EMNLP 2022finding

An adversarial attack generates harmful text that fools a target model. More dangerously, this text is unrecognizable by humans. Existing work detects adversarial text and corrects a target’s prediction by identifying perturbed words and changing them into their synonyms, but many benign words are a…

2021

Machine Translated Text Detection Through Text Similarity with Round-Trip Translation

NAACL 2021long

Translated texts have been used for malicious purposes, i.e., plagiarism or fake reviews. Existing detectors have been built around a specific translator (e.g., Google) but fail to detect a translated text from a strange translator. If we use the same translator, the translated text is similar to it…