← Search

Rafael Alberto Rivera Soto

2 accepted papers

2025

Mitigating Paraphrase Attacks on Machine-Text Detection via Paraphrase Inversion

ACL 2025finding

High-quality paraphrases are easy to produce using instruction-tuned language models or specialized paraphrasing models. Although this capability has a variety of benign applications, paraphrasing attacks—paraphrases applied to machine-generated texts—are known to significantly degrade the performan…

Cited by 0SourcePDFScholar
2024

Few-Shot Detection of Machine-Generated Text using Style Representations

ICLR 2024poster

The advent of instruction-tuned language models that convincingly mimic human writing poses a significant risk of abuse. For example, such models could be used for plagiarism, disinformation, spam, or phishing. However, such abuse may be counteracted with the ability to detect whether a piece of tex…