← Search

Zehua Cheng

4 accepted papers

2025

On Weaponization-Resistant Large Language Models with Prospect Theoretic Alignment

COLING 2025main

Large language models (LLMs) have made significant advancements, but their increasing capabilities present serious risks of misuse, particularly in open-weight models where direct access to the model’s parameters is possible. Current safeguards, designed for closed-weight API models, are inadequate…

Cited by 1SourcePDFScholar
2024

Does DetectGPT Fully Utilize Perturbation? Bridging Selective Perturbation to Fine-tuned Contrastive Learning Detector would be Better

ACL 2024long

The burgeoning generative capabilities of large language models (LLMs) have raised growing concerns about abuse, demanding automatic machine-generated text detectors. DetectGPT, a zero-shot metric-based detector, first introduces perturbation and shows great performance improvement. However, in Dete…

2019

Long Text Analysis Using Sliced Recurrent Neural Networks with Breaking Point Information Enrichment

ICASSP 2019accepted

Sliced recurrent neural networks (SRNNs) are the state-of-the-art efficient solution for long text analysis tasks; however, their slicing operations inevitably result in long-term dependency loss in lower-level networks and thus limit their accuracy. Therefore, we propose a breaking point informatio…

Cited by 0SourceScholar