← Search

Hyeseon Ahn

4 accepted papers

2026

WaterMod: Modular Token-Rank Partitioning for Probability-Balanced LLM Watermarking

AAAI 2026technical

Large language models now draft news, legal analyses, and software code with human-level fluency. At the same time, regulations such as the EU AI Act mandate that each synthetic passage carry an imperceptible, machine-verifiable mark for provenance. Conventional logit-based watermarks satisfy this r

Cited by 0SourcePDFScholar
2025

AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection

EMNLP 2025

Implicit hate speech detection is challenging due to its subtlety and reliance on contextual interpretation rather than explicit offensive words. Current approaches rely on contrastive learning, which are shown to be effective on distinguishing hate and non-hate sentences. Humans, however, detect im

2025

Analyzing Offensive Language Dataset Insights from Training Dynamics and Human Agreement Level

COLING 2025main

Implicit hate speech detection is challenging due to its subjectivity and context dependence, with existing models often struggling in outof-domain scenarios. We propose CONELA, a novel data refinement strategy that enhances model performance and generalization by integrating human annotation agreem…

Cited by 0SourcePDFScholar
2024

SharedCon: Implicit Hate Speech Detection using Shared Semantics

ACL 2024findings

The ever-growing presence of hate speech on social network services and other online platforms not only fuels online harassment but also presents a growing challenge for hate speech detection. As this task is akin to binary classification, one of the promising approaches for hate speech detection is…

Cited by 6SourcePDFScholar