← Search

Zongqi Wang

6 accepted papers

2026

P-GenRM: Personalized Generative Reward Model with Test-time User-based Scaling

ICLR 2026oral

Personalized alignment of large language models seeks to adapt responses to individual user preferences, typically via reinforcement learning. A key challenge is obtaining accurate, user-specific reward signals in open-ended scenarios. Existing personalized reward models face two persistent limitati…

Cited by 0SourcecodeScholar
2026

Reward Modeling from Natural Language Human Feedback

ICML 2026poster

Reinforcement Learning with Verifiable Reward (RLVR) on preference data has become the mainstream approach for training Generative Reward Models (GRMs). Typically, GRMs generate reasoning chains ending with critiques and preference labels, with RLVR using label correctness as the training reward. Ho…

Cited by 0SourceScholar
2025

Invisible Entropy: Towards Safe and Efficient Low-Entropy LLM Watermarking

EMNLP 2025

Logit-based LLM watermarking traces and verifies AI-generated content by maintaining green and red token lists and increasing the likelihood of green tokens during generation. However, it struggles in low-entropy scenarios, where predictable outputs make green token selection difficult without disru

2025

MorphMark: Flexible Adaptive Watermarking for Large Language Models

ACL 2025long

Watermarking by altering token sampling probabilities based on red-green list is a promising method for tracing the origin of text generated by large language models (LLMs). However, existing watermark methods often struggle with a fundamental dilemma: improving watermark effectiveness (the detectab…

2025

Thinking Racial Bias in Fair Forgery Detection: Models, Datasets and Evaluations

AAAI 2025technical

Due to the successful development of deep image generation technology, forgery detection plays a more important role in social and economic security. Racial bias has not been explored thoroughly in the deep forgery detection field. In the paper, we first contribute a dedicated dataset called the Fai…