← Search

Jingheng Pan

3 accepted papers

2025

Chinese Toxic Language Mitigation via Sentiment Polarity Consistent Rewrites

EMNLP 2025

Detoxifying offensive language while preserving the speaker’s original intent is a challenging yet critical goal for improving the quality of online interactions. Although large language models (LLMs) show promise in rewriting toxic content, they often default to overly polite rewrites, distorting t

2025

CogSteer: Cognition-Inspired Selective Layer Intervention for Efficiently Steering Large Language Models

ACL 2025finding

Large Language Models (LLMs) achieve remarkable performance through pretraining on extensive data. This enables efficient adaptation to diverse downstream tasks. However, the lack of interpretability in their underlying mechanisms limits the ability to effectively steer LLMs for specific application…

2024

Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding

ACL 2024findings

Large Vision-Language Models (LVLMs) are increasingly adept at generating contextually detailed and coherent responses from visual inputs. However, their application in multimodal decision-making and open-ended generation is hindered by a notable rate of hallucinations, where generated text inaccura…