← Search

Yimeng Ren

3 accepted papers

2026

Detoxifying Large Language Models via Localized Feature Editing with Sparse Autoencoders

IJCAI 2026

Large Language Models (LLMs) powerful generative capabilities also pose significant risks, underscoring the need for effective detoxification methods to ensure safer deployment. Due to the polysemantic nature of LLM neurons, recent neuron intervention methods inevitably entangle unrelated concepts,

Cited by 0Scholar
2026

From Chaos to Cure: A Prefix Heuristics Guided Model-Agnostic Adaptive Detoxification Framework

AAAI 2026technical

The impressive performance of large language models (LLMs) also brings inherent toxicity risks, prompting the need for effective detoxification to support responsible deployment. Prevailing methods generally follow an inflexible model-specific fashion, addressing only individual models or model fami

Cited by 0SourcePDFScholar
2025

R2DQG: A Quality Meets Diversity Framework for Question Generation over Knowledge Bases

IJCAI 2025

The task of Knowledge-Based Question Generation (KBQG) involves generating natural language questions from structured knowledge sources, posing unique challenges in balancing linguistic diversity and semantic relevance. Existing models often focus on maximizing surface-level similarity to ground-tru