← Search

Jaehyeok Lee

3 accepted papers

2026

Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook

ICML 2026poster

As LLMs are globally deployed, aligning their cultural value orientations is critical for safety and user engagement. However, existing benchmarks face the Construct-Composition-Context (C$^3$) challenge: relying on discriminative, multiple-choice formats that probe value knowledge rather than true …

Cited by 0SourceScholar
2025

Self-Training Meets Consistency: Improving LLMs’ Reasoning with Consistency-Driven Rationale Evaluation

NAACL 2025long

Self-training approach for large language models (LLMs) improves reasoning abilities by training the models on their self-generated rationales. Previous approaches have labeled rationales that produce correct answers for a given question as appropriate for training. However, a single measure risks m…

2025

Unintended Harms of Value-Aligned LLMs: Psychological and Empirical Insights

ACL 2025long

The application scope of Large Language Models (LLMs) continues to expand, leading to increasing interest in personalized LLMs that align with human values. However, aligning these models with individual values raises significant safety concerns, as certain values may correlate with harmful informat…