← Search

Shuai Hao

2 accepted papers

2026

Escaping the Likelihood Trap: Geometric Diversity Optimization for Long-Form Image Captioning

ICML 2026poster

The utility of Vision-Language Models (VLMs) in reasoning and auditing tasks hinges on their ability to exhaustively describe visual scenes. However, current models exhibit a pathology we term the Likelihood Trap: standard alignment objectives, specifically MLE and KL-regularization, drive generatio…

Cited by 0SourceScholar
2025

Understanding PII Leakage in Large Language Models: A Systematic Survey

IJCAI 2025

Large Language Models (LLMs) have demonstrated exceptional success across a variety of tasks, particularly in natural language processing, leading to their growing integration into numerous facets of daily life. However, this widespread deployment has raised substantial privacy concerns, especially