← Search

Xianzhou Zeng

2 accepted papers

2026

UCPO: Uncertainty-Aware Policy Optimization

ICML 2026poster

The key to building trustworthy Large Language Models (LLMs) lies in endowing them with inherent uncertainty expression capabilities to mitigate the hallucinations that restrict their high-stakes applications. However, existing RL paradigms such as GRPO often suffer from Advantage Bias due to binary…

Cited by 0SourceScholar
2025

MHBench: Demystifying Motion Hallucination in VideoLLMs

AAAI 2025technical

Similar to Language or Image LLMs, VideoLLMs are also plagued by hallucination issues. Hallucinations in videos not only manifest in the spatial dimension regarding the perception of the existence of visual objects (static) but also the temporal dimension influencing the perception of actions and ev…