← Search

Yuping Yan

3 accepted papers

2026

IRIS: Implicit Reward-Guided Internal Sifting for Mitigating Multimodal Hallucination

ICML 2026poster

Hallucination remains a fundamental challenge for Multimodal Large Language Models (MLLMs). While Direct Preference Optimization (DPO) is a key alignment framework, existing approaches often rely heavily on costly external evaluators for scoring or rewriting, incurring off-policy learnability gaps a…

Cited by 0SourceScholar
2026

Mitigating Visual Hallucinations via Semantic Curriculum Preference Optimization in MLLMs

ICML 2026poster

Multimodal Large Language Models (MLLMs) have significantly improved the performance of various tasks, but continue to suffer from visual hallucinations, a critical issue where generated responses contradict visual evidence. While Direct Preference Optimization (DPO) is widely used for alignment, it…

Cited by 0SourceScholar
2025

PolypSense3D: A Multi-Source Benchmark Dataset for Depth-Aware Polyp Size Measurement in Endoscopy

NeurIPS 2025poster

Accurate polyp sizing during endoscopy is crucial for cancer risk assessment but is hindered by subjective methods and inadequate datasets lacking integrated 2D appearance, 3D structure, and real-world size information. We introduce PolypSense3D, the first multi-source benchmark dataset specifically…

Cited by 0SourcecodeScholar