← Search

Yilian Liu

2 accepted papers

2026

MIDAS: Multi-Image Dispersion and Semantic Reconstruction for Jailbreaking MLLMs

ICLR 2026poster

Multimodal Large Language Models (MLLMs) have achieved remarkable performance but remain vulnerable to jailbreak attacks that can induce harmful content and undermine their secure deployment. Previous studies have shown that introducing additional inference steps, which disrupt security attention, c…

Cited by 0SourcecodeScholar
2025

Auditing Meta-Cognitive Hallucinations in Reasoning Large Language Models

NeurIPS 2025poster

The development of Reasoning Large Language Models (RLLMs) has significantly improved multi-step reasoning capabilities, but it has also made hallucination problems more frequent and harder to eliminate. While existing approaches address hallucination through external knowledge integration, model pa…

Cited by 0SourcecodeScholar