← Search

Xijia Tao

3 accepted papers

2026

Beyond Confidence: Adaptive and Coherent Decoding for Diffusion Language Models

ICML 2026poster

Diffusion Language Models (DLMs) have recently achieved significant success due to their any-order generation capabilities. However, existing inference methods typically rely on local, immediate-step metrics—such as confidence or entropy—which inherently lack a more reliable perspective, leading to …

Cited by 0SourceScholar
2026

MMSearch-Plus: Benchmarking Provenance-Aware Search for Multimodal Browsing Agents

ICLR 2026poster

Existing multimodal browsing benchmarks often fail to require genuine multimodal reasoning, as many tasks can be solved with text-only heuristics without vision-in-the-loop verification. We introduce MMSearch-Plus, a 311-task benchmark that enforces multimodal understanding by requiring extraction a…

Cited by 0SourcecodeScholar
2025

ImgTrojan: Jailbreaking Vision-Language Models with ONE Image

NAACL 2025long

There has been an increasing interest in the alignment of large language models (LLMs) with human values. However, the safety issues of their integration with a vision module, or vision language models (VLMs), remain relatively underexplored. In this paper, we propose a novel jailbreaking attack aga…