← Search

Ziyue Lin

5 accepted papers

2026

Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation

CVPR 2026

We propose Decoupled Residual Denoising Diffusion models (DRDD) for unified and data-efficient image-to-image (I2I) translation. While diffusion models have advanced I2I translation in terms of quality and diversity, we uncover a previously under-explored property in diffusion models. Crucially, bey

Cited by 0SourcecodeScholar
2026

FlowGen: Synthesizing Diverse Flowcharts to Enhance and Benchmark MLLM Reasoning

ICLR 2026poster

Flowcharts are widely used to represent processes and relationships through intuitive visual representations. However, accurately interpreting these diagrams remains challenging due to their structural complexity and high visual diversity. Existing flowchart datasets often lack fine-grained control…

Cited by 0SourcecodeScholar
2026

Unleashing the Potential of Large Language Models for Text-to-Image Generation Through Autoregressive Representation Alignment

AAAI 2026technical

We present Autoregressive Representation Alignment (ARRA), a new training framework that unlocks global-coherent text-to-image generation in autoregressive LLMs without architectural modifications. Different from prior works that require complex architectural redesigns, ARRA aligns LLM

Cited by 0SourcePDFScholar
2025

MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria

NAACL 2025long

Multimodal large language models (MLLMs) have broadened the scope of AI applications. Existing automatic evaluation methodologies for MLLMs are mainly limited in evaluating objective queries without considering real-world user experiences, inadequately addressing the nuances of creative and associat…