← Search

WEIHE ZHAI

2 accepted papers

2025

ScImage: How good are multimodal large language models at scientific text-to-image generation?

ICLR 2025poster

Multimodal large language models (LLMs) have demonstrated impressive capabilities in generating high-quality images from textual instructions. However, their performance in generating scientific images—a critical application for accelerating scientific progress—remains underexplored. In this work, w…

2024

Towards Faithful Knowledge Graph Explanation Through Deep Alignment in Commonsense Question Answering

EMNLP 2024main

The fusion of language models (LMs) and knowledge graphs (KGs) is widely used in commonsense question answering, but generating faithful explanations remains challenging. Current methods often overlook path decoding faithfulness, leading to divergence between graph encoder outputs and model predicti…

Cited by 1SourcePDFScholar