← Search

Zhongjiang Yao

2 accepted papers

2026

Knowledge-Enhanced Image Captioning with Adaptive Graph-based Multimodal Alignment and LLM

AAAI 2026technical

Image captioning is crucial for multimodal understanding, bridging visual content and natural language. Despite recent advancements in Large Multimodal Models (LMMs), when faced with unseen entities or scenes in the open world, even when attempting to leverage learned knowledge, models still struggl

Cited by 0SourcePDFScholar
2025

Emotion-aware Structural Enhancement Graph Auto-Encoder for Rumor Detection

ICASSP 2025accepted

Social media is a key channel for information dissemination, making effective rumor detection essential to mitigate misinformation’s societal impact. Although large language models excel in inference and text generation, they struggle with understanding propagation relationships and complex reasonin…

Cited by 0SourceScholar