← Search

Shanlin Zhou

4 accepted papers

2025

KIA: Knowledge-Guided Implicit Vision-Language Alignment for Chest X-Ray Report Generation

COLING 2025main

Report generation (RG) faces challenges in understanding complex medical images and establishing cross-modal semantic alignment in radiology image-report pairs. Previous methods often overlook fine-grained cross-modal interaction, leading to insufficient understanding of detailed information. Recent…

2024

On the Essence and Prospect: An Investigation of Alignment Approaches for Big Models

IJCAI 2024poster

Big models have achieved revolutionary breakthroughs in the field of AI, but they also pose potential ethical and societal risks to humans. Addressing such problems, alignment technologies were introduced to make these models conform to human preferences and values. Despite the considerable advancem…

Cited by 12SourcePDFScholar
2023

ToViLaG: Your Visual-Language Generative Model is Also An Evildoer

EMNLP 2023long main

Recent large-scale Visual-Language Generative Models (VLGMs) have achieved unprecedented improvement in multimodal image/text generation. However, these models might also generate toxic content, e.g., offensive text and pornography images, raising significant ethical risks. Despite exhaustive studie…

Cited by 0SourcecodeScholar
2022

CHAE: Fine-Grained Controllable Story Generation with Characters, Actions and Emotions

COLING 2022main

Story generation has emerged as an interesting yet challenging NLP task in recent years. Some existing studies aim at generating fluent and coherent stories from keywords and outlines; while others attempt to control the global features of the story, such as emotion, style and topic. However, these…