← Search

Guangneng Hu

6 accepted papers

2026

Dual-Seed Evolutionary Algorithm for Noise Optimization in Diffusion Models

AAAI 2026technical

Diffusion models have emerged as state-of-the-art generative methods, particularly excelling in conditional tasks such as prompt-driven image synthesis. While recent research emphasizes the pivotal role of noise seeds in enhancing text-image alignment and generating human-preferred outputs,these wor

Cited by 0SourcePDFScholar
2025

ICM-Assistant: Instruction-tuning Multimodal Large Language Models for Rule-based Explainable Image Content Moderation

AAAI 2025technical

Controversial contents largely inundate the Internet, infringing various cultural norms and child protection standards. Traditional Image Content Moderation (ICM) models fall short in producing precise moderation decisions for diverse standards, while recent multimodal large language models (MLLMs),…

2025

SEFE: Superficial and Essential Forgetting Eliminator for Multimodal Continual Instruction Tuning

ICML 2025poster

Multimodal Continual Instruction Tuning (MCIT) aims to enable Multimodal Large Language Models (MLLMs) to incrementally learn new tasks without catastrophic forgetting, thus adapting to evolving requirements. In this paper, we explore the forgetting caused by such incremental training, categorizing…

2024

Improving Vision and Language Concepts Understanding with Multimodal Counterfactual Samples

ECCV 2024poster

"Vision and Language (VL) models have achieved remarkable performance in a variety of multimodal learning tasks. The success of these models is attributed to learning a joint and aligned representation space of visual and text. However, recent popular VL models still struggle with concepts understan…

2024

Towards More Faithful Natural Language Explanation Using Multi-Level Contrastive Learning in VQA

AAAI 2024technical

Natural language explanation in visual question answer (VQA-NLE) aims to explain the decision-making process of models by generating natural language sentences to increase users' trust in the black-box systems. Existing post-hoc methods have achieved significant progress in obtaining a plausible exp…

2023

TITAN : Task-oriented Dialogues with Mixed-Initiative Interactions

IJCAI 2023poster

In multi-domain task-oriented dialogue systems, users proactively propose a series of domain-specific requests that can often be under-or over-specified, sometimes with ambiguous and cross-domain demands. System-sided initiative would be necessary to identify certain situations and appropriately int…