← Search

Bingquan Liu

13 accepted papers

2026

CR³: Boosting Compositional Reasoning in MLLMs Through Rule-Based Reinforcement Learning

AAAI 2026technical

Compositional reasoning is a critical capability for multimodal models, enabling systematic understanding of complex scenes through structured combinations of objects, attributes, and relations. However, existing research on this ability primarily focuses on vision-language models (VLMs, e.g., CLIP

Cited by 0SourcePDFScholar
2025

A Dual Contrastive Learning Framework for Enhanced Multimodal Conversational Emotion Recognition

COLING 2025main

Multimodal Emotion Recognition in Conversations (MERC) identifies utterance emotions by integrating both contextual and multimodal information from dialogue videos. Existing methods struggle to capture emotion shifts due to label replication and fail to preserve positive independent modality contrib…

2025

Expand VSR Benchmark for VLLM to Expertize in Spatial Rules

AAAI 2025technical

Distinguishing spatial relations is a basic part of human cognition which requires fine-grained perception on cross-instance. Although benchmarks like MME, MMBench and SEED comprehensively have evaluated various capabilities which already include visual spatial reasoning(VSR). There is still a la…

2025

LaERC-S: Improving LLM-based Emotion Recognition in Conversation with Speaker Characteristics

COLING 2025main

Emotion recognition in conversation (ERC), the task of discerning human emotions for each utterance within a conversation, has garnered significant attention in human-computer interaction systems. Previous ERC studies focus on speaker-specific information that predominantly stems from relationships…

2024

Mutual Information Assisted Graph Convolution Network for Cold-Start Recommendation

ICASSP 2024accepted

To solve the cold-start issue that cold items have no historical interactions to obtain collaborative feature as their representation, existing methods often represent them totally based on content feature obtained from inherent content (i.e., image, video and attributes). However, these methods wil…

Cited by 0SourceScholar
2024

Preference Aware Dual Contrastive Learning for Item Cold-Start Recommendation

AAAI 2024technical

Existing cold-start recommendation methods often adopt item-level alignment strategies to align the content feature and the collaborative feature of warm items for model training, however, cold items in the test stage have no historical interactions with users to obtain the collaborative feature. Th…

2024

Towards Faithful Knowledge Graph Explanation Through Deep Alignment in Commonsense Question Answering

EMNLP 2024main

The fusion of language models (LMs) and knowledge graphs (KGs) is widely used in commonsense question answering, but generating faithful explanations remains challenging. Current methods often overlook path decoding faithfulness, leading to divergence between graph encoder outputs and model predicti…

Cited by 1SourcePDFScholar
2022

How Pre-trained Language Models Capture Factual Knowledge? A Causal-Inspired Analysis

ACL 2022findings

Recently, there has been a trend to investigate the factual knowledge captured by Pre-trained Language Models (PLMs). Many works show the PLMs’ ability to fill in the missing factual words in cloze-style prompts such as ”Dante was born in [MASK].” However, it is still a mystery how PLMs generate the…

Cited by 54SourcePDFScholar
2022

Pre-training Language Models with Deterministic Factual Knowledge

EMNLP 2022main

Previous works show that Pre-trained Language Models (PLMs) can capture factual knowledge. However, some analyses reveal that PLMs fail to perform it robustly, e.g., being sensitive to the changes of prompts when extracting factual knowledge. To mitigate this issue, we propose to let PLMs learn the…

2021

HopRetriever: Retrieve Hops over Wikipedia to Answer Complex Questions

AAAI 2021technical

Collecting supporting evidence from large corpora of text (e.g., Wikipedia) is of great challenge for open-domain Question Answering (QA). Especially, for multi-hop open-domain QA, scattered evidence pieces are required to be gathered together to support the answer extraction. In this paper, we prop…

Cited by 36SourcePDFScholar
2021

Knowledge-Interactive Network with Sentiment Polarity Intensity-Aware Multi-Task Learning for Emotion Recognition in Conversations

EMNLP 2021finding

Emotion Recognition in Conversation (ERC) has gained much attention from the NLP community recently. Some models concentrate on leveraging commonsense knowledge or multi-task learning to help complicated emotional reasoning. However, these models neglect direct utterance-knowledge interaction. In ad…

Cited by 40SourcePDFScholar