← Search

Jiuxin Cao

10 accepted papers

2025

Denoise-then-Retrieve: Text-Conditioned Video Denoising for Video Moment Retrieval

IJCAI 2025

Current text-driven Video Moment Retrieval (VMR) methods encode all video clips, including irrelevant ones, disrupting multimodal alignment and hindering optimization. To this end, we propose a denoise-then-retrieve paradigm that explicitly filters text-irrelevant clips from videos and then retrieve

Cited by 0SourcePDFScholar
2025

External Reliable Information-enhanced Multimodal Contrastive Learning for Fake News Detection

AAAI 2025technical

With the rapid development of the Internet, the information dissemination paradigm has changed and the efficiency has been improved greatly. While this also brings the quick spread of fake news and leads to negative impacts on cyberspace. Currently, the information presentation formats have evolved…

2025

MAGRET: Machine-generated Text Detection with Rewritten Texts

COLING 2025main

With the quick advancement in text generation ability of Large Language Mode(LLM), concerns about the misuse of machine-generated content have grown, raising potential violations of legal and ethical standards. Some existing studies concentrate on detecting machine-generated text in open-source mode…

Cited by 0SourcePDFScholar
2025

MambaML: Exploring State Space Models for Multi-Label Image Classification

ICCV 2025poster

Mamba, a selective state-space model, has recently seen widespread application across various visual tasks due to its exceptional ability to capture long-range dependencies. While promising results have been demonstrated in image classification, its potential in multi-label image classification rema…

Cited by 0SourcePDFScholar
2025

Positive Text Reframing under Multi-strategy Optimization

COLING 2025main

Differing from sentiment transfer, positive reframing seeks to substitute negative perspectives with positive expressions while preserving the original meaning. With the emergence of pre-trained language models (PLMs), it is possible to achieve acceptable results by fine-tuning PLMs. Nevertheless, g…

2025

PsyAdvisor: A Plug-and-Play Strategy Advice Planner with Proactive Questioning in Psychological Conversations

ACL 2025long

Proactive questioning is essential in psychological conversations as it helps uncover deeper issues and unspoken concerns. Current psychological LLMs are constrained by passive response mechanisms, limiting their capacity to deploy proactive strategies for psychological counseling. To bridge this ga…

2024

CEPT: A Contrast-Enhanced Prompt-Tuning Framework for Emotion Recognition in Conversation

COLING 2024main

Emotion Recognition in Conversation (ERC) has attracted increasing attention due to its wide applications in public opinion analysis, empathetic conversation generation, and so on. However, ERC research suffers from the problems of data imbalance and the presence of similar linguistic expressions fo…

Cited by 3SourcePDFScholar
2024

Causal-Story: Local Causal Attention Utilizing Parameter-Efficient Tuning for Visual Story Synthesis

ICASSP 2024accepted

The excellent text-to-image synthesis capability of diffusion models has driven progress in synthesizing coherent visual stories. The current state-of-the-art method combines the features of historical captions, historical frames, and the current captions as conditions for generating the current fra…

Cited by 0SourceScholar
2023

Scene-Aware Label Graph Learning for Multi-Label Image Classification

ICCV 2023poster

Multi-label image classification refers to assigning a set of labels for an image. One of the main challenges of this task is how to effectively capture the correlation among labels. Existing studies on this issue mostly rely on the statistical label co-occurrence or semantic similarity of labels. H…

Cited by 31PDFScholar
2022

CORN: Co-Reasoning Network for Commonsense Question Answering

COLING 2022main

Commonsense question answering (QA) requires machines to utilize the QA content and external commonsense knowledge graph (KG) for reasoning when answering questions. Existing work uses two independent modules to model the QA contextual text representation and relationships between QA entities in KG,…

Cited by 10SourcePDFScholar