← Search

Bin. Jiang

16 accepted papers

2026

PCASim: Promptable Closed-Loop Adversarial Simulation for Urban Traffic Environment

ICRA 2026poster

Real-world autonomous driving, particularly in urban environments with numerous corner cases, requires rigorous testing to ensure product safety and robustness. However, few studies have explored integrating adversarial scenario generation with the training of safety agents in closed-loop testing, e…

2025

Dynamic Spectral Graph Anomaly Detection

AAAI 2025technical

Graph anomaly detection is crucial for identifying anomalous nodes within graphs and addressing applications like financial fraud detection and social spam detection. Recent spectral graph neural network methods advance graph anomaly detection by focusing on anomalies that notably affect the distrib…

2025

Subdomain Uncertainty Optimization for Cross-Speed Fault Diagnosis

ICASSP 2025accepted

Cross-speed bearing fault diagnosis based on unsupervised domain adaptation can handle data distribution differences across various operating speeds, supporting intelligent maintenance of equipment like wind turbines with variable operating speeds. Existing methods focus on aligning sample distribut…

Cited by 0SourceScholar
2025

VPCI: Self-Supervised Visual Prompt-Guided Cross-Domain Interactive Image Fusion Framework

ICASSP 2025accepted

Image fusion combines information from multi-modality images to produce high-quality fused images with enhanced clarity, contrast, and informativeness. However, limited ground truth fusion data lead to difficulties in effectively training these fusion models. Moreover, current studies lack of fine-g…

Cited by 0SourceScholar
2024

EDformer: Transformer-Based Event Denoising Across Varied Noise Levels

ECCV 2024poster

"Currently, there is relatively limited research on the background activity noise of event cameras in different brightness conditions, and the relevant real-world datasets are extremely scarce. This limitation contributes to the lack of robustness in existing event denoising algorithms when applied…

2024

IPED: An Implicit Perspective for Relational Triple Extraction based on Diffusion Model

NAACL 2024long

Relational triple extraction is a fundamental task in the field of information extraction, and a promising framework based on table filling has recently gained attention as a potential baseline for entity relation extraction. However, inherent shortcomings such as redundant information and incomplet…

2024

Long Term Memory-Enhanced Via Causal Reasoning for Text-To-Video Retrieval

ICASSP 2024accepted

The T2VR task aims to retrieve videos that are semantically relevant to the given query text in a large number of unlabeled videos. Most of the existing methods adopt a representation encoding strategy that can only focus on limited contextual information, and lack the ability to focus on the long m…

Cited by 0SourceScholar
2023

Learning To Dub Movies via Hierarchical Prosody Models

CVPR 2023poster

Given a piece of text, a video clip and a reference audio, the movie dubbing (also known as visual voice clone, V2C) task aims to generate speeches that match the speaker's emotion presented in the video using the desired speaker voice as reference. V2C is more challenging than conventional text-to-…

2022

Stacked Multi-Scale Attention Network for Image Colorization

ICASSP 2022accepted

Deep convolutional networks (CNNs) show their potential in image colorization for producing plausible results. Recently, the attention mechanism further boosts the performances of CNNs by constructing channel and spatial interactions. However, existing attention methods are performed in a single-sca…

Cited by 0SourceScholar
2021

DeFLOCNet: Deep Image Editing via Flexible Low-Level Controls

CVPR 2021poster

User-intended visual content fills the hole regions of an input image in the image editing scenario. The coarse lowlevel inputs, which typically consist of sparse sketch lines and color dots, convey user intentions for content creation (i.e., free-form editing). While existing methods combine an inp…

Cited by 42PDFcodeScholar
2020

METNet: A Mutual Enhanced Transformation Network for Aspect-based Sentiment Analysis

COLING 2020main

Aspect-based sentiment analysis (ABSA) aims to determine the sentiment polarity of each specific aspect in a given sentence. Existing researches have realized the importance of the aspect for the ABSA task and have derived many interactive learning methods that model context based on specific aspect…

Cited by 18SourcePDFScholar
2020

PEDNet: A Persona Enhanced Dual Alternating Learning Network for Conversational Response Generation

COLING 2020main

Endowing a chatbot with a personality is essential to deliver more realistic conversations. Various persona-based dialogue models have been proposed to generate personalized and diverse responses by utilizing predefined persona information. However, generating personalized responses is still a chall…

2020

Rethinking Image Inpainting via a Mutual Encoder-Decoder with Feature Equalizations

ECCV 2020poster

Deep encoder-decoder based CNNs have advanced image inpainting methods for hole filling. While existing methods recover structures and textures step-by-step in the hole regions, they typically use two encoder-decoders for separate recovery. The CNN features of each encoder are learned to capture eit…