← Search

Wenbing Tang

7 accepted papers

2026

Shedding Light on VLN Robustness: A Black-box Framework for Indoor Lighting-based Adversarial Attack

CVPR 2026

Vision-and-Language Navigation (VLN) agents have made remarkable progress, but their robustness remains insufficiently studied. Existing adversarial evaluations often rely on perturbations that manifest as unusual textures rarely encountered in everyday indoor environments. Errors under such contriv

Cited by 0SourcecodeScholar
2026

SimpleDiffusion: A Lightweight and Efficient Conditional Diffusion Model for Multi-Modal Salient Object Detection

AAAI 2026technical

Multi-modal salient object detection (MSOD), which integrates complementary modalities such as depth or thermal data, primarily faces two challenges: accurately preserving salient object details and effectively aligning cross-modal features. Recent advances in using Stable Diffusion to generate imag

Cited by 0SourcePDFScholar
2025

DiMSOD: A Diffusion-Based Framework for Multi-Modal Salient Object Detection

AAAI 2025technical

Multi-modal salient object detection (SOD) through the integration of additional data such as depth or thermal information has become a significant task in computer vision during recent years. Traditionally, the challenges of identifying salient objects in RGB, RGB-D (Depth), and RGB-T (Thermal) ima…

Cited by 0SourcePDFScholar
2025

Multi-modal Salient Object Detection via a Unified Diffusion Model

ICASSP 2025accepted

Salient Object Detection (SOD) aims to identify and segment the most striking elements within an image. Salient object detection methods can be differentiated into several types according to the input data, such as RGB-D (Depth) and RGB-T (Thermal). Previous research primarily focused on saliency de…

Cited by 0SourceScholar
2025

Seg-diffusion: Text-to-Image Diffusion Model for Open-Vocabulary Semantic Segmentation

ICASSP 2025accepted

Open-vocabulary semantic segmentation (OVSS) is a challenging computer vision task that labels each pixel within an image based on text descriptions. Recent advancements in OVSS are largely attributed to the increased model capacity. However, these models often struggle with unfamiliar images or uns…

Cited by 0SourceScholar
2023

GAN-Based Robust Motion Planning for Mobile Robots Against Localization Attacks

RA-L 2023

Motion planning (MP) is essential but challenging for mobile robots. Most of the existing MP methods, at each instant, compute an action based on the states of the robot and the surrounding obstacles, assuming that the robot's localization module is attack-free. Unfortunately, the localization modul

Cited by 9SourceScholar