← Search

Xiaoming Wang

6 accepted papers

2026

Unleashing Vision Transformer Potential in Image Quality Assessment via Global-Local Adaptive Interaction

ICASSP 2026poster

In the field of Blind Image Quality Assessment (BIQA), accurately predicting the perceptual quality of authentically distorted images remains highly challenging due to the diverse and complex distortions present in natural environments. Although existing methods have achieved notable accuracy, their…

Cited by 0SourcePDFScholar
2024

Direct Large Language Model Alignment Through Self-Rewarding Contrastive Prompt Distillation

ACL 2024long

Aligning large language models (LLMs) with human expectations without human-annotated preference data is an important problem. In this paper, we propose a method to evaluate the response preference by using the output probabilities of response pairs under contrastive prompt pairs, which could achiev…

2024

Disturbance Observer Based Terminal Sliding Mode Control for an Electromagnetic Actuated Micropositioner With Prescribed Performance

RA-L 2024

Precise and smooth control of an electromagnetic actuated micropositioner is a challenging task due to the system's nonlinear electromagnetic dynamics, as well as its unique structure and application scenarios. The main challenges include high-order nonlinearity, modeling uncertainties, input buffet

Cited by 4SourceScholar
2024

Mitigating Optimization Conflict in Domain Adversarial Neural Network via Uncertainty-Aware

ICASSP 2024accepted

In prior studies, domain adversarial neural networks (DANNs) are used to align image-level features regardless of foreground and background. However, the conventional discriminator in DANNs may leads the feature extractor to disregard cross-domain features rather than aligning them. This phenomenon…

Cited by 0SourceScholar
2023

Multilateral Semantic Relations Modeling for Image Text Retrieval

CVPR 2023poster

Image-text retrieval is a fundamental task to bridge vision and language by exploiting various strategies to fine-grained alignment between regions and words. This is still tough mainly because of one-to-many correspondence, where a set of matches from another modality can be accessed by a random qu…

Cited by 32SourcePDFScholar
2022

Fully Convolutional One-Stage 3D Object Detection on LiDAR Range Images

NeurIPS 2022accept

We present a simple yet effective fully convolutional one-stage 3D object detector for LiDAR point clouds of autonomous driving scenes, termed FCOS-LiDAR. Unlike the dominant methods that use the bird-eye view (BEV), our proposed detector detects objects from the range view (RV, a.k.a. range image)…

Cited by 130SourcePDFScholar