← Search

QING XU

13 accepted papers

2025

A Generalized Control Revision Method for Autonomous Driving Safety

ICRA 2025

Safety is one of the most crucial challenges of autonomous driving vehicles, and one solution to guarantee safety is to employ an additional control revision module after the planning backbone. Control Barrier Function (CBF) has been widely used because of its strong mathematical foundation on safet

Cited by 0SourceScholar
2025

CocoER: Aligning Multi-Level Feature by Competition and Coordination for Emotion Recognition

CVPR 2025poster

With the explosion of human-machine interaction, emotion recognition has reignited attention. Previous works focus on improving visual feature fusion and reasoning from multiple image levels. Although it is non-trivial to deduce a person's emotion by integrating multi-level feature (head, body a…

2025

Enhancing Multimodal Analogical Reasoning Through Triplet Interaction

ICASSP 2025accepted

Analogical reasoning is fundamental to human cognition and plays a crucial role across various fields. However, previous studies have primarily focused on single-modal analogical reasoning, often ignoring the benefits of incorporating structural knowledge. Research in cognitive psychology has shown…

Cited by 0SourceScholar
2025

Hierarchical Context Interaction and Reasoning with Transformer for Emotion Recognition

ICASSP 2025accepted

Emotion recognition is an important task in computer vision. However, current approaches using hard associations (e.g., element-wise addition or concatenation) suffer from information pollution. To overcome these challenges and utilize information at different scales, we present a novel Transformer-…

Cited by 0SourceScholar
2025

Swin-VasMamba: A Topologically Constrained Model For 3D Vascular Segmentation

ICASSP 2025accepted

Accurate 3D vascular segmentation is essential for diagnosing and treating vascular diseases. This task remains challenging due to the complexity of the 3D data and the morphological diversity of blood vessels. In recent years, state space models (SSMs) have received a great attention for its good p…

Cited by 0SourceScholar
2025

Unilaw-R1: A Large Language Model for Legal Reasoning with Reinforcement Learning and Iterative Inference

EMNLP 2025

Reasoning-focused large language models (LLMs) are rapidly evolving across various domains, yet their capabilities in handling complex legal problems remains underexplored. In this paper, we introduce Unilaw-R1, a large language model tailored for legal reasoning. With a lightweight 7-billion parame

2025

Vision-Driven 2D Supervised Fine-Tuning Framework for Bird's Eye View Perception

IROS 2025

Visual bird’s eye view (BEV) perception, dute to its excellent perceptual capabilities, is progressively replacing costly LiDAR-based perception systems, especially in the realm of urban intelligent driving. However, this type of perception still relies on LiDAR data to construct ground truth databa

Cited by 2SourceScholar
2024

Flaws can be Applause: Unleashing Potential of Segmenting Ambiguous Objects in SAM

NeurIPS 2024poster

As the vision foundation models like the Segment Anything Model (SAM) demonstrate potent universality, they also present challenges in giving ambiguous and uncertain predictions. Significant variations in the model output and granularity can occur with simply subtle changes in the prompt, contradict…

2024

Reinforced Cross-Domain Knowledge Distillation on Time Series Data

NeurIPS 2024poster

Unsupervised domain adaptation methods have demonstrated superior capabilities in handling the domain shift issue which widely exists in various time series tasks. However, their prominent adaptation performances heavily rely on complex model architectures, posing an unprecedented challenge in deplo…

Cited by 0SourcePDFScholar
2023

Distilling Universal and Joint Knowledge for Cross-Domain Model Compression on Time Series Data

IJCAI 2023poster

For many real-world time series tasks, the computational complexity of prevalent deep leaning models often hinders the deployment on resource limited environments (e.g., smartphones). Moreover, due to the inevitable domain shift between model training (source) and deploying (target) stages, compress…

2023

Improving Empathetic Dialogue Generation by Dynamically Infusing Commonsense Knowledge

ACL 2023findings

In empathetic conversations, individuals express their empathy towards others. Previous work has mainly focused on generating empathetic responses by utilizing the speaker’s emotion. Besides, external commonsense knowledge has been applied to enhance the system’s understandings of the speaker’s situ…

2023

VAD: Vectorized Scene Representation for Efficient Autonomous Driving

ICCV 2023poster

Autonomous driving requires a comprehensive understanding of the surrounding environment for reliable trajectory planning. Previous works rely on dense rasterized scene representation (e.g., agent occupancy and semantic map) to perform planning, which is computationally intensive and misses the inst…

Cited by 233PDFcodeScholar
2023

Weakly Supervised Learning of Semantic Correspondence through Cascaded Online Correspondence Refinement

ICCV 2023poster

In this paper, we develop a weakly supervised learning algorithm to learn robust semantic correspondences from large-scale datasets with only image-level labels. Following the spirit of multiple instance learning (MIL), we decompose the weakly supervised correspondence learning problem into three st…

Cited by 1PDFcodeScholar