← Search

Xin Tian

19 accepted papers

2026

Addressing Semantic Blind Spots in Text-to-SQL via Component Pre-generation and AST Matching Rewards

ICML 2026poster

In recent years, significant advancements in large language models have greatly propelled the development of Text-to-SQL tasks. However, due to the token-by-token sequential generation mechanism employed by these models, they encounter a semantic blind spot problem with respect to pending SQL compon…

Cited by 0SourceScholar
2026

MS^2Gait: A Multi-Scale Spatio-Temporal Fusion Network for LiDAR-based Gait Recognition

CVPR 2026

3D LiDAR-based gait recognition has gained increasing attention due to its robustness to illumination, privacy preservation, and capability for long-range and non-contact identity verification. However, existing point cloud-based methods suffer from two critical limitations: they fail to model seman

Cited by 0SourceScholar
2026

Video Mirror Detection with the Motion-in-Depth Cue

AAAI 2026technical

Detecting mirror regions in RGB videos is essential for scene understanding in applications such as scene reconstruction and robotic navigation. Existing video mirror detectors typically rely on cues like inside-outside mirror correspondences and 2D motion inconsistencies. However, these methods oft

Cited by 0SourcePDFScholar
2024

Cross-Scale Domain Adaptation with Comprehensive Information for Pansharpening

IJCAI 2024poster

Deep learning-based pansharpening methods typically use simulated data at the reduced-resolution scale for training. It limits their performance when generalizing the trained model to the full-resolution scale due to incomprehensive information utilization of panchromatic (PAN) images at the full-re…

2023

High-Frequency Transformer Network Based on Window Cross-Attention for Pansharpening

ICASSP 2023accepted

Inspired by the powerful ability to capture long-distance dependencies in the vision transformer, we propose a novel high-frequency transformer network based on window cross-attention to fuse panchromatic (PAN) and multispectral (MS) images for a high-resolution MS image. To overcome the problem bro…

Cited by 0SourceScholar
2023

Query Enhanced Knowledge-Intensive Conversation via Unsupervised Joint Modeling

ACL 2023long

In this paper, we propose an unsupervised query enhanced approach for knowledge-intensive conversations, namely QKConv. There are three modules in QKConv: a query generator, an off-the-shelf knowledge selector, and a response generator. QKConv is optimized through joint training, which produces the…

2023

Robust and Scalable Gaussian Process Regression and Its Applications

CVPR 2023poster

This paper introduces a robust and scalable Gaussian process regression (GPR) model via variational learning. This enables the application of Gaussian processes to a wide range of real data, which are often large-scale and contaminated by outliers. Towards this end, we employ a mixture likelihood mo…

2022

Bi-Directional Object-Context Prioritization Learning for Saliency Ranking

CVPR 2022poster

The saliency ranking task is recently proposed to study the visual behavior that humans would typically shift their attention over different objects of a scene based on their degrees of saliency. Existing approaches focus on learning either object-object or object-scene relations. Such a strategy fo…

Cited by 38PDFcodeScholar
2022

Coherent Point Drift Revisited for Non-Rigid Shape Matching and Registration

CVPR 2022poster

In this paper, we explore a new type of extrinsic method to directly align two geometric shapes with point-to-point correspondences in ambient space by recovering a deformation, which allows more continuous and smooth maps to be obtained. Specifically, the classic coherent point drift is revisited a…

Cited by 17PDFScholar
2022

Q-TOD: A Query-driven Task-oriented Dialogue System

EMNLP 2022main

Existing pipelined task-oriented dialogue systems usually have difficulties adapting to unseen domains, whereas end-to-end systems are plagued by large-scale knowledge bases in practice. In this paper, we introduce a novel query-driven task-oriented dialogue system, namely Q-TOD. The essential infor…

2022

RGL: A Simple yet Effective Relation Graph Augmented Prompt-based Tuning Approach for Few-Shot Learning

NAACL 2022findings

Pre-trained language models (PLMs) can provide a good starting point for downstream applications. However, it is difficult to generalize PLMs to new tasks given a few labeled samples. In this work, we show that Relation Graph augmented Learning (RGL) can improve the performance of few-shot natural l…

2021

When Face Recognition Meets Occlusion: A New Benchmark

ICASSP 2021accepted

The existing face recognition datasets usually lack occlusion samples, which hinders the development of face recognition. Especially during the COVID-19 coronavirus epidemic, wearing a mask has become an effective means of preventing the virus spread. Traditional CNN-based face recognition models tr…

Cited by 0SourceScholar
2020

Attention-Guided Deraining Network Via Stage-Wise Learning

ICASSP 2020accepted

Due to diverse rain shapes, directions, densities as well as different distances to cameras, rain streaks in the air are interweaved and overlapped. However, most existing deraining methods are inherently oblivious this phenomenon and tend to learn a single rain streak layer to simulate this complex…

Cited by 0SourceScholar
2020

Cartoon-Texture Decomposition-Based Variational Pansharpening

ICASSP 2020accepted

Pansharpening is widely used to increase the spatial resolution of a multispectral (MS) image by fusing with a panchromatic (PAN) image that has high-spatial resolution and the same scene. In this paper, the similarities of MS and PAN images in cartoon-texture space are exploited. The cartoon and te…

Cited by 0SourceScholar
2020

Fusionndvi: A Novel Fusion Method for NDVI in Remote Sensing

ICASSP 2020accepted

Normalized difference vegetation index (NDVI) is widely utilized to examine vegetation coverage and estimate crop yield. To obtain a high-resolution (HR) NDVI, fusion techniques, which first generates a HR multispectral (MS) image by fusing a low-resolution (LR) MS image and a HR panchromatic image,…

Cited by 0SourceScholar
2019

Multimodal Retinal Image Registration and Fusion Based on Sparse Regularization via a Generalized Minimax-concave Penalty

ICASSP 2019accepted

We introduce a novel framework for the fusion of retinal OCT and confocal images of mice with uveitis. Input images are semi-automatically registered and then fused to provide more informative retinal images for analysis by ophthalmologists and clinicians. The proposed feature-based registration app…

Cited by 0SourceScholar