← Search

Qiang Gao

22 accepted papers

2026

Beyond Graph Priors: A Co-Evolving Framework Under Uncertainty for Enterprise Resilience Assessment

AAAI 2026technical

Assessing enterprise resilience under uncertainty necessitates capturing both intrinsic attributes and evolving inter-enterprise dependencies. However, real-world enterprise systems pose substantial structural challenges: redundant or loosely correlated links can trigger spurious relational inferenc

Cited by 0SourcePDFScholar
2026

Gaze-Based Teleoperation with Intent Inference Model for Robotic Manipulators

ICRA 2026poster

Eye gaze-based control interfaces provide a non-invasive means of enhancing human-robot collaboration for activities of daily living and can reduce the cognitive burden on operators performing complex tasks. Eye gaze has traditionally been used for "gaze triggering," where fixating on an object acti…

Cited by 0Scholar
2026

Shedding the Facades, Connecting the Domains: Detecting Shifting Multimodal Hate Video with Test-Time Adaptation

AAAI 2026technical

Hate Video Detection (HVD) is crucial for online ecosystems. Existing methods assume identical distributions between training (source) and inference (target) data. However, hateful content often evolves into irregular and ambiguous forms to evade censorship, resulting in substantial semantic drift a

Cited by 0SourcePDFScholar
2025

AdaDARE-gamma: Balancing Stability and Plasticity in Multi-modal LLMs through Efficient Adaptation

CVPR 2025poster

Adapting Multi-modal Large Language Models (MLLMs) to target tasks often suffers from catastrophic forgetting, where acquiring new task-specific knowledge compromises performance on pre-trained tasks. In this paper, we introduce AdaDARE-\gamma, an efficient approach that alleviates catastrophic forg…

Cited by 0SourcePDFScholar
2025

Adversity-aware Few-shot Named Entity Recognition via Augmentation Learning

AAAI 2025technical

Few-shot Named Entity Recognition (NER) spotlights the tag of novel entity types in data-limited scenarios or lower-resource settings. Advances with Pre-trained Language Models (PLMs), including BERT, GPT, and their variants, have driven tremendous strategies to leverage context-dependent representa…

Cited by 0SourcePDFScholar
2025

CodeTool: Enhancing Programmatic Tool Invocation of LLMs via Process Supervision

ACL 2025long

Tool invocation significantly enhances the capabilities of Large Language Models (LLMs), yet challenges persist, particularly in complex task scenarios. Current methods, such as instruction-enhanced reasoning and supervised fine-tuning, often result in unnecessarily long reasoning paths and face dif…

Cited by 0SourcePDFScholar
2025

MBA-RAG: a Bandit Approach for Adaptive Retrieval-Augmented Generation through Question Complexity

COLING 2025main

Retrieval Augmented Generation (RAG) has proven to be highly effective in boosting the generative performance of language model in knowledge-intensive tasks. However, existing RAG framework either indiscriminately perform retrieval or rely on rigid single-label classifiers to select retrieval method…

2025

RadarMask: A Novel End-to-End Sparse Millimeter-Wave Radar Sequence Panoptic Segmentation and Tracking Method

ICRA 2025

In the realms of autonomous driving and robotics, radar sensors are garnering growing interest. Scene understanding is crucial for the safe navigation of autonomous systems. Panoptic segmentation and tracking tasks enable the dynamic, semantic multilevel description of the environment surrounding ve

Cited by 1SourcecodeScholar
2025

Responsive Dynamic Graph Disentanglement for Metro Flow Forecasting

AAAI 2025technical

The metro flow in Urban Rail Transit Systems (URTS) differs from other urban traffic flows because it is characterized by: (1) highly predetermined scheduling; and (2) interactively dynamic dependencies over the fixed physical infrastructure that vary with spatiotemporal and environmental factors. N…

2024

DSL-FIQA: Assessing Facial Image Quality via Dual-Set Degradation Learning and Landmark-Guided Transformer

CVPR 2024poster

Generic Face Image Quality Assessment (GFIQA) evaluates the perceptual quality of facial images which is crucial in improving image restoration algorithms and selecting high-quality face images for downstream tasks. We present a novel transformer-based method for GFIQA which is aided by two unique m…

Cited by 9SourcePDFScholar
2024

Disentanglement-Guided Spatial-Temporal Graph Neural Network for Metro Flow Forecasting (Student Abstract)

AAAI 2024technical

In recent intelligent transportation applications, metro flow forecasting has received much attention from researchers. Most prior arts endeavor to explore spatial or temporal dependencies while ignoring the key characteristic patterns underlying historical flows, e.g., trend and periodicity. Althou…

Cited by 0SourcePDFScholar
2024

Enhancing Cross-Document Event Coreference Resolution by Discourse Structure and Semantic Information

COLING 2024main

Existing cross-document event coreference resolution models, which either compute mention similarity directly or enhance mention representation by extracting event arguments (such as location, time, agent, and patient), lackingmthe ability to utilize document-level information. As a result, they str…

2024

Enhancing Fine-Grained Urban Flow Inference via Incremental Neural Operator

IJCAI 2024poster

Fine-grained urban flow inference (FUFI), which involves inferring fine-grained flow maps from their coarse-grained counterparts, is of tremendous interest in the realm of sustainable urban traffic services. To address the FUFI, existing solutions mainly concentrate on investigating spatial dependen…

2024

Harvesting Events from Multiple Sources: Towards a Cross-Document Event Extraction Paradigm

ACL 2024findings

Document-level event extraction aims to extract structured event information from unstructured text. However, a single document often contains limited event information and the roles of different event arguments may be biased due to the influence of the information source.This paper addresses the li…

2024

Spatial-Temporal Augmentation for Crime Prediction (Student Abstract)

AAAI 2024technical

Crime prediction stands as a pivotal concern within the realm of urban management due to its potential threats to public safety. While prior research has predominantly focused on unraveling the intricate dependencies among urban regions and temporal dynamics, the challenges posed by the scarcity and…

Cited by 1SourcePDFScholar
2023

Attention Localness in Shared Encoder-Decoder Model For Text Summarization

ICASSP 2023accepted

Text summarization is to generate a brief version of a given article while maintaining its essential meaning. Most existing solutions typically relied on the standard attention-based encoder-decoder framework, where each token in the source article, including redundancy, would be contributed to the…

Cited by 0SourceScholar
2023

Cross-Regional Fraud Detection via Continual Learning (Student Abstract)

AAAI 2023technical

Detecting fraud is an urgent task to avoid transaction risks. Especially when expanding a business to new cities or new countries, developing a totally new model will bring the cost issue and result in forgetting previous knowledge. This study proposes a novel solution based on heterogeneous trade g…

Cited by 1SourcePDFScholar
2023

Enhancing Knowledge Transfer for Task Incremental Learning with Data-free Subnetwork

NeurIPS 2023poster

As there exist competitive subnetworks within a dense network in concert with Lottery Ticket Hypothesis, we introduce a novel neuron-wise task incremental learning method, namely Data-free Subnetworks (DSN), which attempts to enhance the elastic knowledge transfer across the tasks that sequentially…

2023

Mobility Prediction via Sequential Trajectory Disentanglement (Student Abstract)

AAAI 2023technical

Accurately predicting human mobility is a critical task in location-based recommendation. Most prior approaches focus on fusing multiple semantics trajectories to forecast the future movement of people, and fail to consider the distinct relations in underlying context of human mobility, resulting in…

Cited by 1SourcePDFScholar
2023

Open Anomalous Trajectory Recognition via Probabilistic Metric Learning

IJCAI 2023poster

Typically, trajectories considered anomalous are the ones deviating from usual (e.g., traffic-dictated) driving patterns. However, this closed-set context fails to recognize the unknown anomalous trajectories, resulting in an insufficient self-motivated learning paradigm. In this study, we investiga…

2022

Dynamic Manifold Learning for Land Deformation Forecasting

AAAI 2022technical

Landslides refer to occurrences of massive ground movements due to geological (and meteorological) factors, and can have disastrous impact on property, economy, and even lead to loss of life. The advances of remote sensing provide accurate and continuous terrain monitoring, enabling the study and an…

Cited by 3SourcePDFScholar
2021

An End-to-End Speech Accent Recognition Method Based on Hybrid CTC/Attention Transformer ASR

ICASSP 2021accepted

This paper proposes a novel accent recognition system in the framework of a transformer-based end-to-end speech recognition system. To incorporate the pronunciation and linguistic knowledge into the network, we first pre-train an ASR model in a hybrid CTC/attention manner. Then, focusing on accent r…

Cited by 0SourceScholar