← Search

Shangguang Wang

12 accepted papers

2026

Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models

CVPR 2026

The rapid advancement of diffusion-based image generation models has raised serious concerns regarding potential copyright and privacy infringements involving human-created data. Membership inference attacks (MIAs) have emerged as a promising tool for identifying unauthorized data usage during model

Cited by 0SourcecodeScholar
2026

MoTRa: Motion-Aware Target Representation Learning for End-to-End Multi-Object Tracking

IJCAI 2026

Multi-object tracking (MOT) has long faced challenges with identity switches, especially for targets with low appearance discriminability and complex motion. Existing end-to-end trackers typically enhance robustness by modeling long-term temporal information across target-level representations, yet

Cited by 0Scholar
2026

MobiEdit: Resource-efficient Knowledge Editing for Personalized On-device LLMs

ICLR 2026poster

Large language models (LLMs) are deployed on mobile devices to power killer applications such as intelligent assistants. LLMs pre-trained on general corpora often hallucinate when handling personalized or unseen queries, leading to incorrect or outdated responses. Knowledge editing addresses this b…

Cited by 0SourcecodeScholar
2026

PDAR-RSITR: A Progressive Decoupling-Aggregation-Refinement Framework for Remote Sensing Image-Text Retrieval

IJCAI 2026

Remote Sensing Image-Text Retrieval (RSITR) aims to achieve precise retrieval between remote sensing images and textual descriptions. However, existing methods neglect the multi-dimensional cognitive attributes inherent in remote sensing data and struggle to handle them simultaneously, leading to su

Cited by 0Scholar
2026

SCo-Cloud: Satellite Constellation Collaboration for Cloud-Aware Onboard-Computed Imaging and Transmission

AAAI 2026technical

Satellite-acquired optical remote sensing imagery is extensively applied in time-critical applications like traffic surveillance and evaluation of natural disasters. However, clouds, as a common atmospheric phenomenon, frequently obscure observation. Current approaches aim to restore visibility in c

Cited by 0SourcePDFScholar
2026

Towards Whole-corpus Reconstruction of Heterogeneous RAG Knowledge Bases

ICML 2026poster

Retrieval-Augmented Generation (RAG) systems are increasingly deployed to provide query-based access to large knowledge bases, thereby introducing concrete privacy risks whereby the underlying corpus may be partially or fully extracted through the deployed service. Existing extraction attacks typica…

Cited by 0SourceScholar
2025

Black-Box Membership Inference Attack for LVLMs via Prior Knowledge-Calibrated Memory Probing

NeurIPS 2025poster

Large vision-language models (LVLMs) derive their capabilities from extensive training on vast corpora of visual and textual data. Empowered by large-scale parameters, these models often exhibit strong memorization of their training data, rendering them susceptible to membership inference attacks (…

Cited by 0SourcecodeScholar
2025

LoRASuite: Efficient LoRA Adaptation Across Large Language Model Upgrades

NeurIPS 2025poster

As Large Language Models (LLMs) are frequently updated, LoRA weights trained on earlier versions quickly become obsolete. The conventional practice of retraining LoRA weights from scratch on the latest model is costly, time-consuming, and environmentally detrimental, particularly as the diversity of…

Cited by 0SourceScholar
2025

Multimodal Knowledge Retrieval-Augmented Iterative Alignment for Satellite Commonsense Conversation

IJCAI 2025

Satellite technology has significantly influenced our daily lives, manifested in applications such as navigation and communication. With its development, a vast amount of multimodal satellite commonsense data has been generated, thus leading to an urgent demand for conversation about satellite data.

Cited by 0SourcePDFScholar
2025

Variational Multi-Modal Hypergraph Attention Network for Multi-Modal Relation Extraction

IJCAI 2025

Multi-modal relation extraction (MMRE) is a challenging task that seeks to identify relationships between entities with textual and visual attributes. However, existing methods struggle to handle the complexities posed by multiple entity pairs within a single sentence that share similar contextual i

2024

SILENCE: Protecting privacy in offloaded speech understanding on resource-constrained devices

NeurIPS 2024poster

Speech serves as a ubiquitous input interface for embedded mobile devices. Cloud-based solutions, while offering powerful speech understanding services, raise significant concerns regarding user privacy. To address this, disentanglement-based encoders have been proposed to remove sensitive informa…

Cited by 0SourcePDFScholar
2023

Mitigating Task Interference in Multi-Task Learning via Explicit Task Routing With Non-Learnable Primitives

CVPR 2023poster

Multi-task learning (MTL) seeks to learn a single model to accomplish multiple tasks by leveraging shared information among the tasks. Existing MTL models, however, have been known to suffer from negative interference among tasks. Efforts to mitigate task interference have focused on either loss/gra…

Cited by 19SourcePDFScholar