← Search

Lin Wu

9 accepted papers

2026

LightAVSeg: Lightweight Audio-Visual Segmentation

ICML 2026poster

Audio-Visual Segmentation (AVS) targets pixel level localization of sounding emitting objects in videos. However, existing models rely on dense cross-modal attention with quadratic computational cost, limiting their suitability for resource efficient deployment. Most efficiency oriented methods focu…

Cited by 0SourceScholar
2026

PERSONALIZED CELL SEGMENTATION: BENCHMARK AND FRAMEWORK FOR REFERENCE-GUIDED CELL TYPE SEGMENTATION

ICASSP 2026poster

Accurate cell segmentation is critical for biological and medical imaging studies. Although recent deep learning models have advanced this task, most methods are limited to generic cell segmentation, lacking the ability to differentiate specific cell types. In this work, we introduce the Personalize…

Cited by 0SourcePDFScholar
2025

AnnaAgent: Dynamic Evolution Agent System with Multi-Session Memory for Realistic Seeker Simulation

ACL 2025finding

Constrained by the cost and ethical concerns of involving real seekers in AI-driven mental health, researchers develop LLM-based conversational agents (CAs) with tailored configurations, such as profiles, symptoms, and scenarios, to simulate seekers. While these efforts advance AI in mental health,…

2025

On the Integration of Spatial-Temporal Knowledge: A Lightweight Approach to Atmospheric Time Series Forecasting

NeurIPS 2025poster

Transformers have gained attention in atmospheric time series forecasting (ATSF) for their ability to capture global spatial-temporal correlations. However, their complex architectures lead to excessive parameter counts and extended training times, limiting their scalability to large-scale forecasti…

Cited by 0SourceScholar
2021

A Chapter-Wise Understanding System for Text-To-Speech in Chinese Novels

ICASSP 2021accepted

In TTS-based audiobook production, multi-role dubbing and emotional expressions can significantly improve the naturalness of audiobooks. However, it requires manual annotation of original novels with explicit speaker and emotion tags in sentence level, which is extremely time-consuming and costly. I…

Cited by 0SourceScholar
2021

Prediction of Egfr Mutation Status in Lung Adenocarcinoma Using Multi-Source Feature Representations

ICASSP 2021accepted

Epidermal growth factor receptor (EGFR) genotyping is essential to treatment guidelines for the use of tyrosine kinase inhibitors in lung adenocarcinoma. However, accurate and noninvasive methods to detect the EGFR gene are ongoing challenges. In this study, we propose a hybrid framework, namely HC-…

Cited by 0SourceScholar
2020

Zero-Shot Object Detection via Learning an Embedding from Semantic Space to Visual Space

IJCAI 2020poster

Zero-shot object detection (ZSD) has received considerable attention from the community of computer vision in recent years. It aims to simultaneously locate and categorize previously unseen objects during inference. One crucial problem of ZSD is how to accurately predict the label of each object pro…

Cited by 0SourcePDFScholar