← Search

Jun Shen

16 accepted papers

2026

FedHPro: Federated Hyper-Prototype Learning via Gradient Matching

ICML 2026poster

Federated Learning (FL) enables collaborative training of distributed clients while protecting privacy. To enhance generalization capability in FL, prototype-based FL is in the spotlight, since shared global prototypes offer semantic anchors for aligning client-specific local prototypes. However, ex…

Cited by 0SourceScholar
2026

Incomplete Multi-View Unsupervised Federated Feature Selection via Cooperative Particle Swarm Optimization and Tensor-Aligned Learning

AAAI 2026technical

With the widespread adoption of multi-view data in numerous fields, multi-view unsupervised feature selection (MUFS) has made notable strides in both feature pruning and missing-view completion. Nonetheless, existing MUFS methods typically rely on centralized servers, which cannot meet real-world de

Cited by 0SourcePDFScholar
2026

KANFIS: A Neuro-Symbolic Framework for Interpretable and Uncertainty-Aware Learning

ICML 2026poster

Adaptive Neuro-Fuzzy Inference System (ANFIS) was designed to combine the learning capabilities of neural network with the reasoning transparency of fuzzy logic. However, conventional ANFIS architectures suffer from structural complexity, where the product-based inference mechanism causes an exponen…

Cited by 0SourceScholar
2026

Towards Zero-Shot Diabetic Retinopathy Grading: Learning Generalized Knowledge via Prompt-Driven Matching and Emulating

AAAI 2026technical

As one of the primary causes of visual impairment, Diabetic Retinopathy (DR) requires accurate and robust grading to facilitate timely diagnosis and intervention. Different from conventional DR grading methods that utilize single-view images, recent clinical studies have revealed that multi-view fun

Cited by 0SourcePDFScholar
2025

DICP: Deep In-Context Prompt for Event Causality Identification

EMNLP 2025

Event causality identification (ECI) is a challenging task that involves predicting causal relationships between events in text. Existing prompt-learning-based methods typically concatenate in-context examples only at the input layer, this shallow integration limits the model’s ability to capture th

2025

FedDifRC: Unlocking the Potential of Text-to-Image Diffusion Models in Heterogeneous Federated Learning

ICCV 2025poster

Federated learning aims at training models collaboratively across participants while protecting privacy. However, one major challenge for this paradigm is the data heterogeneity issue, where biased data preferences across multiple clients, harming the model's convergence and performance. In this pap…

2025

InstructSAM: A Training-free Framework for Instruction-Oriented Remote Sensing Object Recognition

NeurIPS 2025poster

Language-guided object recognition in remote sensing imagery is crucial for large-scale mapping and automated data annotation. However, existing open-vocabulary and visual grounding methods rely on explicit category cues, limiting their ability to handle complex or implicit queries that require adva…

Cited by 0SourcecodeScholar
2025

Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical Tasks

ICLR 2025poster

Several medical Multimodal Large Languange Models (MLLMs) have been developed to address tasks involving visual images with textual instructions across various medical modalities, achieving impressive results. Most current medical generalist models are region-agnostic, treating the entire image as…

Cited by 3SourcePDFScholar
2025

Vox-UDA: Voxel-wise Unsupervised Domain Adaptation for Cryo-Electron Subtomogram Segmentation with Denoised Pseudo-Labeling

AAAI 2025technical

Cryo-Electron Tomography (cryo-ET) is a 3D imaging technology that facilitates the study of macromolecular structures at near-atomic resolution. Recent volumetric segmentation approaches on cryo-ET images have drawn widespread interest in the biological sector. However, existing methods heavily rely…

2024

GuessKT: Improving Knowledge Tracing via Considering Guess Behaviors

ICASSP 2024accepted

Knowledge tracing (KT) aims to predict students’ responses to given questions based on their historical question-answering interactions. Recent studies have proposed multiple types of KT models, mainly relying on learners’ feedback to capture the evolution of their knowledge states. However, these m…

Cited by 0SourceScholar
2024

Iterative Search Attribution for Deep Neural Networks

ICML 2024poster

Deep neural networks (DNNs) have achieved state-of-the-art performance across various applications. However, ensuring the reliability and trustworthiness of DNNs requires enhanced interpretability of model inputs and outputs. As an effective means of Explainable Artificial Intelligence (XAI) researc…

2024

Metric from Human: Zero-shot Monocular Metric Depth Estimation via Test-time Adaptation

NeurIPS 2024poster

Monocular depth estimation (MDE) is fundamental for deriving 3D scene structures from 2D images. While state-of-the-art monocular relative depth estimation (MRDE) excels in estimating relative depths for in-the-wild images, current monocular metric depth estimation (MMDE) approaches still face chall…

Cited by 4SourcePDFScholar
2022

Expression might be enough: representing pressure and demand for reinforcement learning based traffic signal control

ICML 2022spotlight

Many studies confirmed that a proper traffic state representation is more important than complex algorithms for the classical traffic signal control (TSC) problem. In this paper, we (1) present a novel, flexible and efficient method, namely advanced max pressure (Advanced-MP), taking both running an…

2022

Obstacle Avoidance for Microrobots in Simulated Vascular Environment Based on Combined Path Planning

RA-L 2022

In order to increase the feasibility of using microrobots to perform microscale tasks and widen the range of potential applications, the use of autonomous control algorithms such as obstacle avoidance is essential. In this study, aiming at the requirement of microrobots for automatic obstacle avoida

Cited by 23SourceScholar