← Search

Jian Zhou

24 accepted papers

2026

Fast Proteome-Scale Protein Interaction Retrieval via Residue-Level Factorization

ICLR 2026poster

Protein-protein interactions (PPIs) are mediated at the residue level. Most sequence-based PPI models consider residue-residue interactions across two proteins, which can yield accurate interaction scores but are too slow to scale. At proteome scale, identifying candidate PPIs requires evaluating ne…

Cited by 0SourcecodeScholar
2026

LMC-VIO: Lane Model-Constrained Monocular Inertial Visual SLAM for High-Precision Localization in Highway Scenes

ICRA 2026poster

Continuous stability, as one of the core modules of the autopilot system, is particularly important for its performance. However, as the vehicle speed increases, the system positioning error may be amplified, consequently introducing deviations in the positioning consistency of the system. The inher…

Cited by 0Scholar
2026

LagLLM: LLM-empowered lead–lag dependency learning for spatial-temporal time series forecasting

ICML 2026poster

Spatial–temporal time series forecasting is challenging due to complex lead–lag dependencies, which are often ignored or inadequately modeled by existing methods. Thus, we propose LagLLM, the first LLM-empowered framework that explicitly models lead–lag dependencies by unifying data-driven dynamics …

Cited by 0SourceScholar
2026

Neural Dynamic GI: Random-Access Neural Compression for Temporal Lightmaps in Dynamic Lighting Environments

CVPR 2026

High-quality global illumination (GI) in real-time rendering is commonly achieved using precomputed lighting techniques, with lightmap as the standard choice. To support GI for static objects in dynamic lighting environments, multiple lightmaps at different lighting conditions need to be precomputed

Cited by 0SourceScholar
2026

PCRNet: Phase-aware Complex Refinement Network for EEG-based Auditory Attention Decoding

ICML 2026poster

Auditory attention decoding (AAD) based on Electroencephalography (EEG) aims to identify the attended speaker in multi-speaker environments. However, existing methods typically overlook the crucial phase information of EEG signals, which limits their ability to distinguish structured neural patterns…

Cited by 0SourceScholar
2026

Robust Inter-Series Dependency Modeling for Time Series Forecasting via Information-Theoretic Alignment

ICML 2026poster

While iTransformer pioneered general inter-variate dependency (IVD) modeling in Transformers for multivariate time series forecasting (MTSF), subsequent research on such universal paradigms has been surprisingly scarce. Through comprehensive analysis, we identify a critical structural inconsistency …

Cited by 0SourceScholar
2026

TumorChain: Interleaved Multimodal Chain-of-Thought Reasoning for Traceable Clinical Tumor Analysis

ICLR 2026poster

Accurate tumor analysis is central to clinical radiology and precision oncology, where early detection, reliable lesion characterization, and pathology-level risk assessment directly guide diagnosis, staging, and treatment planning. Chain-of-Thought (CoT) reasoning is particularly critical in this s…

Cited by 0SourcecodeScholar
2025

BSDB-Net: Band-Split Dual-Branch Network with Selective State Spaces Mechanism for Monaural Speech Enhancement

AAAI 2025technical

Although the complex spectrum-based speech enhancement (SE) methods have achieved significant performance, coupling amplitude and phase can lead to a compensation effect, where amplitude information is sacrificed to compensate for the phase that is harmful to SE. In addition, to further improve the…

Cited by 0SourcePDFScholar
2025

Co-Fix3D: Enhancing 3D Object Detection With Collaborative Refinement

RA-L 2025

3D object detection in driving scenarios is particularly challenging due to factors such as sensor noise, occlusions, and the inherent sparsity of LiDAR point clouds, which can lead to the loss or incompleteness of key features, in turn affecting perception performance. To address these challenges,

Cited by 0SourcecodeScholar
2025

DeepLA-Net: Very Deep Local Aggregation Networks for Point Cloud Analysis

CVPR 2025poster

Due to the irregular and disordered data structure in 3D point clouds, prior works have focused on designing more sophisticated local representation methods to capture these complex local patterns. However, the recognition performance has saturated over the past few years, indicating that increasing…

2025

Lane Model-Constrained Monocular Inertial Visual SLAM for High-Precision Localization in Highway Scenes

RA-L 2025

Continuous stability, as one of the core modules of the autopilot system, is particularly important for its performance. However, as the vehicle speed increases, the system positioning error may be amplified, consequently introducing deviations in the positioning consistency of the system. The inher

Cited by 0SourceScholar
2025

Linguistic Neuron Overlap Patterns to Facilitate Cross-lingual Transfer on Low-resource Languages

EMNLP 2025

The current Large Language Models (LLMs) face significant challenges in improving their performance on low-resource languagesand urgently need data-efficient methods without costly fine-tuning.From the perspective of language-bridge,we propose a simple yet effective method, namely BridgeX-ICL, to im

2025

ListenNet: A Lightweight Spatio-Temporal Enhancement Nested Network for Auditory Attention Detection

IJCAI 2025

Auditory attention detection (AAD) aims to identify the direction of the attended speaker in multi-speaker environments from brain signals, such as Electroencephalography (EEG) signals. However, existing EEG-based AAD methods overlook the spatio-temporal dependencies of EEG signals, limiting their d

2025

M3ANet: Multi-scale and Multi-Modal Alignment Network for Brain-Assisted Target Speaker Extraction

IJCAI 2025

The brain-assisted target speaker extraction (TSE) aims to extract the attended speech from mixed speech by utilizing the brain neural activities, for example Electroencephalography (EEG). However, existing models overlook the issue of temporal misalignment between speech and EEG modalities, which h

2025

MHANet: Multi-scale Hybrid Attention Network for Auditory Attention Detection

IJCAI 2025

Auditory attention detection (AAD) aims to detect the target speaker in a multi-talker environment from brain signals, such as electroencephalography (EEG), which has made great progress. However, most AAD methods solely utilize attention mechanisms sequentially and overlook valuable multi-scale con

2025

SparseMeXt: Unlocking the Potential of Sparse Representations for HD Map Construction

IROS 2025

Recent advancements in high-definition (HD) map construction have demonstrated the effectiveness of dense representations, which heavily rely on computationally intensive bird’s-eye view (BEV) features. While sparse representations offer a more efficient alternative by avoiding dense BEV processing,

Cited by 4SourceScholar
2023

CancerUniT: Towards a Single Unified Model for Effective Detection, Segmentation, and Diagnosis of Eight Major Cancers Using a Large Collection of CT Scans

ICCV 2023poster

Human readers or radiologists routinely perform full-body multi-organ multi-disease detection and diagnosis in clinical practice, while most medical AI systems are built to focus on single organs with a narrow list of a few diseases. This might severely limit AI's clinical adoption. A certain number…

Cited by 12PDFScholar
2023

Dirichlet Diffusion Score Model for Biological Sequence Generation

ICML 2023poster

Designing biological sequences is an important challenge that requires satisfying complex constraints and thus is a natural problem to address with deep generative modeling. Diffusion generative models have achieved considerable success in many applications. Score-based generative stochastic differe…

2022

Conditional Stroke Recovery for Fine-Grained Sketch-Based Image Retrieval

ECCV 2022poster

"The key to Fine-Grained Sketch Based Image Retrieval (FG-SBIR) is to establish fine-grained correspondence between sketches and images. Since sketches only consist of abstract strokes, stroke recognition ability plays an important role in FG-SBIR. However, existing works usually ignore the unique f…

2021

Contrastive Learning for Compact Single Image Dehazing

CVPR 2021poster

Single image dehazing is a challenging ill-posed problem due to the severe information degeneration. However, existing deep learning based dehazing methods only adopt clear images as positive samples to guide the training of dehazing network while negative information is unexploited. Moreover, most…

Cited by 876PDFcodeScholar
2021

Novelty Detection via Contrastive Learning with Negative Data Augmentation

IJCAI 2021poster

Novelty detection is the process of determining whether a query example differs from the learned training distribution. Previous generative adversarial networks based methods and self-supervised approaches suffer from instability training, mode dropping, and low discriminative ability. We overcome s…

Cited by 17SourcePDFScholar
2021

Temporal Segmentation of Fine-gained Semantic Action: A Motion-Centered Figure Skating Dataset

AAAI 2021technical

Temporal Action Segmentation (TAS) has achieved great success in many fields such as exercise rehabilitation, movie editing, etc. Currently, task-driven TAS is a central topic in human action analysis. However, motion-centered TAS, as an important topic, is little researched due to unavailable datas…

2020

Accelerating Linear Algebra Kernels on a Massively Parallel Reconfigurable Architecture

ICASSP 2020accepted

Much of the recent work on domain-specific architectures has focused on bridging the gap between performance/efficiency and programmability. We consider one such example architecture, Transformer, consisting of light-weight cores interconnected by caches and crossbars that supports run-time reconfig…

Cited by 0SourceScholar