← Search

Zhixin Li

19 accepted papers

2025

Collaborative Dual-Branch Spatial-Frequency Enhancement Network for Low-Light Images

ICASSP 2025accepted

Low-light images are commonly present due to imaging factors such as insufficient light, night shooting and back lit. Existing low-light image enhancement (LLIE) methods typically rely on a low-light input image for enhancement, which seldom leverage information contained in its high-light counterpa…

Cited by 0SourceScholar
2025

JailPO: A Novel Black-Box Jailbreak Framework via Preference Optimization Against Aligned LLMs

AAAI 2025technical

Large Language Models (LLMs) aligned with human feedback have recently garnered significant attention. However, it remains vulnerable to jailbreak attacks, where adversaries manipulate prompts to induce harmful outputs. Exploring jailbreak attacks enables us to investigate the vulnerabilities of LLM…

Cited by 0SourcePDFScholar
2025

PDCE: Patch-wise Dynamic Curve Estimation for Low-Light Image Enhancement

ICASSP 2025accepted

Low-light image enhancement (LLIE) can be reformulated as an image-specific curve estimation (CE) problem. Traditional CE-based methods struggle with issues such as uniform processing across different regions, static parameter estimation, and lack of effective global semantic enhancement. To address…

Cited by 0SourceScholar
2025

Prototypical Graph Alignment for Text-based Person Search

ICASSP 2025accepted

Text-based Person Search is one of the downstream tasks of cross-modal retrieval. The key challenge is aligning features of two extremely irrelavant modalities into the same latent space. Recent works within Prototype Learning introduce a few learnable parameters to map heterogeneous features into t…

Cited by 0SourceScholar
2024

Gradually Spatio-Temporal Feature Activation for Target Tracking

ICASSP 2024accepted

Most existing transformer-based trackers use ViT [1] as the backbone to extract and fuse feature tokens of target templates and search region. Since both the target template and the search region contain background information, their tokens are prone to background interference in interaction that af…

Cited by 0SourceScholar
2023

Overcoming Language Priors for Visual Question Answering via Loss Rebalancing Label and Global Context

UAI 2023poster

Despite the advances in Visual Question Answering (VQA), many VQA models currently suffer from language priors (i.e. generating answers directly from questions without using images), which severely reduces their robustness in real-world scenarios. We propose a novel training strategy called Loss Reb…

Cited by 12SourcePDFScholar
2023

Target-Aware Tracking with Long-Term Context Attention

AAAI 2023technical

Most deep trackers still follow the guidance of the siamese paradigms and use a template that contains only the target without any contextual information, which makes it difficult for the tracker to cope with large appearance changes, rapid target movement, and attraction from similar objects. To al…

2022

Fine-Grained matching with multi-perspective similarity modeling for cross-modal retrieval

UAI 2022poster

Cross-modal retrieval relies on learning inter-modal correspondences. Most existing approaches focus on learning global or local correspondence and fail to explore fine-grained multi-level alignments. Moreover, it remains to be investigated how to infer more accurate similarity scores. In this paper…

Cited by 1SourcePDFScholar