← Search

Zhi Zeng

9 accepted papers

2026

CARE: A Molecular-Guided Foundation Model with Adaptive Region Modeling for Whole Slide Image Analysis

CVPR 2026

Foundation models have achieved success in computational pathology, demonstrating generalization across histopathology tasks. However, existing models overlook the heterogeneous and non-uniform organization of regions of interest (ROIs) because they rely on natural image backbones not tailored for t

Cited by 0SourcecodeScholar
2026

Dual-Branch Asymmetric Discrepancy Learning Based on Fake Image Pattern-Coexistence for AI-Generated Image Detection

AAAI 2026technical

With the rapid advancement of generative models, high-fidelity AI-generated images have become increasingly indistinguishable from real images, posing significant challenges to traditional detection methods that rely on explicit artifacts or uniform feature learning. We hypothesize that detection am

Cited by 0SourcePDFScholar
2025

Each Fake News Is Fake in Its Own Way: An Attribution Multi-Granularity Benchmark for Multimodal Fake News Detection

AAAI 2025technical

Social platforms, while facilitating access to information, have also become saturated with a plethora of fake news, resulting in negative consequences. Automatic multimodal fake news detection is a worthwhile pursuit. Existing multimodal fake news datasets only provide binary labels of real or fake…

2025

IMOL: Incomplete-Modality-Tolerant Learning for Multi-Domain Fake News Video Detection

ACL 2025long

While recent advances in fake news video detection have shown promising potential, existing approaches typically (1) focus on a specific domain (e.g., politics) and (2) assume the availability of multiple modalities, including video, audio, description texts, and related images. However, these metho…

Cited by 0SourcePDFScholar
2025

Truth over Tricks: Measuring and Mitigating Shortcut Learning in Misinformation Detection

NeurIPS 2025poster

Misinformation detectors often rely on superficial cues (i.e., shortcuts) that correlate with misinformation in training data but fail to generalize to the diverse and evolving nature of real-world misinformation. This issue is exacerbated by large language models (LLMs), which can easily generate c…

Cited by 0SourceScholar
2024

Event-Radar: Event-driven Multi-View Learning for Multimodal Fake News Detection

ACL 2024long

The swift detection of multimedia fake news has emerged as a crucial task in combating malicious propaganda and safeguarding the security of the online environment. While existing methods have achieved commendable results in modeling entity-level inconsistency, addressing event-level inconsistency f…

Cited by 11SourcePDFScholar
2024

GesGPT: Speech Gesture Synthesis With Text Parsing From ChatGPT

RA-L 2024

Gesture synthesis has gained significant attention as a critical research field, aiming to produce contextually appropriate and natural gestures corresponding to speech or textual input. Although deep learning-based approaches have achieved remarkable progress, they often overlook the rich semantic

Cited by 13SourceScholar
2016

Region matching and similarity enhancing for image retrieval

ICASSP 2016accepted

Many image retrieval systems adopt the bag-of-words model and rely on matching of local descriptors. However, these descriptors of keypoints, such as SIFT, may lead to false matches, since they do not consider the contextual information of the keypoints. In this paper, we incorporate the cues of mea…

Cited by 0SourceScholar
2015

Transmitting informative components of fisher codes for mobile visual search

ICASSP 2015accepted

Existing techniques usually adopt compact descriptors such as Fisher vector for mobile visual search, since compact descriptors are memory-efficient and suitable for fast transmission. In common Fisher vector methods, in order to make the size of image representations small enough for efficient tran…

Cited by 0SourceScholar