← Search

Yuyu Zhou

4 accepted papers

2025

MM-OPERA: Benchmarking Open-ended Association Reasoning for Large Vision-Language Models

NeurIPS 2025poster

Large Vision-Language Models (LVLMs) have exhibited remarkable progress. However, deficiencies remain compared to human intelligence, such as hallucination and shallow pattern matching. In this work, we aim to evaluate a fundamental yet underexplored intelligence: association, a cornerstone of human…

Cited by 0SourcecodeScholar
2025

Reproducible Vision-Language Models Meet Concepts Out of Pre-Training

CVPR 2025poster

Contrastive Language-Image Pre-training (CLIP) models as a milestone of modern multimodal intelligence, its generalization mechanism grasped massive research interests in the community. While existing studies limited in the scope of pre-training knowledge, hardly underpinned its generalization to co…

Cited by 0SourcePDFScholar
2024

Deformation And Penetration Hybrid Detection-Net For Parcels Inspection In Industrial Supply Chain

ICASSP 2024accepted

The express delivery industry has become integral to modern social life, but supply chain parcels, especially those made of corrugated cardboard, are at risk of damage during transportation. Although corrugated cardboard boxes offer some impact resistance, they can still experience deformation and p…

Cited by 0SourceScholar
2024

Transformer Model with Multi-Type Classification Decisions for Intrusion Attack Detection of Track Traffic and Vehicle

ICASSP 2024accepted

Security vulnerabilities, illustrated by the menace of track traffic or vehicle hacking, present a substantial risk to the Controller Area Network (CAN) bus, enabling unauthorized remote access and intrusion. Nevertheless, existing vehicle intrusion detection models encounter challenges in capturing…

Cited by 0SourceScholar