AAAI 2026technical0 citations

Self-Enhanced Image Clustering with Cross-Modal Semantic Consistency

Zihan Li, Wei Sun, Jing Hu, Jianhua Yin, Xing Wang, Erwei Yin, Jianlong Wu

Abstract

While large language-image pre-trained models like CLIP offer powerful generic features for image clustering, existing methods typically freeze the encoder. This creates a fundamental mismatch between the model

BibTeX
@inproceedings{aaai2026_selfenhancedimag,
  title = {Self-Enhanced Image Clustering with Cross-Modal Semantic Consistency},
  author = {Zihan Li and Wei Sun and Jing Hu and Jianhua Yin and Xing Wang and Erwei Yin and Jianlong Wu},
  booktitle = {AAAI 2026},
  year = {2026}
}