← Search

Sijie Mai

8 accepted papers

2026

Beyond Cosine Similarity: Magnitude-Aware CLIP for No-Reference Image Quality Assessment

AAAI 2026technical

Recent efforts have repurposed the Contrastive Language-Image Pre-training (CLIP) model for No-Reference Image Quality Assessment (NR-IQA) by measuring the cosine similarity between the image embedding and textual prompts such as "a good photo" or "a bad photo." However, this semantic similarity ove

Cited by 0SourcePDFScholar
2025

CyIN: Cyclic Informative Latent Space for Bridging Complete and Incomplete Multimodal Learning

NeurIPS 2025poster

Multimodal machine learning, mimicking the human brain’s ability to integrate various modalities has seen rapid growth. Most previous multimodal models are trained on perfectly paired multimodal input to reach optimal performance. In real‑world deployments, however, the presence of modality is highl…

Cited by 0SourceScholar
2022

Communicative Subgraph Representation Learning for Multi-Relational Inductive Drug-Gene Interaction Prediction

IJCAI 2022poster

Illuminating the interconnections between drugs and genes is an important topic in drug development and precision medicine. Currently, computational predictions of drug-gene interactions mainly focus on the binding interactions without considering other relation types like agonist, antagonist, etc.…

2021

Communicative Message Passing for Inductive Relation Reasoning

AAAI 2021technical

Relation prediction for knowledge graphs aims at predicting missing relationships between entities. Despite the importance of inductive relation prediction, most previous works are limited to a transductive setting and cannot process previously unseen entities. The recent proposed subgraph-based rel…

2021

Which is Making the Contribution: Modulating Unimodal and Cross-modal Dynamics for Multimodal Sentiment Analysis

EMNLP 2021finding

Multimodal sentiment analysis (MSA) draws increasing attention with the availability of multimodal data. The boost in performance of MSA models is mainly hindered by two problems. On the one hand, recent MSA works mostly focus on learning cross-modal dynamics, but neglect to explore an optimal solut…

Cited by 30SourcePDFScholar