← Search

Qinghong Yang

3 accepted papers

2025

FruitMMBench: A Multi-modal Benchmark for Fruit Quality Assessment

ICASSP 2025accepted

The rapid advancement of Large Vision-Language Models (LVLMs) has brought notable improvements in tasks like visual recognition and multi-modal understanding, demonstrating significant potential in real-world applications. However, their performances on issues related to daily life such as fruit qua…

Cited by 0SourceScholar
2023

AltCLIP: Altering the Language Encoder in CLIP for Extended Language Capabilities

ACL 2023findings

CLIP (Contrastive Language–Image Pretraining) is an English multimodal representation model learned from a massive amount of English text-image pairs and has achieved great success in various downstream tasks, including image classification, text-to-image retrieval, and image generation. When extend…

2022

Exploiting Global and Local Hierarchies for Hierarchical Text Classification

EMNLP 2022main

Hierarchical text classification aims to leverage label hierarchy in multi-label text classification. Existing methods encode label hierarchy in a global view, where label hierarchy is treated as the static hierarchical structure containing all labels. Since global hierarchy is static and irrelevant…