← Search

Xiujie Song

3 accepted papers

2025

A Cognitive Evaluation Benchmark of Image Reasoning and Description for Large Vision-Language Models

NAACL 2025long

Large Vision-Language Models (LVLMs), despite their recent success, are hardly comprehensively tested for their cognitive abilities. Inspired by the prevalent use of the Cookie Theft task in human cognitive tests, we propose a novel evaluation benchmark to evaluate high-level cognitive abilities of…

Cited by 3SourcePDFScholar
2023

Transferable and Efficient: Unifying Dynamic Multi-Domain Product Categorization

ACL 2023industry

As e-commerce platforms develop different business lines, a special but challenging product categorization scenario emerges, where there are multiple domain-specific category taxonomies and each of them evolves dynamically over time. In order to unify the categorization process and ensure efficiency…