← Search

Yanyi Chen

3 accepted papers

2025

A Cognitive Evaluation Benchmark of Image Reasoning and Description for Large Vision-Language Models

NAACL 2025long

Large Vision-Language Models (LVLMs), despite their recent success, are hardly comprehensively tested for their cognitive abilities. Inspired by the prevalent use of the Cookie Theft task in human cognitive tests, we propose a novel evaluation benchmark to evaluate high-level cognitive abilities of…

Cited by 3SourcePDFScholar
2023

An Interpretable Model Using Evidence Information for Multi-Hop Question Answering Over Long Texts

ICASSP 2023accepted

Machine Reading Comprehension (MRC) is a challenging task in natural language understanding, especially multi-hop question answering (QA) in long texts. One of the challenges in multi-hop QA requires models to produce interpretable answers based on evidence that is selected from a given long text. B…

Cited by 0SourceScholar
2023

Narrow Down Before Selection: A Dynamic Exclusion Model for Multiple-Choice QA

ICASSP 2023accepted

Multiple-choice question answering (MCQA) is a challenging task that requires selecting the correct answer from a set of options based on a given question. There is a trend to use pre-trained encoder-decoder models to solve MCQA. Previous works concentrate on the decoder and adopt the generated text…

Cited by 0SourceScholar