← Search

Zhenyue Qin

5 accepted papers

2025

GeoDANO: Geometric VLM with Domain Agnostic Vision Encoder

EMNLP 2025

We introduce GeoDANO, a geometric vision-language model (VLM) with a domain-agnostic vision encoder, for solving plane geometry problems. Although VLMs have been employed for solving geometry problems, their ability to recognize geometric features remains insufficiently analyzed. To address this gap

2025

LMOD: A Large Multimodal Ophthalmology Dataset and Benchmark for Large Vision-Language Models

NAACL 2025findings

The prevalence of vision-threatening eye diseases is a significant global burden, with many cases remaining undiagnosed or diagnosed too late for effective treatment. Large vision-language models (LVLMs) have the potential to assist in understanding anatomical information, diagnosing eye diseases, a…

Cited by 6SourcePDFScholar
2024

Visual Prompting in LLMs for Enhancing Emotion Recognition

EMNLP 2024main

Vision Large Language Models (VLLMs) are transforming the intersection of computer vision and natural language processing; however, the potential of using visual prompts for emotion recognition in these models remains largely unexplored and untapped. Traditional methods in VLLMs struggle with spatia…

2023

Anonymization for Skeleton Action Recognition

AAAI 2023technical

Skeleton-based action recognition attracts practitioners and researchers due to the lightweight, compact nature of datasets. Compared with RGB-video-based action recognition, skeleton-based action recognition is a safer way to protect the privacy of subjects while having competitive recognition perf…

2021

Invertible Denoising Network: A Light Solution for Real Noise Removal

CVPR 2021poster

Invertible networks have various benefits for image denoising since they are lightweight, information-lossless, and memory-saving during back-propagation. However, applying invertible models to remove noise is challenging because the input is noisy, and the reversed output is clean, following two di…

Cited by 200PDFcodeScholar