← Search

Yian Li

3 accepted papers

2026

Enhancing Action and Ingredient Modeling for Semantically Grounded Recipe Generation

ICASSP 2026poster

Recent advances in Multimodal Large Language Models (MLMMs) have enabled recipe generation from food images, yet outputs often contain semantically incorrect actions or ingredients despite high lexical scores (e.g., BLEU, ROUGE). To address this gap, we propose a semantically grounded framework that…

Cited by 0SourcePDFScholar
2023

Flow-Attention-based Spatio-Temporal Aggregation Network for 3D Mask Detection

NeurIPS 2023poster

Anti-spoofing detection has become a necessity for face recognition systems due to the security threat posed by spoofing attacks. Despite great success in traditional attacks, most deep-learning-based methods perform poorly in 3D masks, which can highly simulate real faces in appearance and structur…