← Search

Zhuofu Tao

3 accepted papers

2026

FaceShield: Explainable Face Anti-Spoofing with Multimodal Large Language Models

AAAI 2026technical

Face anti-spoofing (FAS) is crucial for protecting facial recognition systems from presentation attacks. Previous methods approached this task as a classification problem, lacking interpretability and reasoning behind the predicted results. Recently, multimodal large language models (MLLMs) have sho

Cited by 0SourcePDFScholar
2026

Position: We Need A Unified Definition of Hallucination (It’s The World Model, Stupid!)

ICML 2026poster

Despite numerous attempts at mitigation since the inception of language models, hallucinations remain a persistent problem even in today's frontier LLMs. Why is this? We review existing definitions of hallucination and fold them into a single, unified definition wherein prior definitions are subsume…

Cited by 0SourceScholar
2022

Retrieve, Caption, Generate: Visual Grounding for Enhancing Commonsense in Text Generation Models

AAAI 2022technical

We investigate the use of multimodal information contained in images as an effective method for enhancing the commonsense of Transformer models for text generation. We perform experiments using BART and T5 on concept-to-text generation, specifically the task of generative commonsense reasoning, or C…