← Search

Chihao Shen

2 accepted papers

2025

Towards Evaluating Proactive Risk Awareness of Multimodal Language Models

NeurIPS 2025poster

Human safety awareness gaps often prevent the timely recognition of everyday risks. In solving this problem, a proactive safety artificial intelligence (AI) system would work better than a reactive one. Instead of just reacting to users' questions, it would actively watch people’s behavior and their…

Cited by 0SourceScholar
2024

Does ChatGPT Know That It Does Not Know? Evaluating the Black-Box Calibration of ChatGPT

COLING 2024main

Recently, ChatGPT has demonstrated remarkable performance in various downstream tasks such as open-domain question answering, machine translation, and code generation. As a general-purpose task solver, an intriguing inquiry arises: Does ChatGPT itself know that it does not know, without any access t…

Cited by 6SourcePDFScholar