← Search

Weisi Fan

3 accepted papers

2025

FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs

NAACL 2025short

Summarization is one of the most common tasks performed by large language models (LLMs), especially in applications like Retrieval-Augmented Generation (RAG). However, existing evaluations of hallucinations in LLM-generated summaries, and evaluations of hallucination detection models both suffer fro…

2024

SummaCoz: A Dataset for Improving the Interpretability of Factual Consistency Detection for Summarization

EMNLP 2024finding

Summarization is an important application of Large Language Models (LLMs). When judging the quality of a summary, factual consistency holds a significant weight. Despite numerous efforts dedicated to building factual inconsistency detectors, the exploration of explanability remains limited among exi…

2022

Design and Evaluation of Object Classifiers for Probabilistic Decision-Making in Autonomous Systems

ICRA 2022poster

Object classification is a key element that enables effective decision-making in many autonomous systems. A more sophisticated system may also utilize the probability distribution over the classes instead of basing its decision only on the most likely class. This paper introduces new performance met…

Cited by 1SourceScholar