← Search

Shayan Ali Akbar

2 accepted papers

2025

SEEval: Advancing LLM Text Evaluation Efficiency and Accuracy through Self-Explanation Prompting

NAACL 2025findings

Large language models (LLMs) have achieved remarkable success in various natural language generation (NLG) tasks, but their performance in automatic text evaluation is not yet ready as human replacements. In this paper, we propose SEEval (Self-Explanation in Evaluation), a novel prompt-based text ev…

Cited by 0SourcePDFScholar
2024

HalluMeasure: Fine-grained Hallucination Measurement Using Chain-of-Thought Reasoning

EMNLP 2024main

Automating the measurement of hallucinations in LLM generated responses is a challenging task as it requires careful investigation of each factual claim in a response. In this paper, we introduce HalluMeasure, a new LLM-based hallucination detection mechanism that decomposes an LLM response into ato…

Cited by 3SourcePDFScholar