← Search

Yisong Su

3 accepted papers

2025

WTU-EVAL: A Whether-or-Not Tool Usage Evaluation Benchmark for Large Language Models

ICASSP 2025accepted

Although Large Language Models (LLMs) excel in NLP tasks, they still need external tools to extend their ability. Current research on tool learning with LLMs often assumes mandatory tool use, which does not always align with real-world situations, where the necessity for tools is uncertain, and inco…

Cited by 0SourceScholar
2023

Exploiting Emotion-Semantic Correlations for Empathetic Response Generation

EMNLP 2023long findings

Empathetic response generation aims to generate empathetic responses by understanding the speaker's emotional feelings from the language of dialogue. Recent methods capture emotional words in the language of communicators and construct them as static vectors to perceive nuanced emotions. However, l…

Cited by 0SourcecodeScholar
2023

MenatQA: A New Dataset for Testing the Temporal Comprehension and Reasoning Abilities of Large Language Models

EMNLP 2023long findings

Large language models (LLMs) have shown nearly saturated performance on many natural language processing (NLP) tasks. As a result, it is natural for people to believe that LLMs have also mastered abilities such as time understanding and reasoning. However, research on the temporal sensitivity of LLM…

Cited by 0SourcecodeScholar