← Search

Jincenzi Wu

8 accepted papers

2026

A Computational Framework for Evaluating Human-likeness in LLMs' Open-ended Human Behaviors

ICML 2026poster

Large Language Models (LLMs) have found widespread application and research in scenarios such as role-playing and sociological simulations. Despite the growing use of LLM-based agents to simulate human activities, the extent to which their behaviors resemble human behavior remains underexplored. As …

Cited by 0SourceScholar
2026

MMSU: A Massive Multi-task Spoken Language Understanding and Reasoning Benchmark

ICLR 2026poster

Speech inherently contains rich acoustic information that extends far beyond the textual language. In real-world spoken communication, effective interpretation often requires integrating semantic meaning (e.g., content), paralinguistic features (e.g., emotions, speed, pitch) and phonological charact…

Cited by 0SourcecodeScholar
2025

InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training

ACL 2025long

Recent advancements in speech large language models (SpeechLLMs) have attracted considerable attention. Nonetheless, current methods exhibit suboptimal performance in adhering to speech instructions. Notably, the intelligence of models significantly diminishes when processing speech-form input as co…

2025

SocialCC: Interactive Evaluation for Cultural Competence in Language Agents

ACL 2025long

Large Language Models (LLMs) are increasingly deployed worldwide, yet their ability to navigate cultural nuances remains underexplored. Misinterpreting cultural content can lead to AI-generated responses that are offensive or inappropriate, limiting their usability in global applications such as cus…

2025

SocialSim: Towards Socialized Simulation of Emotional Support Conversation

AAAI 2025technical

Emotional support conversation (ESC) helps reduce people's psychological stress and provide emotional value through interactive dialogues. Due to the high cost of crowdsourcing a large ESC corpus, recent attempts use large language models for dialogue augmentation. However, existing approaches large…

Cited by 0SourcePDFScholar
2024

COKE: A Cognitive Knowledge Graph for Machine Theory of Mind

ACL 2024long

Theory of mind (ToM) refers to humans’ ability to understand and infer the desires, beliefs, and intentions of others. The acquisition of ToM plays a key role in humans’ social cognition and interpersonal relations. Though indispensable for social intelligence, ToM is still lacking for modern AI and…

2024

Depression Detection in Clinical Interviews with LLM-Empowered Structural Element Graph

NAACL 2024long

Depression is a widespread mental health disorder affecting millions globally. Clinical interviews are the gold standard for assessing depression, but they heavily rely on scarce professional clinicians, highlighting the need for automated detection systems. However, existing methods only capture pa…

2024

ToMBench: Benchmarking Theory of Mind in Large Language Models

ACL 2024long

Theory of Mind (ToM) is the cognitive capability to perceive and ascribe mental states to oneself and others. Recent research has sparked a debate over whether large language models (LLMs) exhibit a form of ToM. However, existing ToM evaluations are hindered by challenges such as constrained scope,…