← Search

Yuyang Sun

5 accepted papers

2026

WRitEer: A Multi-Objective, Preference-Driven Multi-Agent Framework for Human-Like Advanced Text Generation

AAAI 2026technical

Advanced text generation is paramount for enhancing the naturalness of human-computer interaction and improving emotional expressiveness. Current mainstream methods largely rely on large language models (LLMs) for single-turn generation, often lacking the interactivity and multi-dimensional feedback

Cited by 0SourcePDFScholar
2025

ACORD: An Expert-Annotated Retrieval Dataset for Legal Contract Drafting

ACL 2025long

Contract clause retrieval is foundational to contract drafting because lawyers rarely draft contracts from scratch; instead, they locate and revise the most relevant precedent clauses. We introduce the Atticus Clause Retrieval Dataset (ACORD), the first expert-annotated benchmark specifically design…

Cited by 0SourcePDFScholar
2025

Number it: Temporal Grounding Videos like Flipping Manga

CVPR 2025poster

Video Large Language Models (Vid-LLMs) have made remarkable advancements in comprehending video content for QA dialogue. However, they struggle to extend this visual understanding to tasks requiring precise temporal localization, known as Video Temporal Grounding (VTG). To address this, we introduce…

2025

Transtreaming: Adaptive Delay-aware Transformer for Real-time Streaming Perception

AAAI 2025technical

Real-time object detection is critical for the decision-making process for many real-world applications, such as collision avoidance and path planning in autonomous driving. This work presents an innovative real-time streaming perception method, Transtreaming, which addresses the challenge of real-t…

2025

VEU-Bench: Towards Comprehensive Understanding of Video Editing

CVPR 2025highlight

Widely shared videos on the internet are often edited. Recently, although Video Large Language Models (Vid-LLMs) have made great progress in general video understanding tasks, their capabilities in video editing understanding (VEU) tasks remain unexplored. To address this gap, in this paper, we intr…

Cited by 0SourcePDFScholar