← Search

Wonjun Jang

2 accepted papers

2026

OrchestrationBench: LLM-Driven Agentic Planning and Tool Use in Multi-Domain Scenarios

ICLR 2026poster

Recent progress in Large Language Models (LLMs) has transformed them from text generators into agentic systems capable of multi-step reasoning, structured planning, and tool use. However, existing benchmarks inadequately capture their ability to orchestrate complex workflows across multiple domains…

Cited by 0SourcecodeScholar
2022

APEACH: Attacking Pejorative Expressions with Analysis on Crowd-Generated Hate Speech Evaluation Datasets

EMNLP 2022finding

In hate speech detection, developing training and evaluation datasets across various domains is the critical issue. Whereas, major approaches crawl social media texts and hire crowd-workers to annotate the data. Following this convention often restricts the scope of pejorative expressions to a singl…