← Search

Ruihan Tao

1 accepted papers

2025

X-WebAgentBench: A Multilingual Interactive Web Benchmark for Evaluating Global Agentic System

ACL 2025finding

Recently, large language model (LLM)-based agents have achieved significant success in interactive environments, attracting significant academic and industrial attention. Despite these advancements, current research predominantly focuses on English scenarios. In reality, there are over 7,000 languag…