← Search

Zikang Guo

5 accepted papers

2026

MCP-AgentBench: Evaluating Real-World Language Agent Performance with MCP-Mediated Tools

AAAI 2026technical

The Model Context Protocol (MCP) is rapidly emerging as a pivotal open standard, designed to enhance agent-tool integration and interoperability, and is positioned to unlock a new era of powerful, interconnected, and genuinely utilitarian agentic AI. However, despite MCP

Cited by 0SourcePDFScholar
2025

M-RangeDetector: Enhancing Generalization in Machine-Generated Text Detection through Multi-Range Attention Masks

ACL 2025finding

The increasing capability and widespread usage of large language models (LLMs) highlight the desirability of automatic detection of machine-generated text. Existing supervised detectors often overfit within their training domains, as they have primarily learned domain-specific textual features, such…

Cited by 0SourcePDFScholar
2025

MIRROR: Multi-agent Intra- and Inter-Reflection for Optimized Reasoning in Tool Learning

IJCAI 2025

Complex tasks involving tool integration pose significant challenges for Large Language Models (LLMs), leading to the emergence of multi-agent workflows as a promising solution. Reflection has emerged as an effective strategy for correcting erroneous trajectories in agentic workflows. However, exist

Cited by 0SourcePDFScholar
2024

Disentangled Learning with Synthetic Parallel Data for Text Style Transfer

ACL 2024long

Text style transfer (TST) is an important task in natural language generation, which aims to transfer the text style (e.g., sentiment) while keeping its semantic information. Due to the absence of parallel datasets for supervision, most existing studies have been conducted in an unsupervised manner,…

2024

IDEATE: Detecting AI-Generated Text Using Internal and External Factual Structures

COLING 2024main

The effective detection of AI-generated text is a vital principle to ensure responsible use of large language models (LLMs). Previous studies mainly focused on discovering and utilizing internal evidences contained in the text itself to perform the detection, while ignoring external evidences implic…