← Search

Zijian Cheng

2 accepted papers

2026

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

ICML 2026poster

While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynamic nature of user queries, tool sets, and interaction dynamics. To address this generalization gap, we formalize **OpenAgent** (Tool-Use Agent in Ope…

Cited by 0SourceScholar
2026

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning

ICML 2026poster

Multi-model learning has attracted great attention in visual-text tasks. However, visual-tabular data, which plays a pivotal role in high-stakes domains like healthcare and industry, remains underexplored. In this paper, we introduce \textit{VT-Bench}, the first unified benchmark for standardizing v…

Cited by 0SourceScholar