2025
MMAT-1M: A Large Reasoning Dataset for Multimodal Agent Tuning
ICCV 2025poster
Large Language Models (LLMs), enhanced through agent tuning, have demonstrated remarkable capabilities in Chain-of-Thought (CoT) and tool utilization, significantly surpassing the performance of standalone models. However, the multimodal domain still lacks a large-scale, high-quality agent tuning da…