SAM-Veteran: An MLLM-Based Human-like SAM Agent for Reasoning Segmentation
Significant progress has been made in reasoning segmentation by combining multi-modal large language models (MLLMs) with the Segment Anything Model (SAM): the former excel in reasoning and vision–language alignment, while the latter offers powerful pixel-level understanding. However, current paradig…