3MDBench: Medical Multimodal Multi-agent Dialogue Benchmark
Though Large Vision-Language Models (LVLMs) are being actively explored in medicine, their ability to conduct complex real-world telemedicine consultations combining accurate diagnosis with professional dialogue remains underexplored. This paper presents 3MDBench ( M edical M ultimodal M ulti-agent