ICRA 2026poster0 citations
SurgVidLM: Towards Multi-Grained Video Understanding with Large Language Model in Robot-Assisted Surgery
Guankun Wang, Junyi Wang, Wenjin Mo, Long Bai, Kun Yuan, Ming Hu, Jinlin Wu, Junjun He
Semantic Scene UnderstandingComputer Vision for Medical RoboticsMedical Robots and Systems