Multi-State Consistency Visual Language Model Combine Wavelet Transform for Weakly Supervised Robot Visual Segmentation
Robotic visual segmentation is essential for enabling robots to operate in complex environments. Although supervised methods have achieved remarkable progress, their dependence on dense annotations hinders scalability. Weakly supervised semantic segmentation (WSSS) alleviates this issue but suffers …