← Search

Muhammad Bilal

3 accepted papers

2026

Keep It Frozen: Domain-Routed Conditional Residual Modulation for Multi-Domain Vision Transformers

CVPR 2026

Medical imaging remains challenging due to acoustic shadows, motion blur, and indistinct boundaries, while adapting vision models to such domains often requires heavy task-specific fine-tuning and can degrade general-image capability. We propose DCRM-ViT, a domain-conditioned residual modulation fra

Cited by 0SourceScholar
2026

MLLM-HWSI: A Multimodal Large Language Model for Hierarchical Whole Slide Image Understanding

CVPR 2026

Whole Slide Images (WSIs) exhibit hierarchical structure, where diagnostic information emerges from cellular morphology, regional tissue organization, and global context. Existing Computational Pathology (CPath) Multimodal Large Language Models (MLLMs) typically compress an entire WSI into a single

Cited by 0SourcecodeScholar
2024

Beyond Success: Quantifying Demonstration Quality in Learning from Demonstration

IROS 2024poster

Learning from Demonstration (LfD) empowers novice users to teach robots daily life tasks without writing sophisticated code, thereby promoting the democratization of robotics. However, novice users often provide sub-optimal demonstrations, which can potentially impact the robot’s ability to efficien…

Cited by 2SourceScholar