← Search

Shihui Zhang

3 accepted papers

2026

SEATrack: Simple, Efficient, and Adaptive Multimodal Tracker

CVPR 2026

Parameter-efficient fine-tuning (PEFT) in multimodal tracking reveals a concerning trend where recent performance gains are often achieved at the cost of inflated parameter budgets, which fundamentally erodes PEFT's efficiency promise. In this work, we introduce SEATrack, a Simple, Efficient, and Ad

Cited by 0SourcecodeScholar
2025

CorrBEV: Multi-View 3D Object Detection by Correlation Learning with Multi-modal Prototypes

CVPR 2025poster

Camera-only multi-view 3D object detection in autonomous driving has witnessed encouraging developments in recent years, largely attributed to the revolution of fundamental architectures in modeling bird's eye view (BEV). Despite the growing overall average performance, we contend that the explorati…

Cited by 0SourcePDFScholar
2024

Opnet: Deep Occlusion Perception Network with Boundary Awareness for Amodal Instance Segmentation

ICASSP 2024accepted

The Amodal Instance Segmentation (AIS) task aims to infer the visible and occluded regions of an object instance. Existing AIS methods typically focus on directly predicting visible and occluded regions or leveraging prior knowledge to guide predictions. However, these methods often ignore the perce…

Cited by 0SourceScholar