2025
EVT: Efficient View Transformation for Multi-Modal 3D Object Detection
ICCV 2025poster
Multi-modal sensor fusion in Bird's Eye View (BEV) representation has become the leading approach for 3D object detection. However, existing methods often rely on depth estimators or transformer encoders to transform image features into BEV space, which reduces robustness or introduces significant c…