← Search

Jinqing Zhang

4 accepted papers

2026

ResWorld: Temporal Residual World Model for End-to-End Autonomous Driving

ICLR 2026poster

The comprehensive understanding capabilities of world models for driving scenarios have significantly improved the planning accuracy of end-to-end autonomous driving frameworks. However, the redundant modeling of static regions and the lack of deep interaction with trajectories hinder world models f…

Cited by 0SourcecodeScholar
2025

GeoBEV: Learning Geometric BEV Representation for Multi-view 3D Object Detection

AAAI 2025technical

Bird's-Eye-View (BEV) representation has emerged as a mainstream paradigm for multi-view 3D object detection, demonstrating impressive perceptual capabilities. However, existing methods overlook the geometric quality of BEV representation, leaving it in a low-resolution state and failing to restore…

2024

FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection

ECCV 2024poster

"Although multi-view 3D object detection based on the Bird’s-Eye-View (BEV) paradigm has garnered widespread attention as an economical and deployment-friendly perception solution for autonomous driving, there is still a performance gap compared to LiDAR-based methods. In recent years, several cross…

2023

SA-BEV: Generating Semantic-Aware Bird's-Eye-View Feature for Multi-view 3D Object Detection

ICCV 2023poster

Recently, the pure camera-based Bird's-Eye-View (BEV) perception provides a feasible solution for economical autonomous driving. However, the existing BEV-based multi-view 3D detectors generally transform all image features into BEV features, without considering the problem that the large proportion…

Cited by 34PDFcodeScholar