← Search

Yilun Wang

6 accepted papers

2023

Occ3D: A Large-Scale 3D Occupancy Prediction Benchmark for Autonomous Driving

NeurIPS 2023poster

Robotic perception requires the modeling of both 3D geometry and semantics. Existing methods typically focus on estimating 3D bounding boxes, neglecting finer geometric details and struggling to handle general, out-of-vocabulary objects. 3D occupancy prediction, which estimates the detailed occupanc…

2023

VectorMapNet: End-to-end Vectorized HD Map Learning

ICML 2023poster

Autonomous driving systems require High-Definition (HD) semantic maps to navigate around urban roads. Existing solutions approach the semantic mapping problem by offline manual annotation, which suffers from serious scalability issues. Recent learning-based methods produce dense rasterized segmentat…

2023

ViP3D: End-to-End Visual Trajectory Prediction via 3D Agent Queries

CVPR 2023poster

Perception and prediction are two separate modules in the existing autonomous driving systems. They interact with each other via hand-picked features such as agent bounding boxes and trajectories. Due to this separation, prediction, as a downstream module, only receives limited information from the…

2021

DETR3D: 3D Object Detection from Multi-view Images via 3D-to-2D Queries

CoRL 2021poster

We introduce a framework for multi-camera 3D object detection. In contrast to existing works, which estimate 3D bounding boxes directly from monocular images or use depth prediction networks to generate input for 3D object detection from 2D information, our method manipulates predictions directly in…

Cited by 899SourcecodeScholar