← Search

Bisheng Yang

14 accepted papers

2026

CT-FLO: Simple Yet Effective FMCW LiDAR Odometry Using an Linear Continuous-Time Trajectory

RA-L 2026

Frequency-Modulated Continuous-Wave (FMCW) LiDAR is capable of acquiring dense point clouds along with additional Doppler measurements. For discrete-time-based FMCW LiDAR odometry methods, motion distortion correction for both 3D and Doppler measurements should be performed before scan matching. How

Cited by 0SourceScholar
2026

GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting

AAAI 2026technical

3D open-vocabulary scene understanding, which accurately perceives complex semantic properties of objects in space, has gained significant attention in recent years. In this paper, we propose GAGS, a framework that distills 2D CLIP features into 3D Gaussian splatting, enabling open-vocabulary querie

Cited by 0SourcePDFScholar
2026

SCoT: Teaching 3D-LLMs to Think Spatially with Million-scale CoT Annotations

ICLR 2026poster

Recent advances in 3D Large Language Models (3D-LLMs) show strong potential in understanding and interacting with 3D environments, yet their training data typically lack explicit reasoning processes, limiting complex spatial reasoning and task planning. To address this, we annotate SCoT, a million-s…

Cited by 0SourcecodeScholar
2025

CityAnchor: City-scale 3D Visual Grounding with Multi-modality LLMs

ICLR 2025poster

In this paper, we present a 3D visual grounding method called CityAnchor for localizing an urban object in a city-scale point cloud. Recent developments in multiview reconstruction enable us to reconstruct city-scale point clouds but how to conduct visual grounding on such a large-scale urban point…

Cited by 0SourcePDFScholar
2025

Exploiting Motion Prior for Accurate Pose Estimation of Dashboard Cameras

RA-L 2025

Dashboard cameras (dashcams) record millions of driving videos daily, offering a valuable potential data source for various applications, including driving map production and updates. A necessary step for utilizing these dashcam data involves the estimation of camera poses. However, the low-quality

Cited by 1SourceScholar
2025

VistaDream: Sampling multiview consistent images for single-view scene reconstruction

ICCV 2025poster

In this paper, we propose VistaDream, a novel framework to reconstruct a 3D scene from a single-view image. Recent diffusion models enable generating high-quality novel-view images from a single-view input image. Most existing methods only concentrate on building the consistency between the input im…

Cited by 0SourcePDFScholar
2024

CoFiI2P: Coarse-to-Fine Correspondences-Based Image to Point Cloud Registration

RA-L 2024

Image-to-point cloud (I2P) registration is a fundamental task for robots and autonomous vehicles to achieve cross-modality data fusion and localization. Current I2P registration methods primarily focus on estimating correspondences at the point or pixel level, often neglecting global alignment. As a

Cited by 17SourceScholar
2024

Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion

ECCV 2024poster

"∗ Equal contribution Corresponding authorIn this paper, we explore a novel framework, EGIInet (Explicitly Guided Information Interaction Network), a model for View-guided Point cloud Completion (ViPC) task, which aims to restore a complete point cloud from a partial one with a single view image. In…

2024

FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators

ICLR 2024poster

Matching cross-modality features between images and point clouds is a fundamental problem for image-to-point cloud registration. However, due to the modality difference between images and points, it is difficult to learn robust and discriminative cross-modality features by existing metric learning m…

2024

Mobile-Seed: Joint Semantic Segmentation and Boundary Detection for Mobile Robots

RA-L 2024

Precise and rapid delineation of sharp boundaries and robust semantics is essential for numerous downstream robotic tasks, such as robot grasping and manipulation, real-time semantic mapping, and online sensor calibration performed on edge computing units. Although boundary detection and semantic se

Cited by 26SourcecodeScholar
2023

KT-Net: Knowledge Transfer for Unpaired 3D Shape Completion

AAAI 2023technical

Unpaired 3D object completion aims to predict a complete 3D shape from an incomplete input without knowing the correspondence between the complete and incomplete shapes. In this paper, we propose the novel KTNet to solve this task from the new perspective of knowledge transfer. KTNet elaborates a te…

2023

Robust Multiview Point Cloud Registration With Reliable Pose Graph Initialization and History Reweighting

CVPR 2023poster

In this paper, we present a new method for the multiview registration of point cloud. Previous multiview registration methods rely on exhaustive pairwise registration to construct a densely-connected pose graph and apply Iteratively Reweighted Least Square (IRLS) on the pose graph to compute the sca…

2021

AdaFit: Rethinking Learning-Based Normal Estimation on Point Clouds

ICCV 2021poster

This paper presents a neural network for robust normal estimation on point clouds, named AdaFit, that can deal with point clouds with noise and density variations. Existing works use a network to learn point-wise weights for weighted least squares surface fitting to estimate the normals, which has d…

Cited by 56PDFcodeScholar