← Search

Liying Yang

13 accepted papers

2026

Geometry-as-context: Modulating Explicit 3D in Scene-consistent Video Generation to Geometry Context

CVPR 2026

Scene-consistent video generation aims to create videos that explore 3D scenes based on a camera trajectory. Previous methods rely on video generation models with external memory for consistency, or iterative 3D reconstruction and inpainting, which accumulate errors during inference due to incorrect

Cited by 0SourceScholar
2026

PointCHR: Point Cloud Analysis via Curvature-Aware Hyperbolic Rectification

ICML 2026poster

High-curvature regions in 3D point clouds encapsulate critical fine-grained geometric semantics yet exhibit a distinct long-tail sparsity in their spatial distribution. The inherent limitations of polynomial volume growth in Euclidean space frequently render these intricate geometric features challe…

Cited by 0SourceScholar
2026

PointCSP: Cross-Sample Semantic Propagation and Stability Preservation in Self-Supervised Point Cloud Learning

CVPR 2026

Scene-level point cloud self-supervised learning (PC-SSL) has demonstrated potential in enhancing the generalization capability of 3D vision models. Despite the advances in the field through existing methods, the sample-independent modeling paradigm still poses significant limitations in terms of ma

Cited by 0SourceScholar
2025

Dynamic Derivation and Elimination: Audio Visual Segmentation with Enhanced Audio Semantics

CVPR 2025poster

Sound-guided object segmentation has drawn considerable attention for its potential to enhance multimodal perception. Previous methods primarily focus on developing advanced architectures to facilitate effective audio-visual interactions, without fully addressing the inherent challenges posed by aud…

2025

Not All Frame Features Are Equal: Video-to-4D Generation via Decoupling Dynamic-Static Features

ICCV 2025poster

Recently, the generation of dynamic 3D objects from a video has shown impressive results. Existing methods directly optimize Gaussians using whole information in frames. However, when dynamic regions are interwoven with static regions within frames, particularly if the static regions account for a l…

2025

Robust Audio-Visual Segmentation via Audio-Guided Visual Convergent Alignment

CVPR 2025poster

Accurately localizing audible objects based on audio-visual cues is the core objective of audio-visual segmentation. Most previous methods emphasize spatial or temporal multi-modal modeling, yet overlook challenges from ambiguous audio-visual correspondences--such as nearby visually similar but acou…

Cited by 0SourcePDFScholar
2024

Small Multi-Rotor UAV Oriented Direct Thrust Sensor Based on Lightweight Barometers

IROS 2024poster

The multirotor unmanned aerial vehicle (UAV) requires precise control over thrust output when operating in wind-disturbed environments or executing intricate flight missions. Although current commercial force sensors offer high sensitivity and accuracy, they are often heavy and costly. These charact…

Cited by 0SourceScholar
2023

Cooperative Exploration of Heterogeneous UAVs in Mountainous Environments by Constructing Steady Communication

RA-L 2023

Unmanned aerial vehicles (UAVs) must fly at low altitudes to execute certain missions when operating in complex mountainous areas. However, in these environments, UAVs lose their line-of-sight (LOS) communication with the ground station (GS) due to the obstruction of the mountains and are unable to

Cited by 8SourceScholar
2023

Long-Range Grouping Transformer for Multi-View 3D Reconstruction

ICCV 2023poster

Nowadays, transformer networks have demonstrated superior performance in many computer vision tasks. In a multi-view 3D reconstruction algorithm following this paradigm, self-attention processing has to deal with intricate image tokens including massive information when facing heavy amounts of view…

Cited by 17PDFcodeScholar
2023

UMIFormer: Mining the Correlations between Similar Tokens for Multi-View 3D Reconstruction

ICCV 2023poster

In recent years, many video tasks have achieved breakthroughs by utilizing the vision transformer and establishing spatial-temporal decoupling for feature extraction. Although multi-view 3D reconstruction also faces multiple images as input, it cannot immediately inherit their success due to complet…

Cited by 14PDFcodeScholar
2018

Contact Force Control of an Aerial Manipulator in Pressing an Emergency Switch Process

IROS 2018poster

The dangerous work situation in industrial leakage accidents urgently needs a flexible and small robot to help workers perform operations and to protect them from being injured. An aerial manipulator system consisting of a hexa-rotor UAV and a one-DOF manipulator is developed, and is used to press a…

Cited by 44SourceScholar
2018

Grasp a Moving Target from the Air: System & Control of an Aerial Manipulator

ICRA 2018poster

Grasping a moving target has been investigated extensively for fixed-base manipulator. However, such a task becomes much more challenging when the manipulator is free flying in the air with an UAV. Towards moving target grasping, this paper presents an aerial manipulator system composed of a hex-rot…

Cited by 85SourceScholar
2015

Generation of dynamically feasible and collision free trajectory by applying six-order Bezier curve and local optimal reshaping

IROS 2015poster

This paper considers the problem of generating dynamically feasible and collision free trajectory for unmanned aerial vehicles(UAVs) in cluttered environments. General random-based searching algorithms output piecewise linear paths, which cause big discrepancy when used as navigation reference for U…

Cited by 33SourceScholar