← Search

Yanyan Liang

15 accepted papers

2026

Distilling Future Temporal Knowledge with Masked Feature Reconstruction for 3D Object Detection

AAAI 2026technical

Camera-based temporal 3D object detection has shown impressive results in autonomous driving, with offline models improving accuracy by using future frames. Knowledge distillation (KD) can be an appealing framework for transferring rich information from offline models to online models. However, exis

Cited by 0SourcePDFScholar
2026

PointCHR: Point Cloud Analysis via Curvature-Aware Hyperbolic Rectification

ICML 2026poster

High-curvature regions in 3D point clouds encapsulate critical fine-grained geometric semantics yet exhibit a distinct long-tail sparsity in their spatial distribution. The inherent limitations of polynomial volume growth in Euclidean space frequently render these intricate geometric features challe…

Cited by 0SourceScholar
2026

PointCSP: Cross-Sample Semantic Propagation and Stability Preservation in Self-Supervised Point Cloud Learning

CVPR 2026

Scene-level point cloud self-supervised learning (PC-SSL) has demonstrated potential in enhancing the generalization capability of 3D vision models. Despite the advances in the field through existing methods, the sample-independent modeling paradigm still poses significant limitations in terms of ma

Cited by 0SourceScholar
2026

PointMC: Multi-view Consistent Encoding and Center-Global Feature Fusion for Point Clouds Understanding

AAAI 2026technical

Point cloud tasks have recently benefited from Mamba-based architecture, which leverage state space modeling to achieve strong performance. Previous studies have primarily focused on network design while overlooking the importance of position encoding and relying on coarse-grained geometric feature

Cited by 0SourcePDFScholar
2025

Not All Frame Features Are Equal: Video-to-4D Generation via Decoupling Dynamic-Static Features

ICCV 2025poster

Recently, the generation of dynamic 3D objects from a video has shown impressive results. Existing methods directly optimize Gaussians using whole information in frames. However, when dynamic regions are interwoven with static regions within frames, particularly if the static regions account for a l…

2025

Stochastic Trajectory Prediction Under Unstructured Constraints

ICRA 2025

Trajectory prediction facilitates effective planning and decision-making, while constrained trajectory prediction integrates regulation into prediction. Recent advances in constrained trajectory prediction focus on structured constraints by constructing optimization objectives. However, handling uns

Cited by 2SourceScholar
2024

CFPL-FAS: Class Free Prompt Learning for Generalizable Face Anti-spoofing

CVPR 2024highlight

Domain generalization (DG) based Face Anti-Spoofing (FAS) aims to improve the model's performance on unseen domains. Existing methods either rely on domain labels to align domain-invariant feature spaces or disentangle generalizable features from the whole sample which inevitably lead to the distort…

Cited by 34SourcePDFScholar
2024

Coevolving with the Other You: Fine-Tuning LLM with Sequential Cooperative Multi-Agent Reinforcement Learning

NeurIPS 2024poster

Reinforcement learning (RL) has emerged as a pivotal technique for fine-tuning large language models (LLMs) on specific tasks. However, prevailing RL fine-tuning methods predominantly rely on PPO and its variants. Though these algorithms are effective in general RL settings, they often exhibit subop…

2023

Gloss-Free Sign Language Translation: Improving from Visual-Language Pretraining

ICCV 2023poster

Sign Language Translation (SLT) is a challenging task due to its cross-domain nature, involving the translation of visual-gestural language to text. Many previous methods employ an intermediate representation,i.e., gloss sequences, to facilitate SLT, thus transforming it into a two-stage task of sig…

Cited by 62PDFcodeScholar
2023

Lazy Agents: A New Perspective on Solving Sparse Reward Problem in Multi-agent Reinforcement Learning

ICML 2023poster

Sparse reward remains a valuable and challenging problem in multi-agent reinforcement learning (MARL). This paper addresses this issue from a new perspective, i.e., lazy agents. We empirically illustrate how lazy agents damage learning from both exploration and exploitation. Then, we propose a novel…

2023

Long-Range Grouping Transformer for Multi-View 3D Reconstruction

ICCV 2023poster

Nowadays, transformer networks have demonstrated superior performance in many computer vision tasks. In a multi-view 3D reconstruction algorithm following this paradigm, self-attention processing has to deal with intricate image tokens including massive information when facing heavy amounts of view…

Cited by 17PDFcodeScholar
2023

Mixture Uniform Distribution Modeling and Asymmetric Mix Distillation for Class Incremental Learning

AAAI 2023technical

Exemplar rehearsal-based methods with knowledge distillation (KD) have been widely used in class incremental learning (CIL) scenarios. However, they still suffer from performance degradation because of severely distribution discrepancy between training and test set caused by the limited storage memo…

Cited by 11SourcePDFScholar
2023

UMIFormer: Mining the Correlations between Similar Tokens for Multi-View 3D Reconstruction

ICCV 2023poster

In recent years, many video tasks have achieved breakthroughs by utilizing the vision transformer and establishing spatial-temporal decoupling for feature extraction. Although multi-view 3D reconstruction also faces multiple images as input, it cannot immediately inherit their success due to complet…

Cited by 14PDFcodeScholar
2022

Decoupling and Recoupling Spatiotemporal Representation for RGB-D-Based Motion Recognition

CVPR 2022poster

Decoupling spatiotemporal representation refers to decomposing the spatial and temporal features into dimension-independent factors. Although previous RGB-D-based motion recognition methods have achieved promising performance through the tightly coupled multi-modal spatiotemporal representation, the…

Cited by 46PDFcodeScholar