← Search

Yawen Lu

13 accepted papers

2025

Event-Driven Force Measurement of a Variable-Stiffness Robotic Finger Using Masked Autoencoder Pre-Training

RA-L 2025

Force feedback in compliant robotic grippers is essential for precise manipulation tasks but remains challenging because irregular deformation of soft materials renders traditional sensor integration impractical. Event camera has emerged as a powerful alternative to conventional vision sensors throu

Cited by 1SourceScholar
2024

DL3DV-10K: A Large-Scale Scene Dataset for Deep Learning-based 3D Vision

CVPR 2024poster

We have witnessed significant progress in deep learning-based 3D vision ranging from neural radiance field (NeRF) based 3D representation learning to applications in novel view synthesis (NVS). However existing scene-level datasets for deep learning-based 3D vision limited to either synthetic enviro…

Cited by 85SourcePDFScholar
2024

ProMotion: Prototypes As Motion Learners

CVPR 2024poster

In this work we introduce ProMotion a unified prototypical transformer-based framework engineered to model fundamental motion tasks. ProMotion offers a range of compelling attributes that set it apart from current task-specific paradigms. 1. We adopt a prototypical perspective establishing a unified…

Cited by 7SourcePDFScholar
2024

Prototypical Transformer As Unified Motion Learners

ICML 2024poster

In this work, we introduce the Prototypical Transformer (ProtoFormer), a general and unified framework that approaches various motion tasks from a prototype perspective. ProtoFormer seamlessly integrates prototype learning with Transformer by thoughtfully considering motion dynamics, introducing two…

Cited by 17SourcePDFScholar
2023

TransFlow: Transformer As Flow Learner

CVPR 2023highlight

Optical flow is an indispensable building block for various important computer vision tasks, including motion estimation, object tracking, and disparity measurement. In this work, we propose TransFlow, a pure transformer architecture for optical flow estimation. Compared to dominant CNN-based method…

Cited by 99SourcePDFScholar
2022

From Local to Holistic: Self-supervised Single Image 3D Face Reconstruction Via Multi-level Constraints

IROS 2022poster

Single image 3D face reconstruction with accurate geometric details is a critical and challenging task due to the similar appearance on the face surface and fine details in organs. In this work, we introduce a self-supervised 3D face reconstruction approach from a single image that can recover detai…

Cited by 5SourceScholar
2022

Inferring Camera Intrinsics Based on Surfaces of Revolution: A Single Image Geometric Network Approach for Camera Calibration

ICASSP 2022accepted

Camera calibration is a necessary prerequisite in many applications of robotics, especially in robot vision in order to obtain metric reconstruction from a 2D image. In this paper, we address the problem of calibrating from a single image of a surface of revolution (SOR) based on deep learning, in o…

Cited by 0SourceScholar
2021

Matching as Color Images: Thermal Image Local Feature Detection and Description

ICASSP 2021accepted

Feature detection and extraction is considered to be one of the most important aspects when it comes to any computer vision application, especially the autonomous driving field that is highly dependent on it. Thermal imaging is less explored in the field of autonomous driving mainly due to the high…

Cited by 0SourceScholar
2020

Multi-Task Learning for Single Image Depth Estimation and Segmentation Based on Unsupervised Network

ICRA 2020poster

Deep neural networks have significantly enhanced the performance of various computer vision tasks, including single image depth estimation and image segmentation. However, most existing approaches handle them in supervised manners and require a large number of ground truth labels that consume extens…

Cited by 22SourceScholar