← Search

Chao Wen

14 accepted papers

2026

Multi-Strategy Enhanced Particle Swarm Optimization for Variable Curvature Path Planning in Flexible Needle Insertion

ICRA 2026poster

Flexible needles provide enhanced adaptability for navigating puncture pathways and avoiding obstacles when compared to conventional rigid needles. However, developing a three dimensional (3D) curved path for flexible needle is challenging, particularly in achieving both effective obstacle avoidance…

Cited by 0SourceScholar
2025

Consensus Graph Filter Learning for Multiple Graph Clustering

ICASSP 2025accepted

Multi-view Clustering (MVC) has gained significant attention for its ability to utilize consistent and complementary information from multiple views. Graph filter-based MVC methods have recently demonstrated promising performance, attracting growing interest. However, existing graph filter-based met…

Cited by 0SourceScholar
2025

Multi-Strategy Enhanced Particle Swarm Optimization for Variable Curvature Path Planning in Flexible Needle Insertion

RA-L 2025

Flexible needles provide enhanced adaptability for navigating puncture pathways and avoiding obstacles when compared to conventional rigid needles. However, developing a three dimensional (3D) curved path for flexible needle is challenging, particularly in achieving both effective obstacle avoidance

Cited by 0SourceScholar
2025

Precise Action-to-Video Generation Through Visual Action Prompts

ICCV 2025poster

We present visual action prompts, a unified action representation for action-to-video generation of complex high-DoF interactions while maintaining transferable visual dynamics across domains. Action-driven video generation faces a precision-generality tradeoff: existing methods using text, primitiv…

Cited by 0SourcePDFScholar
2025

Program Synthesis Benchmark for Visual Programming in XLogoOnline Environment

ACL 2025long

Large language and multimodal models have shown remarkable success on various benchmarks focused on specific skills such as general-purpose programming, math word problem-solving, and visual question answering. However, it is unclear how well these models perform on tasks that require a combination…

Cited by 0SourcePDFScholar
2024

Joint2Human: High-Quality 3D Human Generation via Compact Spherical Embedding of 3D Joints

CVPR 2024poster

3D human generation is increasingly significant in various applications. However the direct use of 2D generative methods in 3D generation often results in losing local details while methods that reconstruct geometry from generated images struggle with global view consistency. In this work we introdu…

Cited by 6SourcePDFScholar
2024

OHTA: One-shot Hand Avatar via Data-driven Implicit Priors

CVPR 2024poster

In this paper we delve into the creation of one-shot hand avatars attaining high-fidelity and drivable hand representations swiftly from a single image. With the burgeoning domains of the digital human the need for quick and personalized hand avatar creation has become increasingly critical. Existin…

Cited by 9SourcePDFScholar
2023

Decoupled Iterative Refinement Framework for Interacting Hands Reconstruction from a Single RGB Image

ICCV 2023oral

Reconstructing interacting hands from a single RGB image is a very challenging task. On the one hand, severe mutual occlusion and similar local appearance between two hands confuse the extraction of visual features, resulting in the misalignment of estimated hand meshes and the image. On the other h…

Cited by 29PDFcodeScholar
2023

HaMuCo: Hand Pose Estimation via Multiview Collaborative Self-Supervised Learning

ICCV 2023poster

Recent advancements in 3D hand pose estimation have shown promising results, but its effectiveness has primarily relied on the availability of large-scale annotated datasets, the creation of which is a laborious and costly process. To alleviate the label-hungry limitation, we propose a self-supervis…

Cited by 16PDFcodeScholar
2023

Realistic Full-Body Tracking from Sparse Observations via Joint-Level Modeling

ICCV 2023poster

To bridge the physical and virtual worlds for rapidly developed VR/AR applications, the ability to realistically drive 3D full-body avatars is of great significance. Although real-time body tracking with only the head-mounted displays (HMDs) and hand controllers is heavily under-constrained, a caref…

Cited by 28PDFcodeScholar
2022

3D Room Layout Estimation from a Cubemap of Panorama Image via Deep Manhattan Hough Transform

ECCV 2022poster

"Significant geometric structures can be compactly described by global wireframes in the estimation of 3D room layout from a single panoramic image. Based on this observation, we present an alternative approach to estimate the walls in 3D space by modeling long-range geometric patterns in a learnabl…

2022

Needle Tip Pose Estimation for Ultrasound-Guided Steerable Flexible Needle With a Complicated Trajectory in Soft Tissue

RA-L 2022

Visualization of the surgical needle is critical for an image guided insertion. However, it is difficult to obtain a 3D tip pose of a steerable flexible needle under 2D US images. In this letter, an image processing method is used to extract the needle shaft radial cross-sectional centroid coordinat

Cited by 8SourceScholar
2020

Neural Pose Transfer by Spatially Adaptive Instance Normalization

CVPR 2020poster

Pose transfer has been studied for decades, in which the pose of a source mesh is applied to a target mesh. Particularly in this paper, we are interested in transferring the pose of source human mesh to deform the target human mesh, while the source and target meshes may have different identity info…

Cited by 71PDFcodeScholar