← Search

Xiaoming Liu

134 accepted papers

2026

EmoTaG: Emotion-Aware Talking Head Synthesis on Gaussian Splatting with Few-Shot Personalization

CVPR 2026

Audio-driven 3D talking head synthesis has advanced rapidly with Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS). By leveraging rich pre-trained priors, few-shot methods enable instant personalization from just a few seconds of video. However, under expressive facial motion, existing

Cited by 0SourceScholar
2026

FusionAgent: A Multimodal Agent with Dynamic Model Selection for Human Recognition

CVPR 2026

Model fusion is a key strategy for robust recognition in unconstrained scenarios, as different models provide complementary strengths. This is especially important for whole-body human recognition, where biometric cues such as face, gait, and body shape vary across samples and are typically integrat

Cited by 7SourcecodeScholar
2026

MGT-Prism: Enhancing Domain Generalization for Machine-Generated Text Detection via Spectral Alignment

AAAI 2026technical

Large Language Models have shown growing ability to generate fluent and coherent texts that are highly similar to the writing style of humans. Current detectors for Machine-Generated Text (MGT) perform well when they are trained and tested in the same domain but generalize poorly to unseen domains,

Cited by 0SourcePDFScholar
2026

Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection

CVPR 2026

Monocular 3D detection (Mono3D) aims to infer 3D bounding boxes from a single RGB image.Without auxiliary sensors such as LiDAR, this task is inherently ill-posed since the 3D-to-2D projection introduces depth ambiguity.Previous works often predict 3D attributes (e.g., depth, size, and orientation)

Cited by 0SourcecodeScholar
2026

Unlocking Motion from Large Vision Models with a Semantic and Kinematic Duality for Gait Recognition

CVPR 2026

Existing set-based gait recognition methods achieve remarkable performance by capturing global semantic context.However, their order-invariant nature prevents them from modeling the fine-grained kinematic patterns that unfold over time.To unify the global and process-level representations, we propos

Cited by 0SourceScholar
2025

A Quality-Guided Mixture of Score-Fusion Experts Framework for Human Recognition

ICCV 2025poster

Whole-body biometric recognition is a challenging multi-modal task that integrates various biometric modalities, including face, gait, and body. This integration is essential for overcoming the limitations of unimodal systems. Traditionally, whole-body recognition involves deploying different models…

2025

BiggerGait: Unlocking Gait Recognition with Layer-wise Representations from Large Vision Models

NeurIPS 2025poster

Large vision models (LVM) based gait recognition has achieved impressive performance. However, existing LVM-based approaches may overemphasize gait priors while neglecting the intrinsic value of LVM itself, particularly the rich, distinct representations across its multi-layers. To adequately unloc…

Cited by 0SourcecodeScholar
2025

CHARM3R: Towards Unseen Camera Height Robust Monocular 3D Detector

ICCV 2025poster

Monocular 3D object detectors, while effective on data from one ego camera height, struggle with unseen or out-of-distribution camera heights. Existing methods often rely on Plucker embeddings, image transformations or data augmentation. This paper takes a step towards this understudied problem by i…

2025

Compact R-X-Y Stage and Dual-Finger Micromanipulator under Inverted Optical Microscope for Microassembly

IROS 2025

Microassembly plays an important role in fabricating complex structures with small basic components in industrial and biomedical fields. Inverted optical microscope could provide high-quality image feedback for microassembly with its continuously improving resolution. However, a compact stage capabl

Cited by 0SourceScholar
2025

Contactless and Economical Chemical Reaction Platform Based on Ultrasonic Field

IROS 2025

Chemical reactions constitute a cornerstone of fundamental scientific inquiry, yet traditional methodologies and platforms are encumbered by excessive reagent and consumable demands. Emerging alternatives, such as microfluidic systems, while innovative, suffer from intricate fabrication processes an

Cited by 0SourceScholar
2025

DMPBot: A high-speed, high-precision, omnidirectional, insect-scale piezoelectric robot

IROS 2025

Microrobots have garnered significant attention due to their vast potential applications across various fields. Among various types of microrobots, piezoelectric robots stand out due to their exceptional motion accuracy, low power consumption, and simple structural design. This work introduces a nov

Cited by 0SourceScholar
2025

Dual-Bubble Coordinated Acoustic Micromanipulator for Multidirectional Object Rotation*

IROS 2025

Micromanipulation techniques struggle to achieve three-dimensional rotational control at the microscale without compromising biocompatibility or spatial flexibility. Conventional methods based on mechanical contact, optical forces, or confined microfluidics constrain dynamic reconfiguration and surg

Cited by 0SourceScholar
2025

Enhanced Rolling Motion of Magnetic Microparticles by Turning Interface Lubrication

IROS 2025

Micro-nano robots must break the symmetry of the flow field to generate net displacement in the low Reynolds number environment. The spherical micro-robots utilize the frictional forces generated through interaction with the surface. We designed a magnetic microroller robot powered by the rotating A

Cited by 0SourceScholar
2025

H-MoRe: Learning Human-centric Motion Representation for Action Analysis

CVPR 2025highlight

In this paper, we propose H-MoRe, a novel pipeline for learning precise human-centric motion representation. Our approach dynamically preserves relevant human motion while filtering out background movement. Notably, unlike previous methods relying on fully supervised learning from synthetic data, H-…

2025

HAMoBE: Hierarchical and Adaptive Mixture of Biometric Experts for Video-based Person ReID

ICCV 2025poster

Recently, research interest in person re-identification (ReID) has increasingly focused on video-based scenarios, essential for robust surveillance and security in varied and dynamic environments. However, existing video-based ReID methods often overlook the necessity of identifying and selecting th…

Cited by 0SourcePDFScholar
2025

Iron Sharpens Iron: Defending Against Attacks in Machine-Generated Text Detection with Adversarial Training

ACL 2025long

Machine-generated Text (MGT) detection is crucial for regulating and attributing online texts. While the existing MGT detectors achieve strong performance, they remain vulnerable to simple perturbations and adversarial attacks. To build an effective defense against malicious perturbations, we view M…

2025

Magnetically Actuated Steerable Catheter with Redundant DoF for Cardiovascular Interventions

IROS 2025

A magnetically controlled catheter system is proposed to enhance the precision and safety of vascular interventions by reducing procedure time and radiation exposure. The system can also function as a support channel for guidewire deployment. A novel navigation approach is introduced, employing an e

Cited by 0SourceScholar
2025

On-Chip Dynamic Mechanical Characterization: from Cells to Nucleus

IROS 2025

Traditional single-cell mechanical characterization techniques (e.g., atomic force microscopy) often face limitations in throughput, require invasive labeling, or fail to replicate physiological microenvironments, impeding their clinical utility for rapid cancer cell analysis. To address these limit

Cited by 0SourceScholar
2025

RICCARDO: Radar Hit Prediction and Convolution for Camera-Radar 3D Object Detection

CVPR 2025poster

Radar hits reflect from points on both the boundary and internal to object outlines. This results in a complex distribution of radar hits that depends on factors including object category, size and orientation. Current radar-camera fusion methods implicitly account for this with a black-box neural n…

2025

Rethinking Vision-Language Model in Face Forensics: Multi-Modal Interpretable Forged Face Detector

CVPR 2025poster

Deepfake detection is a long-established research topic vital for mitigating the spread of malicious misinformation. Unlike prior methods that provide either binary classification results or textual explanations separately, we introduce a novel method capable of generating both simultaneously. Our m…

2024

Acoustically Driven Micropipette for Hydrodynamic Manipulation of Mouse Oocytes

ICRA 2024poster

Micromanipulation techniques that can achieve controlled fine operations at the micro scale play an important role in biomedical fields including embryo engineering, gene engineering, drug screening, and cell analysis. However, micromanipulation of biological micro-objects, such as cells and micro t…

Cited by 0SourceScholar
2024

Automated Assembly by Two-Fingered Microhand for Fabrication of Soft Magnetic Microrobots

ICRA 2024poster

Micro-assembly is an emerging method to fabricate microrobots with multiple modules or particles. However, there is always a lack of a flexible and efficient method to freely create the desired magnetic soft microrobots. In this paper, an automated assembly system based on a two-fingered microhand i…

Cited by 1SourceScholar
2024

BigGait: Learning Gait Representation You Want by Large Vision Models

CVPR 2024poster

Gait recognition stands as one of the most pivotal remote identification technologies and progressively expands across research and industry communities. However existing gait recognition methods heavily rely on task-specific upstream driven by supervised learning to provide explicit gait representa…

2024

COMPOSE: Comprehensive Portrait Shadow Editing

ECCV 2024poster

"Existing portrait relighting methods struggle with precise control over facial shadows, particularly when faced with challenges such as handling hard shadows from directional light sources or adjusting shadows while remaining in harmony with existing lighting conditions. In many situations, complet…

Cited by 3SourcePDFScholar
2024

Concentrate Attention: Towards Domain-Generalizable Prompt Optimization for Language Models

NeurIPS 2024poster

Recent advances in prompt optimization have notably enhanced the performance of pre-trained language models (PLMs) on downstream tasks. However, the potential of optimized prompts on domain generalization has been under-explored. To explore the nature of prompt generalization on unknown domains, we…

2024

Development of a 3-RRS Micromanipulator Based on Origami-Inspired Spherical Joint

ICRA 2024poster

In recent years, micromanipulation technology has achieved extensive applications in industry and life science. Improving the precision and bandwidth of the micromanipulator and simultaneously reducing size, weight, and cost pose significant challenges to the existing micromanipulator design and fab…

Cited by 0SourceScholar
2024

Dialogue for Prompting: A Policy-Gradient-Based Discrete Prompt Generation for Few-Shot Learning

AAAI 2024technical

Prompt-based pre-trained language models (PLMs) paradigm has succeeded substantially in few-shot natural language processing (NLP) tasks. However, prior discrete prompt optimization methods require expert knowledge to design the base prompt set and identify high-quality prompts, which is costly, ine…

2024

Distilling CLIP with Dual Guidance for Learning Discriminative Human Body Shape Representation

CVPR 2024poster

Person Re-Identification (ReID) holds critical importance in computer vision with pivotal applications in public safety and crime prevention. Traditional ReID methods reliant on appearance attributes such as clothing and color encounter limitations in long-term scenarios and dynamic environments. To…

Cited by 12SourcePDFScholar
2024

Does DetectGPT Fully Utilize Perturbation? Bridging Selective Perturbation to Fine-tuned Contrastive Learning Detector would be Better

ACL 2024long

The burgeoning generative capabilities of large language models (LLMs) have raised growing concerns about abuse, demanding automatic machine-generated text detectors. DetectGPT, a zero-shot metric-based detector, first introduces perturbation and shows great performance improvement. However, in Dete…

2024

KeyPoint Relative Position Encoding for Face Recognition

CVPR 2024poster

In this paper we address the challenge of making ViT models more robust to unseen affine transformations. Such robustness becomes useful in various recognition tasks such as face recognition when image alignment failures occur. We propose a novel method called KP-RPE which leverages key points (e.g.…

2024

On Learning Multi-Modal Forgery Representation for Diffusion Generated Video Detection

NeurIPS 2024poster

Large numbers of synthesized videos from diffusion models pose threats to information security and authenticity, leading to an increasing demand for generated content detection. However, existing video-level detection algorithms primarily focus on detecting facial forgeries and often fail to identif…

2024

ProMark: Proactive Diffusion Watermarking for Causal Attribution

CVPR 2024poster

Generative AI (GenAI) is transforming creative workflows through the capability to synthesize and manipulate images via high-level prompts. Yet creatives are not well supported to receive recognition or reward for the use of their content in GenAI training. To this end we propose ProMark a causal at…

Cited by 14SourcePDFScholar
2024

Remove Projective LiDAR Depthmap Artifacts via Exploiting Epipolar Geometry

ECCV 2024poster

"sensing is a fundamental task for Autonomous Vehicles. Its deployment often relies on aligned RGB cameras and . Despite meticulous synchronization and calibration, systematic misalignment persists in projected . This is due to the physical baseline distance between the two sensors. The artifact is…

Cited by 0SourcePDFScholar
2024

SeaBird: Segmentation in Bird's View with Dice Loss Improves Monocular 3D Detection of Large Objects

CVPR 2024poster

Monocular 3D detectors achieve remarkable performance on cars and smaller objects. However their performance drops on larger objects leading to fatal accidents. Some attribute the failures to training data scarcity or the receptive field requirements of large objects. In this paper we highlight this…

2024

StablePT : Towards Stable Prompting for Few-shot Learning via Input Separation

EMNLP 2024finding

Large language models have shown their ability to become effective few-shot learners with prompting, revoluting the paradigm of learning with data scarcity. However, this approach largely depends on the quality of prompt initialization and always exhibits large variability among different runs. Such…

2024

Stumbling Blocks: Stress Testing the Robustness of Machine-Generated Text Detectors Under Attacks

ACL 2024long

The widespread use of large language models (LLMs) is increasing the demand for methods that detect machine-generated text to prevent misuse. The goal of our study is to stress test the detectors’ robustness to malicious attacks under realistic scenarios. We comprehensively study the robustness of p…

2024

TIGER: Time-Varying Denoising Model for 3D Point Cloud Generation with Diffusion Process

CVPR 2024poster

Recently diffusion models have emerged as a new powerful generative method for 3D point cloud generation tasks. However few works study the effect of the architecture of the diffusion model in the 3D point cloud resorting to the typical UNet model developed for 2D images. Inspired by the wide adopti…

2024

Unified Physical-Digital Face Attack Detection

IJCAI 2024poster

Face Recognition (FR) systems can suffer from physical (i.e., print photo) and digital (i.e., DeepFake) attacks. However, previous related work rarely considers both situations at the same time. This implies the deployment of multiple models and thus more computational burden. The main reasons for t…

Cited by 15SourcePDFScholar
2024

UnlearnCanvas: Stylized Image Dataset for Enhanced Machine Unlearning Evaluation in Diffusion Models

NeurIPS 2024poster

The technological advancements in diffusion models (DMs) have demonstrated unprecedented capabilities in text-to-image generation and are widely used in diverse applications. However, they have also raised significant societal concerns, such as the generation of harmful content and copyright dispute…

2023

ChatGPT-Powered Hierarchical Comparisons for Image Classification

NeurIPS 2023poster

The zero-shot open-vocabulary setting poses challenges for image classification. Fortunately, utilizing a vision-language model like CLIP, pre-trained on image-text pairs, allows for classifying images by comparing embeddings. Leveraging large language models (LLMs) such as ChatGPT can further enhan…

2023

DCFace: Synthetic Face Generation With Dual Condition Diffusion Model

CVPR 2023poster

Generating synthetic datasets for training face recognition models is challenging because dataset generation entails more than creating high fidelity images. It involves generating multiple images of same subjects under different factors (e.g., variations in pose, illumination, expression, aging and…

2023

Hierarchical Fine-Grained Image Forgery Detection and Localization

CVPR 2023poster

Differences in forgery attributes of images generated in CNN-synthesized and image-editing domains are large, and such differences make a unified image forgery detection and localization (IFDL) challenging. To this end, we present a hierarchical fine-grained formulation for IFDL representation learn…

2023

Learning Clothing and Pose Invariant 3D Shape Representation for Long-Term Person Re-Identification

ICCV 2023poster

Long-Term Person Re-Identification (LT-ReID) has become increasingly crucial in computer vision and biometrics. In this work, we aim to extend LT-ReID beyond pedestrian recognition to include a wider range of real-world human activities while still accounting for cloth-changing scenarios over large…

Cited by 39PDFScholar
2023

LightedDepth: Video Depth Estimation in Light of Limited Inference View Angles

CVPR 2023poster

Video depth estimation infers the dense scene depth from immediate neighboring video frames. While recent works consider it a simplified structure-from-motion (SfM) problem, it still differs from the SfM in that significantly fewer view angels are available in inference. This setting, however, suits…

2023

MaLP: Manipulation Localization Using a Proactive Scheme

CVPR 2023poster

Advancements in the generation quality of various Generative Models (GMs) has made it necessary to not only perform binary manipulation detection but also localize the modified pixels in an image. However, prior works termed as passive for manipulation localization exhibit poor generalization perfor…

2023

Programable On-Chip Fabrication of Magnetic Soft Micro-Robot

IROS 2023poster

In the last decade, researchers have been trying to develop many microrobots that mimic the extraordinary abilities of bionts in complex environments. How to fabricate the biomimetic microrobot with satisfying deformability and complex shapes to realize desired precise motion is the key issue. In th…

Cited by 0SourceScholar
2023

RADIANT: Radar-Image Association Network for 3D Object Detection

AAAI 2023technical

As a direct depth sensor, radar holds promise as a tool to improve monocular 3D object detection, which suffers from depth errors, due in part to the depth-scale ambiguity. On the other hand, leveraging radar depths is hampered by difficulties in precisely associating radar returns with 3D estimates…

2023

Rethinking Domain Generalization for Face Anti-Spoofing: Separability and Alignment

CVPR 2023poster

This work studies the generalization issue of face anti-spoofing (FAS) models on domain gaps, such as image resolution, blurriness and sensor variations. Most prior works regard domain-specific signals as a negative impact, and apply metric learning or adversarial losses to remove it from feature re…

2023

Tame a Wild Camera: In-the-Wild Monocular Camera Calibration

NeurIPS 2023poster

3D sensing for monocular in-the-wild images, e.g., depth estimation and 3D object detection, has become increasingly important. However, the unknown intrinsic parameter hinders their development and deployment. Previous methods for the monocular camera calibration rely on specific 3D objects or stro…

2022

A PZT-Driven 6-DOF High-Speed Micromanipulator for Circular Vibration Simulation and Whirling Flow Generation

RA-L 2022

Existing micromanipulation methods, whether contact micromanipulation or non-contact micromanipulation, can hardly meet the requirements of low damage and multiple functions in the biomedical field. This study provides a high-speed micromanipulator that can simulate circular vibrations and generate

Cited by 2SourceScholar
2022

Cluster and Aggregate: Face Recognition with Large Probe Set

NeurIPS 2022accept

Feature fusion plays a crucial role in unconstrained face recognition where inputs (probes) comprise of a set of $N$ low quality images whose individual qualities vary. Advances in attention and recurrent modules have led to feature fusion that can model the relationship among the images in the inpu…

2022

Controllable and Guided Face Synthesis for Unconstrained Face Recognition

ECCV 2022poster

"Although significant advances have been made in face recognition (FR), FR in unconstrained environments remains challenging due to the domain gap between the semi-constrained training datasets and unconstrained testing scenarios. To address this problem, we propose a controllable face synthesis mod…

Cited by 46SourcePDFScholar
2022

Controlled Fabrication of Micro-Chain Robot Using Magnetically Guided Arraying Microfluidic Devices

IROS 2022poster

The magnetic microrobot has become a promising approach in many biomedical applications due to its small volume, flexible motion, and untethered micromachines. The micro-chain robot is one of the most popular magnetic microrobots. However, the uncontrollable magnetic moment direction and quantity of…

Cited by 0SourceScholar
2022

DEVIANT: Depth EquiVarIAnt NeTwork for Monocular 3D Object Detection

ECCV 2022poster

"Modern neural networks use building blocks such as convolutions that are equivariant to arbitrary 2D translations. However, these vanilla blocks are not equivariant to arbitrary 3D translations in the projective manifold. Even then, all monocular 3D detectors use vanilla blocks to obtain the 3D coo…

2022

Face Relighting With Geometrically Consistent Shadows

CVPR 2022poster

Most face relighting methods are able to handle diffuse shadows, but struggle to handle hard shadows, such as those cast by the nose. Methods that propose techniques for handling hard shadows often do not produce geometrically consistent shadows since they do not directly leverage the estimated face…

Cited by 52PDFcodeScholar
2022

MOST-GAN: 3D Morphable StyleGAN for Disentangled Face Image Manipulation

AAAI 2022technical

Recent advances in generative adversarial networks (GANs) have led to remarkable achievements in face image synthesis. While methods that use style-based GANs can generate strikingly photorealistic face images, it is often difficult to control the characteristics of the generated faces in a meaningf…

Cited by 35SourcePDFScholar
2022

Multi-Domain Learning for Updating Face Anti-Spoofing Models

ECCV 2022poster

"In this work, we study multi-domain learning for face anti-spoofing (MD-FAS), where a pre-trained FAS model needs to be updated to perform equally well on both source and target domains while only using target domain data for updating. We present a new model for MD-FAS, which addresses the forgetti…

2022

On-Chip Automatic Trapping and Rotating for Zebrafish Embryo Injection

RA-L 2022

Zebrafish embryo injection is often required in biomedical research using zebrafish. In the injecting operation, trapping and rotating the zebrafish embryo to achieve a proper posture is essential for the high success rate. We proposed an on-chip platform capable of efficient and automatic trapping

Cited by 8SourceScholar
2022

Reverse Engineering of Imperceptible Adversarial Image Perturbations

ICLR 2022poster

It has been well recognized that neural network based image classifiers are easily fooled by images with tiny perturbations crafted by an adversary. There has been a vast volume of research to generate and defend such adversarial attacks. However, the following problem is left unexplored: How to rev…

2021

Efficient Single-Cell Mechanical Measurement by Integrating a Cell Arraying Microfluidic Device With Magnetic Tweezer

RA-L 2021

Cell stiffness is an essential label-free biomarker used to diagnose and sort cells at the single-cell level. Here, we integrated magnetic tweezers on an efficient cell arraying microfluidic device to evaluate the mechanical properties of single cells. Two motion modes under pulsed electromagnetic f

Cited by 20SourceScholar
2021

Full-Velocity Radar Returns by Radar-Camera Fusion

ICCV 2021poster

A distinctive feature of Doppler radar is the measurement of velocity in the radial direction for radar points. However, the missing tangential velocity component hampers object velocity estimation as well as temporal integration of radar sweeps in dynamic scenes. Recognizing that fusing camera with…

Cited by 28PDFScholar
2021

GrooMeD-NMS: Grouped Mathematically Differentiable NMS for Monocular 3D Object Detection

CVPR 2021poster

Modern 3D object detectors have immensely benefited from the end-to-end learning idea. However, most of them use a post-processing algorithm called Non-Maximal Suppression (NMS) only during inference. While there were attempts to include NMS in the training pipeline for tasks such as 2D object detec…

Cited by 110PDFcodeScholar
2021

In-Situ Bonding of Multi-Layer Microfluidic Devices Assisted by an Automated Alignment System

RA-L 2021

Three-dimensional multi-layer microfluidic device (MMD) fabricated by polydimethylsiloxane (PDMS) is a solution for on-chip high-complexity serial or parallel processes. In this letter, we propose and set up an automated 8-DOF alignment system assisted by computer vision, which is capable of automat

Cited by 1SourceScholar
2021

Radar-Camera Pixel Depth Association for Depth Completion

CVPR 2021poster

While radar and video data can be readily fused at the detection level, fusing them at the pixel level is potentially more beneficial. This is also more challenging in part due to the sparsity of radar, but also because automotive radar beams are much wider than a typical pixel combined with a large…

Cited by 92PDFcodeScholar
2021

Towards High Fidelity Face Relighting With Realistic Shadows

CVPR 2021poster

Existing face relighting methods often struggle with two problems: maintaining the local facial details of the subject and accurately removing and synthesizing shadows in the relit image, especially hard shadows. We propose a novel deep face relighting method that addresses both problems. Our method…

Cited by 65PDFcodeScholar
2020

Automated Tracking System with Head and Tail Recognition for Time-Lapse Observation of Free-Moving C. elegans

ICRA 2020poster

In this paper, an automated tracking system with head and tail recognition for time-lapse observation of free-moving C. elegans is presented. In microscale field, active C. elegans can move out of the view easily without an automated tracking system because of the narrow field of view and rapid spee…

Cited by 4SourceScholar
2020

CurricularFace: Adaptive Curriculum Learning Loss for Deep Face Recognition

CVPR 2020poster

As an emerging topic in face recognition, designing margin-based loss functions can increase the feature margin between different classes for enhanced discriminability. More recently, the idea of mining-based strategies is adopted to emphasize the misclassified samples, achieving promising results.…

Cited by 686PDFcodeScholar
2020

Improving Face Recognition from Hard Samples via Distribution Distillation Loss

ECCV 2020poster

Large facial variations are the main challenge in face recognition. To this end, previous variation-specific methods make full use of task-related prior to design special network losses, which are typically not general among different tasks and scenarios. In contrast, the existing generic methods fo…

2020

Jointly De-biasing Face Recognition and Demographic Attribute Estimation

ECCV 2020poster

We address the problem of bias in automated face recognition and demographic attribute estimation algorithms, where errors are lower on certain cohorts belonging to specific demographic groups. We present a novel de-biasing adversarial network (DebFace) that learns to extract disentangled feature re…

2020

LUVLi Face Alignment: Estimating Landmarks' Location, Uncertainty, and Visibility Likelihood

CVPR 2020poster

Modern face alignment methods have become quite accurate at predicting the locations of facial landmarks, but they do not typically estimate the uncertainty of their predicted locations nor predict whether landmarks are visible. In this paper, we present a novel framework for jointly predicting land…

Cited by 196PDFcodeScholar
2020

Noise Modeling, Synthesis and Classification for Generic Object Anti-Spoofing

CVPR 2020poster

Using printed photograph and replaying videos of biometric modalities, such as iris, fingerprint and face, are common attacks to fool the recognition systems for granting access as the genuine user. With the growing online person-to-person shopping (e.g., Ebay and Craigslist), such attacks also thre…

Cited by 39PDFScholar
2020

On Disentangling Spoof Trace for Generic Face Anti-Spoofing

ECCV 2020poster

Prior studies show that the key to face anti-spoofing lies in the subtle image pattern, termed “spoof trace”, e.g., color distortion, 3D mask edge, Moir´e pattern, and many others. Designing a generic anti-spoofing model to estimate those spoof traces can improve not only the generalization of the s…

2019

Automatic Cell Assembly by Two-fingered Microhand

IROS 2019poster

We have successfully achieved manipulation and assembly of microbeads having the size of 100μm diameter by hemispherical end-effectors with high stability and accuracy. The motivation of achieving assembly of actual cells lies in the great significance of it in tissue regeneration and cell analysis.…

Cited by 0SourceScholar
2019

Feature Transfer Learning for Face Recognition With Under-Represented Data

CVPR 2019poster

Despite the large volume of face recognition datasets, there is a significant portion of subjects, of which the samples are insufficient and thus under-represented. Ignoring such significant portion results in insufficient training data. Training with under-represented data leads to biased classifie…

Cited by 396PDFScholar
2019

Gait Recognition via Disentangled Representation Learning

CVPR 2019oral

Gait, the walking pattern of individuals, is one of the most important biometrics modalities. Most of the existing gait recognition methods take silhouettes or articulated body models as the gait features. These methods suffer from degraded recognition performance when handling confounding variables…

Cited by 324PDFScholar
2019

Gotta Adapt 'Em All: Joint Pixel and Feature-Level Domain Adaptation for Recognition in the Wild

CVPR 2019poster

Recent developments in deep domain adaptation have allowed knowledge transfer from a labeled source domain to an unlabeled target domain at the level of intermediate features or input pixels. We propose that advantages may be derived by combining them, in the form of different insights that lead to…

Cited by 54PDFScholar
2018

Disentangling Features in 3D Face Shapes for Joint Face Reconstruction and Recognition

CVPR 2018poster

This paper proposes an encoder-decoder network to disentangle shape features during 3D face shape reconstruction from single 2D images, such that the tasks of learning discriminative shape features for face recognition and reconstructing accurate 3D face shapes can be done simultaneously. Unlike exi…

Cited by 129SourcePDFScholar
2018

FSRNet: End-to-End Learning Face Super-Resolution With Facial Priors

CVPR 2018poster

Face Super-Resolution (SR) is a domain-specific superresolution problem. The facial prior knowledge can be leveraged to better super-resolve face images. We present a novel deep end-to-end trainable Face Super-Resolution Network (FSRNet), which makes use of the geometry prior, i.e., facial landmark…

2018

Learning Deep Models for Face Anti-Spoofing: Binary or Auxiliary Supervision

CVPR 2018poster

Face anti-spoofing is crucial to prevent face recognition systems from a security breach. Previous deep learning approaches formulate face anti-spoofing as a binary classification problem. Many of them struggle to grasp adequate spoofing cues and generalize poorly. In this paper, we argue the impor…

Cited by 789SourcePDFScholar
2018

Resistive Pulse Study of Liposome Stability: Towards Precision and Efficient Drug Delivery

IROS 2018poster

In this work, the authors report the investigation of liposomes' stability as a drug delivery vehicle, using the method of resistive pulse method. The main objects of interest are the 50nm diameter liposomes, while the 100nm diameter liposomes are widely used for its stability. However, certain drug…

Cited by 0SourceScholar
2017

Monocular Video-Based Trailer Coupler Detection Using Multiplexer Convolutional Neural Network

ICCV 2017poster

This paper presents an automated monocular-camera-based computer vision system for autonomous self-backing-up a vehicle towards a trailer, by continuously estimating the 3D trailer coupler position and feeding it to the vehicle control system, until the alignment of the tow hitch with the trailers c…

Cited by 21PDFScholar
2017

Non-contact transportation and rotation of micro objects by vibrating glass needle circularly under water

ICRA 2017poster

In micromanipulation, lots of methods have been developed to manipulate objects in microscale. However, few of them can be applied in both the transportation and the rotation of the micro objects. In this paper, we present a novel method to realize the non-contact transportation and rotation of the…

Cited by 4SourceScholar
2017

Robotics-based micro-reeling of magnetic microfibers to fabricate helical structure for smooth muscle cells culture

ICRA 2017poster

Helical structure assembled by hydrogel microfibers is significant for culture of smooth muscle cells. However, the helical structure is only fabricated at the macroscale, while the fabrication of helical microstructure is still a challenge due to the lack of assembly method. In this paper, we propo…

Cited by 2SourceScholar
2017

Spatio-Temporal Alignment of Non-Overlapping Sequences From Independently Panning Cameras

CVPR 2017poster

This paper addresses the problem of spatio-temporal alignment of multiple video sequences. We identify and tackle a novel scenario of this problem referred to as Nonoverlapping Sequences (NOS). NOS are captured by multiple freely panning handheld cameras whose field of views (FOV) might have no dire…

Cited by 1PDFScholar
2016

High-Speed Bioassembly of Cellular Microstructures With Force Characterization for Repeating Single-Step Contact Manipulation

RA-L 2016

In vitro tissues are significant biological substitute for drug test, cell morphogenesis exploration and organ transplantation. In this letter, a novel microrobotic bioassembly method is proposed to engineer 3-D cellular structure, which can be utilized to culture in vitro tissue with microstructura

Cited by 2SourceScholar
2016

Microbubbles for High-Speed Assembly of Cell-Laden Vascular-Like Microtube

RA-L 2016

Vascular-like microtube takes an important role in delivering oxygen and nutrient to keep cell alive in the generated tissue. In this paper, we present an automated micromanipulation system to assemble 2-D gel micro-rings to vascular-like microtubes by means of generating and controlling microbubble

Cited by 1SourceScholar
2015

Automated bubble-based assembly of cell-laden microgels into vascular-like microtubes

IROS 2015poster

Fabrication of artificial blood vessels in micro scale significantly benefits the regeneration of functional human vascular networks. In this paper, we develop an efficient multi-microrobotic system with an innovative motorized sample holder (MSH) and two manipulators. Air is injected into the solut…

Cited by 3SourceScholar
2015

Pose-Invariant 3D Face Alignment

ICCV 2015poster

Face alignment aims to estimate the locations of a set of landmarks for a given image.This problem has received much attention as evidenced by the recent advancement in both the methodology and performance. However, most of the existing works neither explicitly handle face images with arbitrary pose…

Cited by 227PDFScholar