← Search

Xianxu Hou

9 accepted papers

2025

FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs

CVPR 2025poster

Multimodal large language models (MLLMs) have demonstrated remarkable capabilities in various tasks. However, effectively evaluating these MLLMs on face perception remains largely unexplored. To address this gap, we introduce FaceBench, a dataset featuring hierarchical multi-view and multi-level att…

2025

High-Fidelity Editable Portrait Synthesis with 3D GAN Inversion

ICASSP 2025accepted

The 3D generative adversarial network (GAN) inversion converts an image into 3D representation to attain high-fidelity reconstruction and facilitate realistic image manipulation within the 3D latent space. However, previous approaches face challenges regarding the trade-off between the reconstructio…

Cited by 0SourceScholar
2024

A Dataset and Model for Realistic License Plate Deblurring

IJCAI 2024poster

Vehicle license plate recognition is a crucial task in intelligent traffic management systems. However, the challenge of achieving accurate recognition persists due to motion blur from fast-moving vehicles. Despite the widespread use of image synthesis approaches in existing deblurring and recogniti…

2024

Dynamic Data Sampler for Cross-Language Transfer Learning in Large Language Models

ICASSP 2024accepted

Large Language Models (LLMs) have gained significant attention in the field of natural language processing (NLP) due to their wide range of applications. However, training LLMs for languages other than English poses significant challenges, due to the difficulty in acquiring large-scale corpus and th…

Cited by 0SourceScholar
2023

StyleGene: Crossover and Mutation of Region-Level Facial Genes for Kinship Face Synthesis

CVPR 2023highlight

High-fidelity kinship face synthesis has many potential applications, such as kinship verification, missing child identification, and social media analysis. However, it is challenging to synthesize high-quality descendant faces with genetic relations due to the lack of large-scale, high-quality anno…

2022

C2AM: Contrastive Learning of Class-Agnostic Activation Map for Weakly Supervised Object Localization and Semantic Segmentation

CVPR 2022poster

While class activation map (CAM) generated by image classification network has been widely used for weakly supervised object localization (WSOL) and semantic segmentation (WSSS), such classifiers usually focus on discriminative object regions. In this paper, we propose Contrastive learning for Class…

Cited by 142PDFcodeScholar
2022

CLIMS: Cross Language Image Matching for Weakly Supervised Semantic Segmentation

CVPR 2022poster

It has been widely known that CAM (Class Activation Map) usually only activates discriminative object regions and falsely includes lots of object-related backgrounds. As only a fixed set of image-level object labels are available to the WSSS (weakly supervised semantic segmentation) model, it could…

Cited by 183PDFcodeScholar
2022

RamGAN: Region Attentive Morphing GAN for Region-Level Makeup Transfer

ECCV 2022poster

"In this paper, we propose a region adaptive makeup transfer GAN, called RamGAN, for precise region-level makeup transfer. Compared to face-level transfer methods, our RamGAN uses spatial-aware Region Attentive Morphing Module (RAMM) to encode Region Attentive Matrices (RAMs) for local regions like…

2020

End-to-End Illuminant Estimation Based on Deep Metric Learning

CVPR 2020poster

Previous deep learning approaches to color constancy usually directly estimate illuminant value from input image. Such approaches might suffer heavily from being sensitive to the variation of image content. To overcome this problem, we introduce a deep metric learning approach named Illuminant-Guide…

Cited by 44PDFcodeScholar