← Search

Yuge Huang

19 accepted papers

2026

GloTok: Global Perspective Tokenizer for Image Reconstruction and Generation

AAAI 2026technical

Existing state-of-the-art image tokenization methods leverage diverse semantic features from pre-trained vision models for additional supervision, to expand the distribution of latent representations and thereby improve the quality of image reconstruction and generation. These methods employ a local

Cited by 0SourcePDFScholar
2026

TranX-Adapter: Bridging Artifacts and Semantics within MLLMs for Robust AI-generated Image Detection

ICML 2026poster

Rapid advances in AI-generated image (AIGI) technology enable highly realistic synthesis, threatening public information integrity and security. Recent studies have demonstrated that incorporating texture-level artifact features alongside semantic features into multimodal large language models (MLLM…

Cited by 0SourceScholar
2025

Data Synthesis with Diverse Styles for Face Recognition via 3DMM-Guided Diffusion

CVPR 2025poster

Identity-preserving face synthesis aims to generate synthetic face images of virtual subjects that can substitute real-world data for training face recognition models. While prior arts strive to create images with consistent identities and diverse styles, they face a trade-off between them. Identify…

2025

EyeSeg: An Uncertainty-Aware Eye Segmentation Framework for AR/VR

IJCAI 2025

Human-machine interaction through augmented reality (AR) and virtual reality (VR) is increasingly prevalent, requiring accurate and efficient gaze estimation which hinges on the accuracy of eye segmentation to enable smooth user experiences. We introduce EyeSeg, a novel eye segmentation framework de

Cited by 0SourcePDFScholar
2025

SlerpFace: Face Template Protection via Spherical Linear Interpolation

AAAI 2025technical

Contemporary face recognition systems use feature templates extracted from face images to identify persons. To enhance privacy, face template protection techniques are widely employed to conceal sensitive identity and appearance information stored in the template. This paper identifies an emerging p…

Cited by 6SourcePDFScholar
2025

Stylized-Face: A Million-level Stylized Face Dataset for Face Recognition

ICCV 2025poster

Stylized face recognition is the task of recognizing generated faces with the same ID across diverse stylistic domains (e.g., anime, painting, cyberpunk styles). This emerging field plays a vital role in the governance of generative image, serving the primary objective: Recognize the ID information…

2025

UIFace: Unleashing Inherent Model Capabilities to Enhance Intra-Class Diversity in Synthetic Face Recognition

ICLR 2025poster

Face recognition (FR) stands as one of the most crucial applications in computer vision. The accuracy of FR models has significantly improved in recent years due to the availability of large-scale human face datasets. However, directly using these datasets can inevitably lead to privacy and legal pr…

2024

$\text{ID}^3$: Identity-Preserving-yet-Diversified Diffusion Models for Synthetic Face Recognition

NeurIPS 2024poster

Synthetic face recognition (SFR) aims to generate synthetic face datasets that mimic the distribution of real face data, which allows for training face recognition models in a privacy-preserving manner. Despite the remarkable potential of diffusion models in image generation, current diffusion-based…

Cited by 4SourcePDFScholar
2024

Privacy-Preserving Face Recognition Using Trainable Feature Subtraction

CVPR 2024poster

The widespread adoption of face recognition has led to increasing privacy concerns as unauthorized access to face images can expose sensitive personal information. This paper explores face image protection against viewing and recovery attacks. Inspired by image compression we propose creating a visu…

2023

Privacy-Preserving Face Recognition Using Random Frequency Components

ICCV 2023poster

The ubiquitous use of face recognition has sparked increasing privacy concerns, as unauthorized access to sensitive face images could compromise the information of individuals. This paper presents an in-depth study of the privacy protection of face images' visual information and against recovery. Dr…

Cited by 21PDFcodeScholar
2022

Evaluation-Oriented Knowledge Distillation for Deep Face Recognition

CVPR 2022oral

Knowledge distillation (KD) is a widely-used technique that utilizes large networks to improve the performance of compact models. Previous KD approaches usually aim to guide the student to mimic the teacher's behavior completely in the representation space. However, such one-to-one corresponding con…

Cited by 47PDFcodeScholar
2022

Privacy-Preserving Face Recognition with Learnable Privacy Budgets in Frequency Domain

ECCV 2022poster

"Face recognition technology has been used in many fields due to its high recognition accuracy, including the face unlocking of mobile devices, community access control systems, and city surveillance. As the current high accuracy is guaranteed by very deep network structures, facial images often nee…

2021

Consistent Instance False Positive Improves Fairness in Face Recognition

CVPR 2021poster

Demographic bias is a significant challenge in practical face recognition systems. Several methods have been proposed to reduce the bias, which rely on accurate demographic annotations. However, such annotations are usually not available in real scenarios. Moreover, these methods are explicitly desi…

Cited by 66PDFcodeScholar
2021

SDD-FIQA: Unsupervised Face Image Quality Assessment With Similarity Distribution Distance

CVPR 2021poster

In recent years, Face Image Quality Assessment (FIQA) has become an indispensable part of the face recognition system to guarantee the stability and reliability of recognition performance in an unconstrained scenario. For this purpose, the FIQA method should consider both the intrinsic property and…

Cited by 140PDFcodeScholar
2021

Scribble-Supervised Semantic Segmentation Inference

ICCV 2021poster

In this paper, we propose a progressive segmentation inference (PSI) framework to tackle with scribble-supervised semantic segmentation. In virtue of latent contextual dependency, we encapsulate two crucial cues, contextual pattern propagation and semantic label diffusion, to enhance and refine pixe…

Cited by 42PDFScholar
2021

Wasserstein Coupled Graph Learning for Cross-Modal Retrieval

ICCV 2021poster

Graphs play an important role in cross-modal image-text understanding as they characterize the intrinsic structure which is robust and crucial for the measurement of cross-modal similarity. In this work, we propose a Wasserstein Coupled Graph Learning (WCGL) method to deal with the cross-modal retri…

Cited by 29PDFScholar
2020

CurricularFace: Adaptive Curriculum Learning Loss for Deep Face Recognition

CVPR 2020poster

As an emerging topic in face recognition, designing margin-based loss functions can increase the feature margin between different classes for enhanced discriminability. More recently, the idea of mining-based strategies is adopted to emphasize the misclassified samples, achieving promising results.…

Cited by 686PDFcodeScholar
2020

Improving Face Recognition from Hard Samples via Distribution Distillation Loss

ECCV 2020poster

Large facial variations are the main challenge in face recognition. To this end, previous variation-specific methods make full use of task-related prior to design special network losses, which are typically not general among different tasks and scenarios. In contrast, the existing generic methods fo…