← Search

Stylianos Moschoglou

13 accepted papers

2026

Archon: A Unified Multimodal Model for Holistic Digital Human Generation

CVPR 2026

Digital humans are fundamental to immersive interaction, yet creating a unified model for holistic modalities, including text, audio, motion, and visual content, remains an open challenge. In this paper, we present Archon, a fully pretrained, human-centric unified multimodal model for holistic avata

Cited by 0SourceScholar
2026

Physical Simulator In-the-Loop Video Generation

CVPR 2026

Recent advances in diffusion-based video generation have achieved remarkable visual realism but still struggle to obey basic physical laws such as gravity, inertia, and collision. Generated objects often move inconsistently across frames, exhibit implausible dynamics, or violate physical constraints

Cited by 0SourcecodeScholar
2025

SpinMeRound: Consistent Multi-View Identity Generation Using Diffusion Models

ICCV 2025poster

Despite recent progress in diffusion models, generating realistic head portraits from novel viewpoints remains a significant challenge in computer vision. Most current approaches are constrained to limited angular ranges, predominantly focusing on frontal or near-frontal views. Moreover, although th…

Cited by 0SourcePDFScholar
2024

AnimateMe: 4D Facial Expressions via Diffusion Models

ECCV 2024poster

"The field of photorealistic 3D avatar reconstruction and generation has garnered significant attention in recent years; however, animating such avatars remains challenging. Recent advances in diffusion models have notably enhanced the capabilities of generative models in 2D animation. In this work,…

Cited by 2SourcePDFScholar
2024

Arc2Face: A Foundation Model for ID-Consistent Human Faces

ECCV 2024oral

"This paper presents , an identity-conditioned face foundation model, which, given the ArcFace embedding of a person, can generate diverse photo-realistic images with an unparalleled degree of face similarity than existing models. Despite previous attempts to decode face recognition features into de…

2023

FitMe: Deep Photorealistic 3D Morphable Model Avatars

CVPR 2023poster

In this paper, we introduce FitMe, a facial reflectance model and a differentiable rendering optimization pipeline, that can be used to acquire high-fidelity renderable human avatars from single or multiple images. The model consists of a multi-modal style-based generator, that captures facial appea…

Cited by 35SourcePDFScholar
2023

Handy: Towards a High Fidelity 3D Hand Shape and Appearance Model

CVPR 2023poster

Over the last few years, with the advent of virtual and augmented reality, an enormous amount of research has been focused on modeling, tracking and reconstructing human hands. Given their power to express human behavior, hands have been a very important, but challenging component of the human body.…

2023

Relightify: Relightable 3D Faces from a Single Image via Diffusion Models

ICCV 2023poster

Following the remarkable success of diffusion models on image generation, recent works have also demonstrated their impressive ability to address a number of inverse problems in an unsupervised way, by properly constraining the sampling process based on a conditioning input. Motivated by this, in th…

Cited by 28PDFcodeScholar
2022

3D Human Tongue Reconstruction From Single "In-the-Wild" Images

CVPR 2022oral

3D face reconstruction from a single image is a task that has garnered increased interest in the Computer Vision community, especially due to its broad use in a number of applications such as realistic 3D avatar creation, pose invariant face recognition and face hallucination. Since the introduction…

Cited by 7PDFcodeScholar
2022

MimicME: A Large Scale Diverse 4D Database for Facial Expression Analysis

ECCV 2022poster

"Recently, Deep Neural Networks (DNNs) have been shown to outperform traditional methods in many disciplines such as computer vision, speech recognition and natural language processing. A prerequisite for the successful application of DNNs is the big number of data. Even though various facial datase…

2020

AvatarMe: Realistically Renderable 3D Facial Reconstruction "In-the-Wild"

CVPR 2020poster

Over the last years, with the advent of Generative Adversarial Networks (GANs), many face analysis tasks have accomplished astounding performance, with applications including, but not limited to, face generation and 3D face reconstruction from a single "in-the-wild" image. Nevertheless, to the best…

Cited by 201PDFcodeScholar
2020

P-nets: Deep Polynomial Neural Networks

CVPR 2020poster

Deep Convolutional Neural Networks (DCNNs) is currently the method of choice both for generative, as well as for discriminative learning in computer vision and machine learning. The success of DCNNs can be attributed to the careful selection of their building blocks (e.g., residual blocks, rectifier…

Cited by 95PDFcodeScholar
2020

Synthesizing Coupled 3D Face Modalities by Trunk-Branch Generative Adversarial Networks

ECCV 2020poster

Generating realistic 3D faces is of high importance for computer graphics and computer vision applications. Generally, research on 3D face generation revolves around linear statistical models of the facial surface. Nevertheless, these models cannot represent faithfully either the facial texture or t…