← Search

Ognjen (Oggi) Rudovic

4 accepted papers

2022

Streaming on-Device Detection of Device Directed Speech from Voice and Touch-Based Invocation

ICASSP 2022accepted

When interacting with smart devices such as mobile-phones or wearables, the user typically invokes a virtual assistant (VA) by saying a keyword or by pressing a button on the device. However, in many cases, the VA can accidentally be invoked by the keyword-like speech or accidental button press, whi…

Cited by 0SourceScholar
2017

Deep Structured Learning for Facial Action Unit Intensity Estimation

CVPR 2017poster

We consider the task of automated estimation of facial expression intensity. This involves estimation of multiple output variables (facial action units --- AUs) that are structurally dependent. Their structure arises from statistically induced co-occurrence patterns of AU intensity levels. Modeling…

Cited by 155PDFScholar
2017

DeepCoder: Semi-Parametric Variational Autoencoders for Automatic Facial Action Coding

ICCV 2017poster

Human face exhibits an inherent hierarchy in its representations (i.e., holistic facial expressions can be encoded via a set of facial action units (AUs) and their intensity). Variational (deep) auto-encoders (VAE) have shown great results in unsupervised extraction of hierarchical latent representa…

Cited by 56PDFScholar
2017

PUnDA: Probabilistic Unsupervised Domain Adaptation for Knowledge Transfer Across Visual Categories

ICCV 2017poster

This paper introduces a probabilistic latent variable model to address unsupervised domain adaptation problems. This is achieved by learning projections from each domain to a latent space along the classifier in the latent space to simultaneously minimizing a notion of domain disparity while maximiz…

Cited by 51PDFScholar