Method converts facial expressions and voice of a source speaker into a target speaker.
problem Separate conversion of facial and acoustic features leads to unnatural results.
method Uses three neural networks: conversion, waveform generation, and image reconstruction.
result Significantly higher naturalness achieved when converting both features together.
Study uses stacked hourglass networks to improve facial landmark detection for medical diagnosis.
problem Improving accuracy of facial landmark detection for medical diagnosis.
method Conducted a study on landmark localisation methods using stacked hourglass networks.
result State-of-the-art stacked hourglass architecture outperforms traditional methods.
We present a method for synthesizing a frontal, neutral-expression image of a person's face given an input face photograph. This is achieved by learning to generate facial landmarks and textures from features extracted from a facial-recognition network. Unlike previous approaches, our encoding feature vector is largely…
Classifies Tinder dating profiles using facial embeddings.
problem Automatically review and classify online dating profiles.
method Uses FaceNet facial embeddings to extract features and train logistic regression models.
result Simple logistic regression on 20 profiles achieved 65% validation accuracy.
Polynomial fusion layer improves speech-driven facial animation.
problem Recent facial synthesis relies on low-dimensional representations and concatenation, ignoring higher-order interactions.
method Proposes a polynomial fusion layer to model higher-order interactions of facial encodings.
result Demonstrates improved video quality, audiovisual synchronisation, and blink generation.
Content based image retrieval, a technique which uses visual contents of image to search images from large scale image databases according to users' interests. This paper provides a comprehensive survey on recent technology used in the area of content based face image retrieval. Nowadays digital devices and photo shari…
New CNN architecture detects face spoofing with deep local features.
problem Face recognition systems are vulnerable to face spoofing attacks.
method Two-step CNN architecture: first learns features from facial regions, then fine-tunes on whole images.
result Improves face spoofing detection performance and convergence speed.
Unsupervised model predicts facial attractiveness with high accuracy.
problem Capturing the complexity of facial attractiveness through machine learning.
method Infer probabilistic models of facial preferences using Maximum Entropy and neural networks.
result High prediction accuracy in gender classification of sculpting subjects.
Efficient facial feature learning with shared representations reduces redundancy and improves accuracy.
problem Redundancy and high computational load in training deep ensemble models.
method Wide Ensemble-based Convolutional Neural Networks (ESRs) with varying branching levels.
result ESRs reduce residual generalization error and outperform state-of-the-art methods on facial expression recognition.
The paper uses facial keypoints to estimate post-surgical pain intensity.
problem Accurately assessing pain levels from self-reported ratings is challenging.
method The approach analyzes 2D and 3D facial keypoints to estimate pain intensity.
result The pain estimation model uses multiple instance learning.
Enhances 2D face recognition with 3D features using active illumination.
problem Improving robustness of 2D face recognition to spoofing attacks and low-light conditions.
method Projecting a high spatial frequency pattern onto the face to recover 3D information and a 2D image simultaneously.
result Significantly boosts face recognition performance and dramatically improves robustness to spoofing attacks.
Combines multiple data types to predict emotions in images.
problem Predicting emotions in images using various data types.
method Combines facial features, scene extraction, audio tonality, human pose, text-based tagging, and CNN predictions.
result Improves accuracy in emotion prediction compared to baseline methods.
Unified facial behavior analysis network improves performance across tasks.
problem Independent study of facial behavior tasks.
method Single multi-task, multi-domain, multi-label network (FaceBehaviorNet).
result Joint training of facial behavior tasks yields better performance.
Paper introduces a method for generating interlocutor-aware facial gestures in dyadic settings.
problem Generating appropriate non-verbal behavior for conversational agents in dyadic settings.
method Probabilistic method using multi-modal cues from the interlocutor to synthesize facial gestures.
result The model successfully leverages multi-modal input from the interlocutor to generate more appropriate behavior.
We propose a novel method for automatic pain intensity estimation from facial images based on the framework of kernel Conditional Ordinal Random Fields (KCORF). We extend this framework to account for heteroscedasticity on the output labels(i.e., pain intensity scores) and introduce a novel dynamic features, dynamic ra…
Develops a fast non-invasive tool for diagnosing pediatric sleep apnea.
problem Diagnosing pediatric obstructive sleep apnea using an overnight sleep study is often impractical.
method Combines persistent homology, geometric shape analysis, and convolutional neural networks to classify facial images.
result Facial features associated with obstructive sleep apnea can be recognized for diagnosis.
We address the task of simultaneous feature fusion and modeling of discrete ordinal outputs. We propose a novel Gaussian process(GP) auto-encoder modeling approach. In particular, we introduce GP encoders to project multiple observed features onto a latent space, while GP decoders are responsible for reconstructing the…
A model separates visual style from digit type on MNIST and facial features from shape on CelebA.
problem Learning compact, independent factors of data.
method Explicitly encoded in a generative model with two latent spaces: spatial transformations and intrinsic appearance.
result The model separates visual style from digit type on MNIST and facial features from shape on CelebA.
AI misidentifies facial expressions in videos, often misinterpreting happiness as sadness.
problem Automated facial emotion recognition in videos, especially with mixed modalities (visual and audio).
method Applied state-of-the-art visual and temporal networks, explored feature fusion methods.
result Machine learning models misclassify emotions, particularly happiness as sadness.
This paper improves facial expression recognition using CNNs and coherence constraints.
problem Facial expression recognition from static images and video sequences is challenging.
method Investigates the use of Convolutional Neural Networks (CNNs) with coherence constraints in a semi-supervised setting.
result Coherence constraints improve facial expression recognition quality, especially in the presence of occlusions.
Facial landmark localization and occlusion estimation for driver safety.
problem Robust facial landmark localization and occlusion estimation under harsh lighting and occlusion.
method Occluded Stacked Hourglass approach based on Stacked Hourglass network.
result State-of-the-art results in face detection, head pose, and occlusion estimation on various datasets.
FEAFA dataset annotates facial expressions with high detail.
problem Lack of detailed facial expression annotations in existing datasets.
method Manual annotation of 122 participants' facial expressions.
result FEAFA dataset provides detailed annotations for facial expressions.
NaimishNet detects facial key points using deep learning.
problem Facial key points detection in computer vision.
method Adapted LeNet deep CNN model for facial key points data.
result NaimishNet outperforms state-of-the-art approaches.
Limited annotated data available for the recognition of facial expression and action units embarrasses the training of deep networks, which can learn disentangled invariant features. However, a linear model with just several parameters normally is not demanding in terms of training data. In this paper, we propose an el…
GANs can bias synthetic data, affecting minority and female faces.
problem GANs can amplify biases in synthetic data augmentation.
method Examine GANs on face-shots with gender and skin tone biases.
result GANs generate biased synthetic data, skewing minority modes and features.
This work generates synthetic 3D thermal facial data using 2D facial data and deep learning.
problem Creating large datasets for deep learning in computer vision.
method 3D facial modelling techniques and deep learning methodologies.
result Synthetic 3D thermal facial data created for deep learning applications.
The human face constantly conveys information, both consciously and subconsciously. However, as basic as it is for humans to visually interpret this information, it is quite a big challenge for machines. Conventional semantic facial feature recognition and analysis techniques are already in use and are based on physiol…
Fawkes protects images from unauthorized facial recognition models.
problem Unauthorized training of facial recognition models poses privacy risks.
method Fawkes adds imperceptible pixel-level changes (cloaks) to images before release.
result Fawkes can protect images from misidentification by 95% and 80% even when clean images are leaked.
Study reveals biases in facial landmark detection methods for dementia patients.
problem Challenges in facial landmark detection for older adults with dementia.
method Evaluation of seven facial landmark detection methods on frontal, profile, and various face regions.
result Significant performance differences between dementia patients and non-patients, and biases across face regions.
Survey examines public views on facial recognition technology.
problem Public acceptance and privacy concerns with facial recognition.
method Cross-national survey of 4 countries.
result Public views vary significantly across countries.
cVAE enhances salient latent features using contrastive learning.
problem Identifying salient latent features in datasets with enriched variation.
method Contrastive Variational Autoencoder (cVAE) combining contrastive learning and deep generative models.
result cVAE effectively uncovers salient latent features across diverse datasets.
This review explores automated pain detection from facial expressions using FACS.
problem Automated detection of pain from facial expressions is challenging and underexplored.
method Examines progress in AFER and APD, focusing on FACS-based methods and deep learning.
result Limited studies on APD, highlighting the need for clinical applications.
New method transfers emotions in facial images.
problem Transforming facial images to different emotions.
method Infinite task learning and vector-valued reproducing kernel Hilbert spaces.
result Achieves low reconstruction cost and high emotion classification accuracy.
Paper evaluates CNN-based facial landmark detection methods.
problem Evaluate characteristics and performance of CNN-based facial landmark detection methods.
method Divided into regression and heatmap approaches, investigated using a hybrid loss function and discrimination network.
result Proposed model outperforms other models in all tested datasets.
Improved animated faces using audiovisual and modality dropout.
problem Creating realistic animated faces using speech and visual cues.
method Training a deep learning model with modality dropout to balance audio and visual inputs.
result Modality dropout improves viewer preference for audiovisual-driven animation.
Bayesian Neural Networks improve uncertainty modeling in facial emotion recognition.
problem High aleatoric uncertainty and visual ambiguity in facial emotion recognition.
method Bayesian Neural Networks approximated using MC-Dropout, MC-DropConnect, or Ensemble methods.
result Bayesian Neural Networks produce more human-like output probabilities.
Paper studies facial keypoint detection using various algorithms.
problem Challenges in accurately detecting facial keypoints from complex images.
method Preprocess data with PCA and LBP, apply multiple algorithms including linear regression, tree models, neural networks, and CNNs.
result Demonstrates the effectiveness of the proposed framework through comprehensive experiments.
Study aims to develop a humanoid robot dialogue system.
problem Current dialogue systems lack attention to non-verbal cues.
method Participated in a competition to develop a system with facial expressions and gaze control.
result Developed a humanoid robot dialogue system.
AttGAN edits facial attributes by changing only what you want, preserving details.
problem Facial attribute editing with preservation of details.
method Encoder-decoder architecture with attribute classification and reconstruction learning.
result Outperforms state-of-the-arts on realistic attribute editing with preserved details.
Paper proposes a method to combine multiple facial analysis models for better performance.
problem Facial analysis models from different sources have low transferability.
method Two-step process: 1) Auto-encoder for common embedding, 2) Distillation for lightweight model.
result Lightweight model outperforms state-of-the-art on 15 facial analysis tasks.
First steps towards a mathematical theory of deep convolutional neural networks for feature extraction were made---for the continuous-time case---in Mallat, 2012, and Wiatowski and Bölcskei, 2015. This paper considers the discrete case, introduces new convolutional neural network architectures, and proposes a mathemati…
New system detects pain from facial AU combinations using MIL and MCIL.
problem Detecting pain from facial expressions reliably.
method Weakly supervised learning, multiple instance learning, multiple clustered instance learning.
result 87% pain recognition accuracy on UNBC-McMaster Shoulder Pain Expression dataset.
AutoTune learns wireless identifiers for facial recognition in real-world settings.
problem Facial recognition requires extensive user training, making it impractical for widespread deployment.
method Uses ambient wireless identifiers to train deep neural networks for facial recognition without user effort.
result Demonstrates a system that continuously refines facial recognition using wireless identifiers over time.
This work introduces a framework to detect unintended bias in facial analysis models.
problem Detecting unintended biases in facial analysis models used in critical applications.
method Image counterfactual sensitivity analysis using generative adversarial networks.
result Identifies factors affecting facial classifier predictions, revealing unintended biases.
DepthNets learns 3D face geometry and transformations without supervision.
problem Learning 3D face geometry and transformations from a single image.
method Unsupervised learning of facial keypoints depth, using backpropable loss for 3D transformations.
result DepthNets can predict 3D transformations and re-target faces to new poses or geometries.
New NMF algorithm uses Toeplitz matrix for facial recognition.
problem Facial recognition performance improvement.
method Proposes TNMF algorithm with Toeplitz penalty for NMF.
result TNMF outperforms ZNMF and other constrained NMF algorithms.
This paper explores how facial recognition systems can be fooled by adversarial attacks.
problem Adversarial attacks on facial recognition systems.
method Applying Fast Gradient Sign Method and crafting various black-box attack algorithms.
result High levels of perturbation can significantly decrease classifier confidence and misclassification rates.
Enhances facial emotion recognition with gradient and Laplacian images.
problem Improving the performance of facial emotion recognition systems.
method Proposes using gradient and Laplacian of input images with a CNN.
result Enhances FER systems by 3 to 5%.