Paper proposes a method to encrypt faces while maintaining visual similarity.
problem Protecting personal data from unauthorized face recognition.
method Targeted identity-protection iterative method (TIP-IM) to generate adversarial identity masks.
result TIP-IM provides 95%+ protection success rate against face recognition models.
Paper proposes using synthetic data to improve face recognition accuracy.
problem Improving face recognition accuracy using real data alone.
method Proposes a GAN that disentangles identity attributes and generates photo-realistic synthetic images.
result Synthetic images generated by the model are photo-realistic and can increase face recognition accuracy.
Paper proposes a framework to protect user anonymity in emotion recognition.
problem Preserving user anonymity in face-based emotion recognition systems.
method Adversarial learning framework using CNN architecture.
result The proposed approach minimizes identity-specific information and maximizes emotion-dependent information.
A new method for person recognition using cosine loss.
problem Recognizing the same identity across time and space with complicated scenes and similar appearance.
method Proposes a congenerous cosine loss to train a network for robust and representative features.
result The proposed method achieves better classification accuracy than previous state-of-the-arts.
Face recognition system trained with noisy labels.
problem Label noise in training deep learning classifiers.
method Review and apply recent methods to manage noisy annotations.
result Improved performance of face recognition system with noisy labels.
Max pooling selects frames with similar poses for better face recognition.
problem Measuring similarity of faces with varying poses in videos.
method Pose-Selective Max Pooling: Select frames closest to centroid of K-means cluster.
result Max correlation among selected features yields better performance than VGG-face.
Bayesian weight priors improve neural network learning of identity relations.
problem Neural networks struggle to learn abstract and systematic relations, especially identity relations.
method Extended RBP approach using Bayesian weight priors as a regularization term.
result Bayesian weight priors lead to perfect generalization for identity relations and do not hinder standard neural network learning.
Generative adversarial network synthesizes sketches into realistic images.
problem Improving the quality and realism of facial sketches.
method Hybrid GAN with quality guided encoder and identity preserving network.
result Synthesized images are more realistic and maintain identity.
Neural network framework for language recognition considers sequence information and improves accuracy.
problem Challenging task of automatic language identification in noisy conditions.
method Proposes a neural network framework with bidirectional LSTM and attention modeling for relevance weighting.
result Significant improvements over conventional methods in noisy conditions and multi-speaker speech.
Subject Cross Validation improves Human Activity Recognition performance by up to 16%.
problem Overestimation of Human Activity Recognition performance using k-fold cross validation.
method Investigated Subject Cross Validation vs. k-fold cross validation for Human Activity Recognition.
result Subject Cross Validation increases performance by up to 16%.
Motivated by vision tasks such as robust face and object recognition, we consider the following general problem: given a collection of low-dimensional linear subspaces in a high-dimensional ambient (image) space and a query point (image), efficiently determine the nearest subspace to the query in ℓ1 distance. We …
Optimizes master faces for 2D and 3D face verification using evolutionary algorithms and neural networks.
problem Impersonation attacks using master faces for face-based identity authentication.
method Evolutionary algorithm in latent space of StyleGAN, neural network to direct search, 2D and 3D face reconstruction.
result Obtains high impersonation rates with fewer master faces for 2D and 3D face verification.
Unified model for age-invariant face recognition with photorealistic face synthesis.
problem Reliable face recognition across ages remains challenging due to significant intra-class variations.
method Unified deep architecture for cross-age face synthesis and recognition, continuous face rejuvenation/aging, disentangled age-invariant face representations.
result Superior performance on CAFR and other cross-age datasets, promising generalizability to unconstrained face recognition.
Simpler CNN model with spatial attention and temporal pooling outperforms complex models.
problem Emotion recognition from videos with small face deformations and identity variations.
method Spatial attention mechanism and temporal softmax pooling applied to a pre-trained CNN.
result The approach achieves higher accuracy than state-of-the-art methods on the EmotiW dataset.
The study compares adversarial and multi-task learning for speech recognition, finding invariant representations are key.
problem Improving speech recognition performance with speaker information.
method Investigated multi-task learning and adversarial learning for speech recognition, comparing their effects on error rates.
result Deep models already develop speaker-invariant representations, and adversarial learning has a minor impact.
In this paper we study speaker linking (a.k.a.\ partitioning) given constraints of the distribution of speaker identities over speech recordings. Specifically, we show that the intractable partitioning problem becomes tractable when the constraints pre-partition the data in smaller cliques with non-overlapping speakers…
Predicts classifier accuracy scaling with more classes.
problem Estimating classifier performance with increasing number of classes.
method Uses statistical moments and data from a subset of classes.
result Expected accuracy can be estimated from a subset of classes.
Acoustic Neighbor Embeddings map speech and text to fixed dimensions for phonetic confusability.
problem Mapping speech and text to fixed dimensions for phonetic confusability.
method Adapting SNE to sequential inputs, training two encoder neural networks.
result More accurate results with low-dimensional embeddings in word recognition tasks.
A framework learns image embeddings robust to transformations for better recognition.
problem Learning robust image embeddings resistant to transformations like viewpoint, scale, and illumination.
method Discriminate-and-Rectify Encoders using orbit sets, deep parametrizations, and a novel orbit-based loss.
result Learned embeddings are robust to geometric transformations and improve one-shot classification.
Motivated by vision tasks such as robust face and object recognition, we consider the following general problem: given a collection of low-dimensional linear subspaces in a high-dimensional ambient (image) space, and a query point (image), efficiently determine the nearest subspace to the query in ℓ1 distance. In…
Synthesizes faces from facial features, invariant to pose and expression.
problem Creating realistic face images from facial features.
method Learning facial landmarks and textures from facial-recognition features, training on frontal, neutral-expression images.
result Generated images are invariant to lighting, pose, and expression.
Proposes a linear model for facial action recognition without requiring large datasets.
problem Limited annotated data for facial expression and action units.
method Exploits low-rank property across frames and group sparsity to subtract neutral faces and recognize actions.
result One-shot automatic method on raw face videos performs competitively and better than previous methods.
Researchers found PP-GANs can hide sensitive data in sanitized images, undermining privacy checks.
problem Lack of formal proofs of privacy in PP-GANs for image sanitization.
method Subverted PP-GANs for facial expression recognition to hide sensitive data in sanitized images.
result It is possible to hide sensitive identification data in sanitized PP-GAN output images, even allowing reconstruction of entire input images.
Pixel-wise relevance method shows how CNNs classify faces, varying across datasets and tasks.
problem Interpreting black-box CNN face recognition models.
method Layer-wise relevance propagation (LRP) applied to VGG-16 models trained for face recognition.
result Relevance maps are generally stable across random initializations and tasks, but less so across pretraining datasets.
A novel method to prevent bias in authentication models.
problem Bias in data-driven authentication models trained in one domain but required to apply in others.
method Two-stage method involving one-versus-rest disentangle learning and additive adversarial learning.
result Demonstrated effectiveness and superiority of the proposed method through comprehensive evaluation.
Anti-transfer learning prevents misleading representations for speech tasks.
problem Misleading representations learned from orthogonal tasks in speech processing.
method Penalizes similarity between activations of a network and another trained on an orthogonal task.
result Improves classification accuracy and invariance to the orthogonal task.
Photo-identification technique improved for new dolphin individuals.
problem Traditional photo-identification of dolphins is laborious and manual.
method Metric embedding learning using triplet loss function in Euclidean space.
result Compact representation of fin images generalizes well to new identities.
End-to-end framework learns new classes dynamically.
problem Challenges in recognizing unseen classes in real-world settings.
method Dynamic cascade of classifiers that incrementally learn features.
result Outperforms existing methods on real-world datasets.
Estimates the capacity of face representations, providing upper bounds for automatic face recognition.
problem Estimating how many identities a face representation can resolve.
method Formulated as packing bounds on a low-dimensional manifold embedded in a deep representation space, accounting for manifold structure and noise.
result Demonstrated upper bounds of 2.7×10^4 and 8.4×10^4 for FaceNet and SphereFace at a FAR of 1%, respectively.
Neural networks struggle with abstract patterns, new RBP structures improve performance.
problem Neural networks fail to learn abstract patterns based on identity rules.
method Proposed Relation Based Pattern (RBP) extensions to neural network structures.
result Neural networks with RBP structures achieve perfect performance on synthetic and real-world sequence prediction tasks.
New method for continual learning without task boundaries.
problem Traditional continual learning is task-based and impractical for real-world applications.
method Developed an online continual learning system using Memory Aware Synapses.
result Valid approach demonstrated in self-supervised learning and robot collision avoidance.
Enhances PLDA for speaker recognition by considering channel variables.
problem Speaker recognition challenges with same-channel trials.
method Generalizes PLDA model to include channel variable dependency.
result Improves performance in same-channel versus different-channel trials.
Deep learning improves gait biometric recognition accuracy.
problem Low accuracy in existing CSI-based gait identification systems.
method Developed an end-to-end deep CSI learning system using deep neural networks.
result Achieved a top-1 accuracy of 97.12% for a dataset of 30 people.
If V and W are varieties of algebras such that any V-algebra A has a reduct U(A) in W, there is a forgetful functor U: V->W that acts by A |-> U(A) on objects, and identically on homomorphisms. This functor U always has a left adjoint F: W->V by general considerations. One calls F(B) the V-algebra freely generated by t…
This article reviews zero-shot recognition techniques for unseen categories.
problem Scaling recognition to many classes with few training samples.
method Comprehensive review of existing zero-shot recognition techniques.
result Highlighting limitations and future directions in zero-shot recognition.
Improved speech recognition using EEG and video.
problem Enhancing continuous speech recognition systems.
method Implemented a CTC-based ASR model using EEG features.
result EEG features improve continuous visual speech recognition.
A new generator for GANs separates style and variation.
problem Improving GANs' ability to control and disentangle image attributes.
method Borrowing from style transfer, a new generator architecture.
result Improves GANs' quality and disentanglement of latent factors.
Solves recognition problem of frontal singularities.
problem Recognition of frontal singularities.
method Specified geometric frontal singularities, provided explicit normal forms, combined results from K. Saji and applied to tangent surfaces of null curves.
result Classification of singularities in tangent surfaces of null curves.
CSI-Net learns WiFi signals for body characterization and pose recognition.
problem Unified learning of body characteristics and pose recognition.
method Unified Deep Neural Network (DNN) for WiFi signal representation and multi-task learning.
result CSI-Net solves biometrics estimation and person recognition.
Study shows emotion affects speaker recognition and vice versa.
problem Dependencies between emotion and speaker recognition.
method Transfer learning and fine-tuning for emotion classification.
result Fine-tuning improves emotion recognition performance by 30.40% on IEMOCAP, 7.99% on MSP-Podcast, and 8.61% on Crema-D.
End-to-end speech recognition using EEG without speech input.
problem Speech recognition without direct speech input.
method Implemented attention model and CTC-based ASR systems for EEG signals; fused EEG with noisy speech features.
result Demonstrated end-to-end speech recognition using EEG signals.
New proof for sphere recognition algorithm.
problem Sphere recognition algorithm proof.
method New proof of a lemma in Abigail Thompson's algorithm.
result New proof of a lemma in Abigail Thompson's proof of the Recognition Algorithm for 3-spheres.
Continuous speech recognition from brain activity without vocalization.
problem Recognizing silent speech from EEG signals.
method Implemented a CTC ASR model using EEG signals.
result Demonstrated feasibility of EEG for continuous silent speech recognition.
VoxCeleb 2019 challenge assesses speaker recognition in uncontrolled settings.
problem Evaluate speaker recognition technology in unconstrained data.
method Public dataset, challenge, and workshop at Interspeech 2019.
result Baseline results and discussions provided.
Paper proposes Roweisposes for 3D action recognition using generalized eigenvalue problem.
problem Need for basic methods in 3D action recognition.
method Roweisposes uses Roweis discriminant analysis for generalized subspace learning.
result Roweisposes is effective for 3D action recognition.
Paper tackles zero-shot activity recognition using video features and text embeddings.
problem Zero-shot activity recognition with videos.
method Auto-encoder model for multimodal joint embedding, 3D convolutional action recognition for visual features, GloVe word embeddings for textual features.
result Improved zero-shot recognition results with top-n accuracy and mean Nearest Neighbor Overlap.
Aims to eliminate domain bias in authentication without domain labels.
problem Authentication models are biased due to domain differences.
method Discover latent domains and eliminate domain difference alternately, using a meta-learning framework.
result Eliminates domain difference in authentication without domain labels.
Survey of open set recognition techniques and their limitations.
problem Recognition tasks with unknown classes during testing.
method Comprehensive review of techniques, datasets, and evaluation criteria.
result Highlighting the limitations and future directions in open set recognition.