Deep learning models improve sound separation across various types of sounds.
problem Developing a universal method to separate arbitrary sounds of different types.
method Created a dataset of mixtures containing arbitrary sounds, investigated mask-based separation architectures, and tested different framewise analysis-synthesis bases.
result STFT outperformed learnable bases in universal sound separation tasks.
This work disentangles speech and non-speech components from found data.
problem Building robust acoustic models from found data with non-standard variations.
method Latent Stochastic Models and Multinode Latent Space Variational Autoencoders (VAE).
result Speech and music can be separated in the latent space of a VAE, independent of the language.
With deep learning approaches becoming state-of-the-art in many speech (as well as non-speech) related machine learning tasks, efforts are being taken to delve into the neural networks which are often considered as a black box. In this paper it is analyzed how recurrent neural network (RNNs) cope with temporal dependen…
Unified deep learning framework improves SV in noisy, reverberant, and long non-speech segments.
problem Robust speaker verification in adverse environments, especially short speech segments.
method Feature Pyramid Module (FPM)-based Multi-scale Aggregation (MSA), Self-adaptive Soft VAD (SAS-VAD), Masking-based Speech Enhancement (SE).
result The proposed method outperforms baseline systems in challenging conditions.
Automatic conflict detection has grown in relevance with the advent of body-worn technology, but existing metrics such as turn-taking and overlap are poor indicators of conflict in police-public interactions. Moreover, standard techniques to compute them fall short when applied to such diversified and noisy contexts. W…
Study improves radio show segmentation using audio embeddings.
problem Automated segmentation of radio shows.
method Created audio embeddings from multi-class classification tasks on different datasets, evaluated performance against text-only baseline.
result Audio embeddings from non-speech sound event classification significantly outperformed text-only baseline by 32.3% in F1-measure.
Paper proposes a new framework for SAD using GANs.
problem Speech Activity Detection (SAD) in diverse conditions.
method Joint learning with GANs and temporal discriminator.
result Framework outperforms state-of-the-art SAD approaches.
Personal VAD detects target speaker voice activity efficiently.
problem Efficiently detect target speaker voice activity for reduced computational cost and battery usage.
method Trains a neural network conditioned on speaker embedding or verification score, outputs probabilities for three speech classes.
result Trained model with 130K parameters outperforms combined standard VAD and speaker recognition networks.
Speech, Music and Noise classification/segmentation is an important preprocessing step for audio processing/indexing. To this end, we propose a novel 1D Convolutional Neural Network (CNN) - SwishNet. It is a fast and lightweight architecture that operates on MFCC features which is suitable to be added to the front-end …
Study uses PPG signals for detecting speech events and speaker characteristics.
problem Detecting speech events and speaker characteristics from PPG signals.
method End-to-end convolutional neural network architectures for gender and person verification.
result Promising results showing potential of PPG for speech processing tasks.
Paper proposes a new method for robust speaker verification.
problem Improving robustness in speaker verification systems.
method Combines soft VAD and self-adaptive VAD with DNN-based VAD.
result Significant improvement in verification performance in real-world environments.
Improved animated faces using audiovisual and modality dropout.
problem Creating realistic animated faces using speech and visual cues.
method Training a deep learning model with modality dropout to balance audio and visual inputs.
result Modality dropout improves viewer preference for audiovisual-driven animation.
New graph types help identify complex relationships.
problem Understanding complex relationships in data.
method Introducing separable and essentially separable graphs to characterize and identify graphical models.
result Developed algorithms to identify equivalence classes of essentially separable graphs.
This paper improves universal sound separation using sound classification.
problem Separating acoustic sources from an open domain, regardless of their class.
method Utilizing semantic embeddings from a sound classifier to condition a separation network.
result Classifier embeddings provide nearly one dB of SNR gain, and iterative models achieve significant performance.
Shallow neural nets classify objects perfectly if their distribution is linearly separable.
problem Designing efficient neural networks for classification.
method Constructed shallow sigmoid-type neural networks.
result Achieves 100% accuracy for datasets following a linear separability condition.
Neural networks can separate non-separable data using feature maps.
problem Non-separable data in neural networks.
method Characterization of feedforward neural networks and use of feature maps.
result ReLU neural networks can separate concentric data.
DSI measures dataset separability for neural networks.
problem Difficulty in separating different classes of data in neural networks.
method Created the Distance-based Separability Index (DSI) to quantify dataset separability.
result DSI effectively measures dataset separability and indicates similar distributions of different classes.
Separates estimation and control in risk-sensitive investment problems with partial observation.
problem Risk-sensitive investment problems with incomplete observation.
method Investigates separability of a general class of risk-sensitive investment management problems using a finite-dimensional filter.
result The separated problem is strictly equivalent to the original control problem.
Proposes a two-step method for sound source separation.
problem Improving sound source separation performance.
method First, learn a latent space transform. Second, train a separation module in the latent space.
result The proposed method achieves better performance than joint learning approaches.
Boosts neural network performance by improving weight separability.
problem Improving the separability of weight vectors in neural networks.
method Proposes a new evaluation metric and feed-backward reconstruction loss to encourage weight separability.
result Improves visual recognition performance across various tasks.
Bayesian approach uses generative models as priors for better source separation.
problem Artifacts in source separation for richly structured data.
method Bayesian approach with generative models as priors and noise-annealed Langevin dynamics.
result Achieves state-of-the-art performance for MNIST digit separation.
New examples show some convex-cocompact subgroups are separable.
problem Whether all convex-cocompact subgroups are separable.
method Using Manning-Mj-Sageev construction, examples of separable subgroups of arbitrary finite rank are given.
result Examples of separable convex-cocompact subgroups of arbitrary finite rank exist.
Logarithmic separation profile in hyperbolic groups shows hierarchical structure.
problem Understanding hierarchical structure in hyperbolic groups with logarithmic separation.
method Proving groups with logarithmic separation split over cyclic groups and providing counterexamples.
result Not all groups with hierarchical structure have logarithmic separation profile.
New machine learning method detects quantum separability in large-scale systems.
problem Deciding quantum separability of large-scale bipartite density matrices.
method Frank-Wolfe-based algorithm for finding nearest separable density matrices and classification of density matrices as separable or entangled.
result The method scales up to thousands of density matrices and achieves high quantum entanglement detection accuracy.
Non-relatively hyperbolic separating curve graph for surfaces.
problem Classifying hyperbolicity of separating curve graphs.
method Proof of non-relatively hyperbolic property.
result Separating curve graph is not relatively hyperbolic for surfaces with genus ≥ 3 and one boundary component.
The study shows subgroup separability conditions for specific groups.
problem Conditions for subgroup separability in free-by-cyclic and deficiency 1 groups.
method Analyzes polynomially growing monodromy and asymptotic probability of random groups.
result Random deficiency 1 groups are not subgroup separable with positive probability.
The paper examines subgroup separability for surface and virtual braid groups.
problem Subgroup separability of surface and virtual braid groups.
method Study of subgroup separability (LERF) properties.
result Properties of subgroup separability for surface and virtual braid groups are explored.
Prob-PIT improves speech separation by considering output-label permutations as random variables.
problem Overconfident output-label assignment in PIT leads to unreliable speech separation.
method Prob-PIT treats output-label permutations as a discrete latent random variable with a uniform prior distribution and maximizes the log-likelihood function.
result Prob-PIT significantly outperforms PIT in terms of Signal to Distortion Ratio and Signal to Interference Ratio.
The study of random surfaces reveals asymptotic lengths of separating geodesics.
problem Understanding geometric properties of random hyperbolic surfaces.
method Analysis of Weil-Petersson measure and asymptotic behavior of lengths.
result The shortest separating closed geodesics have lengths about 2logg. We present a monophonic source separation system that is trained by only observing mixtures with no ground truth separation information. We use a deep clustering approach which trains on multi-channel mixtures and learns to project spectrogram bins to source clusters that correlate with various spatial features. We sho…
A new measure DCSI quantifies separability for density-based clustering.
problem Quantifying meaningful clusters in data sets.
method Developed a new separability measure DCSI based on separation and connectedness.
result Correctly identifies touching or overlapping classes that do not correspond to meaningful density-based clusters.
A new NMF variant tackles underdetermined problems with sparse and separable assumptions.
problem Underdetermined blind source separation, especially multispectral image unmixing.
method Sparse Separable Nonnegative Matrix Factorization (SSNMF) combining separability and sparsity assumptions. Algorithm based on SNPA and sparse nonnegative least squares.
result In noiseless settings, the algorithm recovers true underlying sources.
This work examines a semi-blind single-channel source separation problem. Our specific aim is to separate one source whose local structure is approximately known, from another a priori unspecified background source, given only a single linear combination of the two sources. We propose a separation technique based on lo…
One approach to monitoring a dynamic system relies on decomposition of the system into weakly interacting subsystems. An earlier paper introduced a notion of weak interaction called separability, and showed that it leads to exact propagation of marginals for prediction. This paper addresses two questions left open by t…
Develops large-sample theory for non-stationary source separation.
problem Lack of large-sample results for non-stationary source separation methods.
method Large-sample theory for NSS-JD method under specific assumptions.
result Consistency of unmixing estimator and its convergence to Gaussian distribution.
Sound source separation has attracted attention from Music Information Retrieval(MIR) researchers, since it is related to many MIR tasks such as automatic lyric transcription, singer identification, and voice conversion. In this paper, we propose an intuitive spectrogram-based model for source separation by adapting U-…
This work learns sparse tensor representations using mixtures of separable dictionaries.
problem Learning sparse representations of tensor data with structured models.
method Proposes and explores learning a mixture of separable dictionaries with sufficient conditions for local identifiability.
result Developed computational algorithms for batch and online learning.
Criterion for subgroup separability in outer automorphism groups.
problem Subgroup separability in outer automorphism groups.
method Criterion for separability of subgroups.
result Strengthening and generalizing a previous result on mapping class groups.
Study on hyperbolic groups, focusing on separability and splittings.
problem Coarse separability and splittings in hyperbolic groups.
method Quantitative analysis of volume growth and cut-sets, focusing on thickened spheres.
result One-ended hyperbolic groups that are not virtually surface groups are coarsely separable by a subset of subexponential growth if and only if they split over a virtually cyclic subgroup.
The paper solves optimal bounds for separating data points in high dimensions.
problem Correcting AI errors and analyzing vulnerabilities in high-dimensional data.
method General stochastic separation theorems with optimal probability estimates.
result Explicit and optimal estimates of separation probabilities for important classes of distributions.
Wavesplit separates speech from mixtures using clustering.
problem Permutation problem in speech separation.
method End-to-end system infers speaker representations and estimates signals.
result Robust separation of long recordings, new benchmarks set.
Suppose that all hyperbolic groups are residually finite. The following statements follow: In relatively hyperbolic groups with peripheral structures consisting of finitely generated nilpotent subgroups, quasiconvex subgroups are separable; Geometrically finite subgroups of non-uniform lattices in rank one symmetric sp…
The paper introduces toric separable geometries and finds new extremal metrics.
problem Finding explicit extremal Kähler metrics on toric manifolds.
method Introducing toric separable geometries and analyzing their moduli space.
result Explicit computation of scalar curvature and derivation of necessary conditions for extremality.
Singing voice separation attempts to separate the vocal and instrumental parts of a music recording, which is a fundamental problem in music information retrieval. Recent work on singing voice separation has shown that the low-rank representation and informed separation approaches are both able to improve separation qu…
Separating an audio scene into isolated sources is a fundamental problem in computer audition, analogous to image segmentation in visual scene analysis. Source separation systems based on deep learning are currently the most successful approaches for solving the underdetermined separation problem, where there are more …
Adaptive algorithm reduces regret in causal bandits.
problem Minimize regret in causal bandits with unknown d-separators.
method Adaptive algorithm exploiting d-separators without prior knowledge.
result Significantly smaller regret than previous methods.
Improved speech separation and enhancement using neural beamforming.
problem Challenging speech separation and enhancement in reverberant environments.
method Sequential neural beamforming combining spectral and spatial separation methods.
result Average improvement of 2.75 dB in scale-invariant signal-to-noise ratio and 14.2% absolute reduction in speech recognition metric.
Study optimizes fairness in predictive models by balancing utility and separation.
problem Balancing fairness and utility in predictive models.
method Information-theoretic approach using conditional mutual information (CMI).
result Reduces separation violations while maintaining or improving utility.