A new NMF variant tackles underdetermined problems with sparse and separable assumptions.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Generative source separation methods such as non-negative matrix factorization (NMF) or auto-encoders, rely on the assumption of an output probability density. Generative Adversarial Networks (GANs) can learn data distributions without needing a parametric assumption on the output density. We show on a speech source se…
The separability assumption (Donoho & Stodden, 2003; Arora et al., 2012) turns non-negative matrix factorization (NMF) into a tractable problem. Recently, a new class of provably-correct NMF algorithms have emerged under this assumption. In this paper, we reformulate the separable NMF problem as that of finding the ext…
Develops large-sample theory for non-stationary source separation.
New algorithms SVCA and SSPA improve robustness to noise in nonnegative matrix factorization.
Algorithm clusters mixtures with bounded covariances under specific separation conditions.
The paper solves optimal bounds for separating data points in high dimensions.
New graph types help identify complex relationships.
The paper proves local laws for non-separable sample covariance matrices.
In this paper, we presented a novel semi-supervised one-class classification algorithm which assumes that class is linearly separable from other elements. We proved theoretically that class is linearly separable if and only if it is maximal by probability within the sets with the same mean. Furthermore, we presented an…
SS-SARSA tackles recovering bandits by treating rounds as states.
Quantum computers outperform classical methods in density modeling.
Source separation problems are ubiquitous in the physical sciences; any situation where signals are superimposed calls for source separation to estimate the original signals. In this tutorial I will discuss the Bayesian approach to the source separation problem. This approach has a specific advantage in that it require…
Nonnegative matrix factorization (NMF) is a linear dimensionality technique for nonnegative data with applications such as image analysis, text mining, audio source separation and hyperspectral unmixing. Given a data matrix and a factorization rank , NMF looks for a nonnegative matrix with columns and a …
Directed graphical models provide a useful framework for modeling causal or directional relationships for multivariate data. Prior work has largely focused on identifiability and search algorithms for directed acyclic graphical (DAG) models. In many applications, feedback naturally arises and directed graphical models …
In this paper, we study the nonnegative matrix factorization problem under the separability assumption (that is, there exists a cone spanned by a small subset of the columns of the input nonnegative data matrix containing all columns), which is equivalent to the hyperspectral unmixing problem under the linear mixing mo…
Separating mixed distributions is a long standing challenge for machine learning and signal processing. Most current methods either rely on making strong assumptions on the source distributions or rely on having training samples of each source in the mixture. In this work, we introduce a new method---Neural Egg Separat…
A new NMF model for co-clustering and data approximation.
We simplify information measure computation using learned features.
Proposes a framework to balance supervised and unsupervised learning using random matrix theory.
A new convex model tackles noisy separable NMF with provable correctness.
Independent component analysis (ICA) is a powerful method for blind source separation based on the assumption that sources are statistically independent. Though ICA has proven useful and has been employed in many applications, complete statistical independence can be too restrictive an assumption in practice. Additiona…
DDICA separates nonlinear mixed signals robustly.
Mirror flow optimizes separable data problems, converging to a maximum margin classifier.
Adaptive algorithm reduces regret in causal bandits.
New framework for understanding BSS robustness under model violations.
We introduce a convex approach for mixed linear regression over features. This approach is a second-order cone program, based on L1 minimization, which assigns an estimate regression coefficient in for each data point. These estimates can then be clustered using, for example, -means. For problem…
New algorithm learns POMDPs without computational oracles.
The paper establishes a nearly-sharp statistical threshold for efficient learning in Latent MDPs with separated components.
SepIt improves speech separation for multiple speakers.
Numerous algorithms are used for nonnegative matrix factorization under the assumption that the matrix is nearly separable. In this paper, we show how to make these algorithms efficient for data matrices that have many more rows than columns, so-called "tall-and-skinny matrices". One key component to these improved met…
Nonnegative matrix factorization (NMF) has been shown to be identifiable under the separability assumption, under which all the columns(or rows) of the input data matrix belong to the convex cone generated by only a few of these columns(or rows) [1]. In real applications, however, such separability assumption is hard t…
Recently, a family of tractable NMF algorithms have been proposed under the assumption that the data matrix satisfies a separability condition Donoho & Stodden (2003); Arora et al. (2012). Geometrically, this condition reformulates the NMF problem as that of finding the extreme rays of the conical hull of a finite set …
New algorithms improve blind source separation for linear-quadratic mixtures.
Nonnegative matrix factorization (NMF) under the separability assumption can provably be solved efficiently, even in the presence of noise, and has been shown to be a powerful technique in document classification and hyperspectral unmixing. This problem is referred to as near-separable NMF and requires that there exist…
For many interesting tasks, such as medical diagnosis and web page classification, a learner only has access to some positively labeled examples and many unlabeled examples. Learning from this type of data requires making assumptions about the true distribution of the classes and/or the mechanism that was used to selec…
Lowered regularity assumption for a phase-dependent Helfrich energy equation.
Study uniform rates for estimating Gaussian mixtures without separation assumption.
This paper is an attempt to separate cardiac and respiratory signals from an electrical bio-impedance (EBI) dataset. For this two well-known algorithms, namely Principal Component Analysis (PCA) and Independent Component Analysis (ICA), were used to accomplish the task. The ability of the PCA and the ICA methods first …
We consider an online version of the robust Principle Component Analysis (PCA), which arises naturally in time-varying source separations such as video foreground-background separation. This paper proposes a compressive online robust PCA with prior information for recursively separating a sequences of frames into spars…
Paper relaxes independence assumption for non-centered data.
We consider perturbed quadharmonic operators, , acting on sections of a Hermitian vector bundle over a complete Riemannian manifold, with the potential satisfying a bound from below by a non-positive function depending on the distance from a point. Under a bounded geometry assumption on the Hermitian vecto…
Proposes a weaker faithfulness assumption for causal discovery.
In this paper, we propose a provably correct algorithm for convolutive nonnegative matrix factorization (CNMF) under separability assumptions. CNMF is a convolutive variant of nonnegative matrix factorization (NMF), which functions as an NMF with additional sequential structure. This model is useful in a number of appl…
DISCoVeR learns disentangled representations by separating shared and condition-specific factors.
Root Cause Analysis for Anomalies is challenging because of the trade-off between the accuracy and its explanatory friendliness, required for industrial applications. In this paper we propose a framework for simple and friendly RCA within the Bayesian regime under certain restrictions (that Hessian at the mode is diago…
We derive and analyze a new, efficient, pool-based active learning algorithm for halfspaces, called ALuMA. Most previous algorithms show exponential improvement in the label complexity assuming that the distribution over the instance space is close to uniform. This assumption rarely holds in practical applications. Ins…
Study provides guarantees for kernel clustering under non-parametric mixtures.