Paper explores geometry of covariance matrices using associated bundles.
problem Geometry of fixed-rank covariance matrices.
method Associated bundle approach to Bures--Wasserstein geometry.
result Established a one-to-one correspondence between geodesics.
We reduce variance in Bures-Wasserstein variational inference.
problem High variance in Monte Carlo approximations of Bures-Wasserstein gradients.
method Control variates to reduce variance in the forward step.
result Proposed estimator reduces variance by orders of magnitude.
Study on deep matrix factorization with Bures-Wasserstein loss, focusing on critical points and convergence.
problem Analyzing critical points and convergence of generative deep linear networks trained with Bures-Wasserstein loss.
method Characterization of critical points and minimizers of Bures-Wasserstein distance, analysis of Hessian at low-rank matrices, convergence results for gradient flow and descent.
result Established convergence results for gradient flow and finite step size gradient descent under certain assumptions.
This paper studies geodesics between covariance matrices of different ranks using the Bures-Wasserstein metric.
problem Geodesics between covariance matrices of varying ranks.
method Analyzes the Bures-Wasserstein distance on covariance matrices, completing previous work on geodesics and providing explicit formulas.
result The set of all minimizing geodesics between two covariance matrices is parametrized by a closed unit ball in R(k−r)imes(l−r). Test partial effects in Frechet regression on Bures-Wasserstein manifolds.
problem Assessing partial effects in Frechet regression on complex manifolds.
method Sample splitting strategy to estimate covariance matrices and test statistic convergence.
result The test statistic converges to a weighted mixture of chi squared components.
This work tackles regression on non-Euclidean spaces, specifically positive-definite matrices with the Bures-Wasserstein metric.
problem Regression on non-Euclidean spaces, specifically positive-definite matrices with the Bures-Wasserstein metric.
method Developed a sufficient condition for the existence of a minimizer of the conditional barycenter problem, characterized the optimization landscape, and developed a projection-free algorithm for approximate computation of first-order stationary points.
result The objective is free of local maxima under the sufficient condition, and the algorithm enables the use of stochastic Riemannian optimization methods for large-scale setups.
Paper introduces a generalized Bures-Wasserstein geometry for SPD matrices.
problem Understanding the geometry of SPD matrices for machine learning.
method Proposes a generalized Bures-Wasserstein geometry parameterized by a symmetric positive definite matrix.
result The GBW geometry outperforms the BW geometry in machine learning applications.
Paper develops statistical tests for covariance matrix regression on manifold.
problem Regression with random covariance matrices in Fréchet space.
method Develops Wasserstein F-tests for Bures-Wasserstein manifold.
result Asymptotic null distribution and power of the test.
Improved stability for matrix recovery from rank-one measurements.
problem Phase retrieval problem of recovering rank-one positive semidefinite matrices.
method Developed a smoothing Newton method based on Bures-Wasserstein gradient descent.
result Superlinear convergence with rigorous guarantees and stable implementation.
Geometric approach to quantum thermodynamics models state spaces and processes.
problem Quantum thermodynamics in the regime of non-equilibrium states.
method Contact geometry and principal fiber bundles to model quantum state spaces and processes.
result Geometric formulation reveals the fundamental thermodynamic relations and unattainability of the third law.
This paper bridges variational inference and Wasserstein gradient flows.
problem Combining variational inference and Wasserstein gradient flows for more efficient approximations.
method Recasting Bures-Wasserstein gradient flow as a Euclidean gradient flow and using path-derivative gradient estimator.
result A new gradient estimator for f-divergences that can be implemented using machine learning libraries. Optimal transport (OT)-based methods have a wide range of applications and have attracted a tremendous amount of attention in recent years. However, most of the computational approaches of OT do not learn the underlying transport map. Although some algorithms have been proposed to learn this map, they rely on kernel-ba…
Improved VI with Price's gradient estimator for target log-density.
problem Approximating target distributions from unnormalized log-densities.
method Stochastic gradient-based variational inference with Price's gradient estimator.
result Identifies Price's gradient as the key to WVI's superior performance.
New PAC-Bayes bounds use Wasserstein distances to improve generalization.
problem Lack of geometric properties in existing PAC-Bayes bounds.
method Developed new PAC-Bayes bounds with Wasserstein distances.
result Optimization guarantees translate to good generalization abilities.
This work proposes new methods for variational inference using gradient flows on Gaussian measures.
problem Developing algorithmic guarantees for variational inference.
method Proposes principled methods for variational inference using gradient flows on the Bures--Wasserstein space of Gaussian measures.
result Strong theoretical guarantees for log-concave posteriors.
Proposes a new method to improve regression models with reweighted samples.
problem Improves regression models' performance under low sample sizes and covariate perturbations.
method Reparametrizes sample weights using a doubly non-negative matrix and solves the reweighted estimate efficiently.
result Adversarial reweighting strategy delivers promising results on various datasets.
Python package for SPD matrix distances, reproducible and extensible.
problem Computing distances between SPD matrices for various applications.
method Unified, extensible framework supporting multiple SPD metrics.
result Reproducible and accessible SPD matrix comparison tool.
New method for optimal transport with missing data, debiased and efficient.
problem Solving optimal transport between two distributions with missing values.
method Debiasing Wasserstein distance for empirical Gaussian distributions, entropic regularized optimal transport using ISVT.
result Efficient and consistent estimation of entropic regularized optimal transport.
New SPD metrics improve stability and efficiency in neural networks.
problem Designing stable and efficient Riemannian metrics on SPD manifolds.
method Cholesky decomposition to derive SPD metrics.
result Proposed metrics provide closed-form operators, computational efficiency, and improved numerical stability.
The paper presents two schemes for sampling matrices from specific distributions on a manifold.
problem Sampling matrices from Gibbs distributions on the manifold of positive semi-definite matrices with fixed rank.
method Two explicit schemes based on Euler-Maruyama discretization of the Riemannian Langevin equation with Brownian motion on the manifold.
result Numerical validation of the schemes using specific energy functions and metrics.
Extends metrics for SPD matrices to infinite dimensions.
problem Lack of generalized forms for Riemannian metrics.
method Unitized Hilbert-Schmidt operators and extended Mahalanobis norm.
result Improved performance in high-dimensional comparisons.
We propose a novel framework for graph mean computation.
problem Defining a mean for graph data is difficult.
method Embeddings in the space of smooth graph signal distributions, using the Wasserstein metric.
result Existence and uniqueness of the graph mean established, and an iterative algorithm provided.
Proposes a variational NNCC formulation for infinite dimensions.
problem Optimization and gradient flows in infinite-dimensional settings.
method Variational formulation of NNCC on c-convex domains.
result Wasserstein spaces inherit NNCC from their base space.
ITSPACE improves covariance alignment faster than other methods.
problem Optimizing covariance matrices for machine learning tasks.
method Proximal majorization-minimization method that directly optimizes the Bures-Wasserstein objective.
result ITSPACE achieves lower BW gap solutions faster than other methods.
Inversion-free natural gradient method for Riemannian manifolds.
problem Hindered by the need for Euclidean space, Fisher information matrix inversion, and computational cost.
method Intrinsic, inversion-free natural gradient method on Riemannian manifolds, using moving approximation of inverse FIM.
result Almost-sure convergence rates and sub-quadratic storage complexity for large-scale applications.
Investigates O(n)-invariant metrics on SPD matrices, extending kernel metrics.
problem Limited coverage of O(n)-invariant metrics by kernel metrics.
method Characterization of O(n)-invariant metrics, intermediate classes construction.
result Introduction of cometric-stability as a key property for geodesics.
BWFlow improves graph generation by smoothly interpolating graph components.
problem Disjoint modeling of graph nodes and edges leads to irregular and non-smooth probability paths.
method Modeling graphs as MRFs and using optimal transport displacement for a smooth probability path.
result BWFlow achieves better training convergence and efficient sampling in graph generation.
The paper proves geometric and spectral alignment for deep neural networks.
problem Understanding the singular spectra of deep neural network layers.
method Proves deterministic quotient-geometric estimates for singular spectra of Frobenius-normalized layer factors.
result Exact power-law spectra form a trace-normalized Cartan orbit under Frobenius normalization.
JKO scheme adds deceleration in rapidly changing metric curvature directions.
problem Understanding the implicit bias of the JKO scheme in Wasserstein gradient flow.
method Characterized the implicit bias of the JKO scheme at second order in η, modifying the energy functional.
result JKO scheme adds deceleration in directions where metric curvature of J is rapidly changing.