Improved sampling method using regularized Stein Variational Gradient Flow.
problem Improving the accuracy of sampling methods in machine learning.
method Proposed Regularized Stein Variational Gradient Flow to interpolate between SVGD and Wasserstein Gradient Flow.
result Established theoretical properties and provided preliminary numerical evidence of improved performance.
Paper explores Fisher-Rao gradient flows and their kernel approximations.
problem Understanding and analyzing approximations of Fisher-Rao gradient flows.
method Rigorous investigation of Fisher-Rao and Wasserstein type gradient flows, focusing on kernel approximations.
result Proves evolutionary Γ-convergence for kernel-approximated Fisher-Rao flows, providing theoretical guarantees.
New gradient flows for non-negative and probability measures combining optimal transport and interaction forces.
problem Optimizing non-negative and probability measures using interaction forces and optimal transport.
method Interaction-Force Transport (IFT) gradient flows and their spherical variant, developed via infimal convolution of Wasserstein and spherical MMD tensors, with a particle-based optimization algorithm.
result The spherical IFT gradient flow provides global exponential convergence guarantees for both MMD and KL energy.
This paper explores gradient flows for sampling distributions without normalization constants.
problem Sampling from distributions with unknown normalization constants.
method Gradient flows in the space of probability measures, focusing on Kullback-Leibler divergence, Fisher-Rao metric, and affine invariance.
result Gradient flows derived from Kullback-Leibler divergence do not depend on the normalization constant.
This paper bridges variational inference and Wasserstein gradient flows.
problem Combining variational inference and Wasserstein gradient flows for more efficient approximations.
method Recasting Bures-Wasserstein gradient flow as a Euclidean gradient flow and using path-derivative gradient estimator.
result A new gradient estimator for f-divergences that can be implemented using machine learning libraries. Constructs explicit solutions to Spin(7)-structures gradient flow.
problem Finding explicit solutions to Spin(7)-structures gradient flow.
method Expressed Spin(7)-torsion tensor and gradient flow in terms of torsion forms; used these formulae to find solutions.
result Found explicit solutions including a shrinking soliton on SU(3) and another on a T7-bundle over S1. Mean curvature flow is not a gradient flow on two nondegenerate metric spaces.
problem Whether mean curvature flow is a gradient flow on nondegenerate metric spaces of simple closed plane curves.
method Examined two nondegenerate metric spaces: uniformness-preserving and curvature-weighted structures.
result Mean curvature flow is not a gradient flow on either metric space.
Unified framework for analyzing gradient flows of measures with exponential decay of entropy.
problem Analyzing exponential decay of entropy functionals in gradient flows of measures.
method Characterization of global exponential decay behaviors using Hellinger-Kantorovich geometry, shape-mass decomposition, and Polyak-Łojasiewicz-type inequalities.
result Unified theoretical framework for gradient flows with complete analysis of exponential decay behaviors.
Gradient flow in parameters equals linear interpolation in outputs.
problem Understanding and optimizing training algorithms in deep learning.
method Proving equivalence between gradient flow in parameter space and linear interpolation in output space, and deriving formulas for global minima.
result Gradient flow in parameters can be transformed into linear interpolation in outputs, leading to global minima.
The article reviews how gradient flow systems on hypergraphs connect to information geometry and nonequilibrium physics.
problem Understanding the geometry of perturbed gradient flow systems on hypergraphs.
method Formulating modern nonequilibrium principles within the framework of perturbed gradient flow systems on hypergraphs.
result New concepts like moduli spaces and thermodynamical area are introduced to understand speed limits.
A new method for Gaussian filtering using gradient flows and Wasserstein metrics.
problem Approximating Gaussian and mixture-of-Gaussians filtering for complex systems.
method Variational approximation via gradient-flow representation on Wasserstein metric space.
result Competitive performance in posterior representation and parameter estimation for systems with multiplicative noise and multi-modal distributions.
A new gradient flow framework for distributionally robust optimization.
problem Optimizing under uncertainty with worst-case distributional constraints.
method Gradient flow theory applied to distributionally robust optimization.
result Practical algorithms for sampling from worst-case distributions.
Graph neural networks are explained through energy gradient flow and framelet decomposition.
problem Understanding and improving graph neural networks.
method Viewing framelet-based models as gradient flows of energy, proposing a generalized energy via framelet decomposition.
result The proposed model leads to more flexible dynamics, enhancing graph neural networks.
This paper studies gradient flows for sampling using various metrics and their affine invariance.
problem Sampling from probability distributions with unknown normalizations.
method Gradient flows in the space of probability measures, focusing on Kullback-Leibler divergence and affine invariance of metrics.
result Gradient flows of Kullback-Leibler divergence do not depend on the normalization constant, and affine invariance is achieved for certain metrics.
Forward-Euler fails for simulating Wasserstein gradient flows with KL divergence.
problem Simulating Wasserstein gradient flows with forward-Euler discretization fails for KL divergence.
method Forward-Euler discretization for Wasserstein gradient flows with KL divergence.
result Forward-Euler discretization can be incorrect for Wasserstein gradient flows with KL divergence.
Gradient flow on diffeomorphisms for image registration, with well-posedness proven.
problem Image registration with metric tensor deformation penalization.
method Gradient flow on Sobolev diffeomorphisms for a specific energy functional.
result Well-posedness of the gradient flow established.
New method for constrained sampling using gradient flows.
problem Sampling from constrained domains.
method Introducing a boundary condition for gradient flow to confine particles within the domain.
result Provable continuous-time convergence in total variation for constrained sampling.
Gradient flows of neural networks converge to optimal values or diverge, with thresholds and asymptotic behaviors.
problem Understanding the convergence and divergence of gradient flows in neural networks.
method Analysis of gradient flows on loss landscapes of neural networks using o-minimal structures.
result Gradient flows either converge to optimal values or diverge to infinity, with thresholds and asymptotic behaviors.
Gradient flow studies Spin(7)-structures on compact 8-manifolds.
problem Formulating and studying the gradient flow of Spin(7)-structures.
method Negative gradient flow of an energy functional of Spin(7)-structures.
result Short-time existence and uniqueness of solutions to the flow.
The paper constructs optimal confidence bands for kernel gradient flow estimators.
problem Estimating generalization error and constructing confidence bands for kernel gradient flows.
method Established convergence rates and constructed optimal confidence bands under capacity-source condition.
result Optimal confidence bands for kernel gradient flows have shrinkage rates close to minimax optimal rates.
The paper tackles safe reinforcement learning with convex regularization.
problem Safe reinforcement learning in complex, high-dimensional settings with safety constraints.
method Doubly-regularized RL framework combining reward and parameter regularization, formulated as a convex regularized objective with parametrized policies on an infinite-dimensional statistical manifold.
result Exponential convergence guarantees under sufficient regularization, robust theoretical insights and guarantees for safe RL.
Gradient flow expands curves to round shapes.
problem Expanding curves to round shapes.
method Steepest descent L2-gradient flow of entropy.
result Flow converges to a round expanding circle for various initial curves.
New gradient flows improve high-dimensional sampling.
problem Sampling from high-dimensional target densities.
method Introducing Radon--Wasserstein gradient flows.
result Linear scaling in particles and dimensions.
Paper analyzes inclusive KL inference using Wasserstein gradient flows.
problem Analyzing inclusive KL inference with mathematical tools.
method Gradient flows derived from PDE analysis.
result Unified view of existing sampling algorithms as inclusive-KL inference.
New method uses diffusion models to solve inverse problems.
problem Solving ill-posed inverse problems with powerful priors.
method Formulate posterior sampling as a regularized Wasserstein gradient flow in latent space.
result Demonstrates improved performance on standard benchmarks.
Gradient flows on graphons converge to curves on graphon space.
problem Optimizing functions on large, exchangeable graphs.
method Euclidean gradient flow on edge weights converges to a curve on graphon space.
result Gradient flows on graphons can be described as curves of maximal slope on graphon space.
Study on test risk dynamics in learning theory with stochastic gradient flow.
problem Understanding test risk in stochastic gradient flow dynamics.
method Path integral formulation for small learning rates, explicit computation for weak features.
result Explicit corrections due to stochastic term in dynamics, good agreement with simulations.
New method for natural policy gradients converges linearly.
problem Improving natural policy gradient methods for better convergence.
method Fisher-Rao gradient flow applied to state-action distributions.
result Linear convergence rate with geometry-dependent factor.
This paper analyzes convergence of large-scale Transformers with weight decay.
problem Understanding optimization guarantees in large-scale Transformer training.
method Construct mean-field limit, show gradient flow convergence to PDE, demonstrate global minimum consistency.
result Gradient flow reaches global minimum in large-scale Transformers with small weight decay.
The paper shows how gradient flow on over-parametrized tensor decomposition behaves like deflation.
problem Understanding the training dynamics of gradient flow on tensor decomposition.
method Empirical observation and mathematical proof of gradient flow dynamics for orthogonally decomposable tensors.
result Gradient flow dynamics for orthogonally decomposable tensors follows a tensor deflation process, recovering all tensor components.
We use the Yang-Mills gradient flow on the space of connections over a closed Riemann surface to construct a Morse-Bott chain complex. The chain groups are generated by Yang-Mills connections. The boundary operator is defined by counting the elements of appropriately defined moduli spaces of Yang-Mills gradient flow li…
This paper studies gradient flows in asymmetric metric spaces and proves existence results.
problem Investigating gradient flows in asymmetric metric spaces.
method Discrete approximation and natural convexity assumption on potential function.
result Existence of curves of maximal slope in asymmetric metric spaces.
This study explains gradient flow dynamics in neural networks for small initialisation.
problem Understanding the training dynamics of neural networks for small initialisation.
method Analysis of gradient flow dynamics for one-hidden layer ReLU networks with orthogonal inputs.
result Gradient flow converges to zero loss and characterizes implicit bias towards minimum variation norm.
The paper analyzes neural network dynamics after weights escape the origin.
problem Understanding gradient flow dynamics of neural networks after the origin.
method Analyzes gradient flow of homogeneous neural networks with locally Lipschitz gradients.
result Characterizes the first saddle point encountered after escaping the origin.
Study describes bifurcations of gradient flows on 2-sphere with holes.
problem Analyzing gradient flows on a 2-sphere with up to six singular points.
method Using separatrix diagrams to specify saddle-node and saddle connections.
result Identified all possible topological structures of bifurcations.
The paper defines and analyzes higher-order Yang-Mills-Higgs functionals and their gradient flows.
problem Analyzing the behavior of higher-order Yang-Mills-Higgs functionals and their gradient flows.
method Gauge fixing technique, L2-bound of the Higgs field, local L2-derivative estimates, energy estimates, blow-up analysis. result Solutions to the gradient flow do not hit finite time singularities under certain conditions.
Paper formulates particle flow using variational inference and Fisher-Rao gradient flow.
problem Estimating posterior densities in probabilistic models.
method Variational formulation of particle flow, Fisher-Rao gradient flow, Gaussian and Gaussian mixture approximations.
result Gaussian and Gaussian mixture approximations of Fisher-Rao particle flow reduce to Exact Daum and Huang particle flow under linear Gaussian assumptions.
We study the L2 gradient flow of the Yang--Mills functional on the space of connection 1-forms on a principal G-bundle over the sphere S2 from the perspective of Morse theory. The resulting Morse homology is compared to the heat flow homology of the space ΩG of based loops in the compact Lie group G. An iso…
New construction of Fukaya-Seidel categories using complex gradient flow equation.
problem Constructing Fukaya-Seidel categories for specific models.
method Using the complex gradient flow equation and neck-stretching limits.
result Alternative proof of Seidel's spectral sequence for Lagrangian Floer cohomology.
This study provides an explicit expansion of KL divergence's gradient flow in Fisher-Rao geometry.
problem Sampling techniques struggle to traverse between modes in non-convex potential functions.
method Explicit expansion of KL divergence's gradient flow in Fisher-Rao geometry.
result The convergence rate to π is independent of the potential function.
Gradient flow with weight decay shows grokking effect in deep learning.
problem Understanding the grokking effect in deep learning.
method Analyzing gradient flow dynamics with weight decay.
result Weight decay causes slow norm reduction, explaining grokking.
The paper develops statistical inference for gradient flows in optimization.
problem Uncertainty quantification along the entire optimization path.
method Uniform central limit theorem and algorithm-aware covariance estimator.
result Asymptotically valid confidence intervals for target parameter.
Study on multi-head softmax attention dynamics for in-context learning.
problem Understanding and optimizing multi-head softmax attention models for multi-task linear regression.
method Gradient flow analysis and spectral mapping technique.
result Gradient flow converges to optimal multi-head softmax attention model, with task allocation emerging during training.
Gradient flow of elastic energy converges to elastica.
problem Optimizing closed curves to minimize elastic energy.
method Proving the existence of a unique global solution and convergence via Łojasiewicz--Simon gradient inequality.
result Convergence to elastica established for the H2(ds)-gradient flow of modified elastic energy. Paper introduces a differentially private generative model using gradient flow and sliced Wasserstein distance.
problem Protecting privacy in sensitive training data for generative models.
method Gradient flow in the space of probability measures, Gaussian-smoothed Sliced Wasserstein Distance, and numerical scheme for SDE.
result Demonstrates higher-fidelity data generation at low privacy budget compared to existing methods.
Identifies a gradient flow to solve kernel learning problems with noise reduction.
problem Kernel learning problem with Gaussian noise.
method Riemannian gradient flow with continuous Lyapunov functionals.
result Flow reduces noise and finds stationary points.
Gradient flow method solves isoperimetric inequality for maps.
problem Finding maps with optimal enclosed area.
method Sobolev gradient flow for area-normalised Dirichlet energy.
result Solutions converge to a circle as time goes to infinity.
New particle-based VI algorithm expands function class and improves scalability.
problem Limited function class in particle-based VI algorithms restricts flexibility and scalability.
method Introduces a functional regularization term to expand the function class and proposes PFG algorithm.
result Proposed PFG algorithm has larger function class, improved scalability, better adaptation to ill-conditioned distributions, and provable convergence.