A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Establish optimal Lipschitz lower bounds for functions on manifolds with negative curvature, revealing interplay between width, boundary area, and topology.
problem Width estimates and rigidity of manifolds with negative curvature
method Gromov's μ-bubble method
result Sharp lower bound for boundary area in hyperbolic bands
We focus on estimating \emph{a priori} generalization error of two-layer ReLU neural networks (NNs) trained by mean squared error, which only depends on initial parameters and the target function, through the following research line. We first estimate \emph{a priori} generalization error of finite-width two-layer ReLU …
This article concerns the expressive power of depth in neural nets with ReLU activations and bounded width. We are particularly interested in the following questions: what is the minimal width wmin(d) so that ReLU nets of width wmin(d) (and arbitrary depth) can approximate any continuous functio…
The study explores positive scalar curvature metrics on non-orientable manifolds and their covers.
problem Existence of positive scalar curvature metrics on non-orientable manifolds and their covers.
method Extends Schoen-Yau inductive descent approach to non-orientable manifolds.
result Examples of non-orientable manifolds with positive scalar curvature metrics on their orientation double covers but not on homotopy equivalent manifolds.
This article concerns the expressive power of depth in deep feed-forward neural nets with ReLU activations. Specifically, we answer the following question: for a fixed din≥1, what is the minimal width w so that neural nets with ReLU activations, input dimension din, hidden layer widths at most w, and …
We prove that a bounded open set U in Euclidean n-space has k-width less than C(n) Volume(U)^{k/n}. Using this estimate, we give lower bounds for the k-dilation of degree 1 maps between certain domains in Euclidean space. In particular, we estimate the smallest (n-1)-dilation of any degree 1 map between two n-dimension…
This paper studies the expressive power of graph neural networks falling within the message-passing framework (GNNmp). Two results are presented. First, GNNmp are shown to be Turing universal under sufficient conditions on their depth, width, node attributes, and layer expressiveness. Second, it is discovered that GNNm…
This paper establishes the (nearly) optimal approximation error characterization of deep rectified linear unit (ReLU) networks for smooth functions in terms of both width and depth simultaneously. To that end, we first prove that multivariate polynomials can be approximated by deep ReLU networks of width $\mathcal{O}(N…
Better uncertainty estimates for neural networks using Gaussian process priors.
problem Poor uncertainty estimates in neural networks, especially on out-of-distribution data.
method Characterize the function-space prior of an ensemble of infinitely-wide neural networks as a Gaussian process and use it to build a probabilistic model.
result The approach improves calibration of neural networks, especially under distributional shift.
Given a Riemannian metric on a homotopy n-sphere, sweep it out by a continuous one-parameter family of closed curves starting and ending at point curves. Pull the sweepout tight by, in a continuous way, pulling each curve as tight as possible yet preserving the sweepout. We show: Each curve in the tightened sweepout …
We prove a sharp area estimate for catenoids that allows us to rule out the phenomenon of multiplicity in min-max theory in several settings. We apply it to prove that i) the width of a three-manifold with positive Ricci curvature is realized by an orientable minimal surface ii) minimal genus Heegaard surfaces in such …
Let M be a closed connected spin manifold such that its spinor Dirac operator has non-vanishing (Rosenberg) index. We prove that for any Riemannian metric on V=M×[−1,1] with scalar curvature bounded below by σ>0, the distance between the boundary components of V is at most Cn/σ, where $C_n = \…
We define the Wirtinger width of a knot. Then we prove the Wirtinger width of a knot equals its Gabai width. The algorithmic nature of the Wirtinger width leads to an efficient technique for establishing upper bounds on Gabai width. As an application, we use this technique to calculate the Gabai width of approximately …
The paper explores geometry and positive scalar curvature on non-compact manifolds.
problem Understanding the relationship between geometry and positive scalar curvature on non-compact manifolds.
method Analysis of volume growth, scalar curvature integral, and width in different dimensions.
result Proves minimal volume growth and integral of scalar curvature in three dimensions, and volume growth with stronger conditions in higher dimensions.
We propose a Generalized Dantzig Selector (GDS) for linear models, in which any norm encoding the parameter structure can be leveraged for estimation. We investigate both computational and statistical aspects of the GDS. Based on conjugate proximal operator, a flexible inexact ADMM framework is designed for solving GDS…
In the framework of Multifractal Diffusion Entropy Analysis we propose a method for choosing an optimal bin-width in histograms generated from underlying probability distributions of interest. The method presented uses techniques of Rényi's entropy and the mean squared error analysis to discuss the conditions under whi…
Lectures on deep learning properties in infinite and large-width networks.
problem Understanding deep neural networks in extreme width conditions.
method Analysis of random deep neural networks, connections to linear models, kernels, and Gaussian processes, perturbative and non-perturbative treatments.
result Properties and behaviors of deep neural networks in the infinite-width limit and large-width regime.
A number of results for C2-smooth surfaces of constant width in Euclidean 3-space E3 are obtained. In particular, an integral inequality for constant width surfaces is established. This is used to prove that the ratio of volume to cubed width of a constant width surface is reduced by shrinking it along…
Analysis of non-asymptotic estimation error and structured statistical recovery based on norm regularized regression, such as Lasso, needs to consider four aspects: the norm, the loss function, the design matrix, and the noise model. This paper presents generalizations of such estimation error analysis on all four aspe…
While studying the existence of closed geodesics and minimal hypersurfaces in compact manifolds, the concept of width was introduced in different contexts. Generally, the width is realized by the energy of the closed geodesics or the volume of minimal hypersurfaces, which are found by the Minimax argument. Recently, Ma…