Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

130260390520 · Jun 202019922001200920172026
48 results for free multiplicative convolution

Study on the geometric Dyson Brownian motion of non-square matrix products.

problem Understanding the spectrum of a product of non-square random matrices.
method Proportional depth-width limit followed by mean-field limit, solving Burgers equation.
result Free log-normal law is obtained in the identity-start case.

We analyze the eigenvalue distribution of a neural network's kernel under specific scaling.

problem Analyzing the eigenvalue distribution of the Neural Tangent Kernel (NTK) of a neural network.
method Asymptotic analysis of the NTK matrix under given scaling conditions.
result The eigenvalue distribution is described as a free multiplicative convolution of the Marchenko-Pastur distribution and a deterministic distribution.

Paper analyzes the free energy of CNNs with skip connections in Bayesian learning.

problem Dependency of CNNs with skip connections on the number of parameters.
method Examines the Bayesian free energy of CNNs with and without skip connections.
result The upper bound of free energy of Bayesian CNN with skip connections does not depend on overparametrization.

Ensemble learning is a method of combining multiple trained models to improve model accuracy. We propose the usage of such methods, specifically ensemble average, inside Convolutional Neural Network (CNN) architectures by replacing the single convolutional layers with Inner Average Ensembles (IEA) of multiple convoluti…

2018-08-30abs ↗pdf ↗

New method predicts neural network performance using free probability theory.

problem Stability and performance prediction of feed-forward neural networks.
method Free Probability Theory and homotopy method for Jacobian spectral density computation.
result FPT metrics correlate highly with final test accuracies of neural networks.

In this letter, we generalize the convolutional NMF by taking the ββ-divergence as the contrast function and present the correct multiplicative updates for its factors in closed form. The new updates unify the ββ-NMF and the convolutional NMF. We state why almost all of the existing updates are inexact and approximat…

2018-03-14abs ↗pdf ↗

Let (M,ω)(M,ω) be a connected symplectic manifold on which a connected Lie group GG acts properly and in a Hamiltonian fashion with moment map $μ:M \lra \mf g^*$. Our purpose is investigate multiplicity-free actions, giving criteria to decide a multiplicity freenes of the action. As an application we give the complete cl…

2004-02-17abs ↗pdf ↗

Study on Bayesian deep linear networks with multiple outputs and convolutional layers.

problem Characterize feature learning in finite-width Bayesian deep linear networks.
method Exact and analytical formulas for joint and posterior distributions, using large deviation theory.
result Quantitative description of feature learning in infinite-width regime.

Yes, they do. This paper provides the first empirical demonstration that deep convolutional models really need to be both deep and convolutional, even when trained with methods such as distillation that allow small or shallow models of high accuracy to be trained. Although previous research showed that shallow feed-for…

2016-03-17abs ↗pdf ↗

A (quasi-)Hamiltonian manifold is called multiplicity free if all of its symplectic reductions are 0-dimensional. In this paper, we classify multiplicity free Hamiltonian actions for (twisted) loop groups or, equivalently, multiplicity free (twisted) quasi-Hamiltonian manifolds for simply connected compact Lie groups. …

2016-12-12abs ↗pdf ↗

Analog arrays are a promising upcoming hardware technology with the potential to drastically speed up deep learning. Their main advantage is that they compute matrix-vector products in constant time, irrespective of the size of the matrix. However, early convolution layers in ConvNets map very unfavorably onto analog a…

2018-07-03abs ↗pdf ↗

Paper connects algebraic and analytic methods for braid group representations.

problem Constructing representations of braid groups using algebraic and analytic approaches.
method Katz-Long-Moody construction and multiplicative middle convolution for KZ-type equations.
result Multiplicative middle convolution preserves unitarity and provides an algorithm to determine the signature of a Hermitian matrix.

We present an alternative layer to convolution layers in convolutional neural networks (CNNs). Our approach reduces the complexity of convolutions by replacing it with binary decisions. Those binary decisions are used as indexes to conditional distributions where each weight represents a leaf in a decision tree. This m…

2019-05-24abs ↗pdf ↗

Winograd convolution is widely used in deep neural networks (DNNs). Existing work for DNNs considers only the subset Winograd algorithms that are equivalent to Toom-Cook convolution. We investigate a wider range of Winograd algorithms for DNNs and show that these additional algorithms can significantly improve floating…

2019-05-13abs ↗pdf ↗

In this work we detail the application of a fast convolution algorithm computing high dimensional integrals to the context of multiplicative noise stochastic processes. The algorithm provides a numerical solution to the problem of characterizing conditional probability density functions at arbitrary time, and we applie…

2011-07-07abs ↗pdf ↗

We compute the sheaf of automorphisms of a multiplicity free Hamiltonian manifold over its momentum polytope and show that its higher cohomology groups vanish. Together with a theorem of Losev, arXiv:math/0612561, this implies a conjecture of Delzant: a compact multiplicity free Hamiltonian manifold is uniquely determi…

2010-02-23abs ↗pdf ↗

ViLT is a faster vision-and-language model without convolution or region supervision.

problem Efficiency and expressive power limitations in current VLP models.
method A convolution-free Vision-and-Language Transformer (ViLT) that processes visual inputs similarly to textual inputs.
result ViLT is up to tens of times faster with competitive or better performance.

Two new inverse-free ELM algorithms for incremental and decremental learning are proposed.

problem Efficiently updating and removing multiple hidden nodes in ELM.
method Improved inverse-free recursive algorithms for Tikhonov regularization.
result Inverse-free algorithms for ELM with multiple hidden nodes and redundant nodes.

Study shows deterministic equivalent for neural network kernel convergence.

problem Understanding convergence of neural network kernels.
method Analyzes empirical spectral distribution of Conjugate Kernel, proving convergence to a deterministic limit.
result Obtains a deterministic equivalent for the Stieltjes transform and resolvent of the Conjugate Kernel.

CNN predicts stock fluctuations using company news headlines.

problem Predicting next-day stock fluctuations based on company-specific news.
method Convolutional Neural Network (CNN) with reduced filter dimensions and multiple hidden layers. Fine-tuned word embeddings and various filter widths.
result 61.7% classification accuracy achieved using pre-learned embeddings.

The introduction of convolutional layers greatly advanced the performance of neural networks on image tasks due to innately capturing a way of encoding and learning translation-invariant operations, matching one of the underlying symmetries of the image domain. In comparison, there are a number of problems in which the…

2016-12-14abs ↗pdf ↗

We show that Tolman's example (of a six dimensional Hamiltonian T2T^2-space with isolated fixed points and no compatible Kähler structure) can be constructed from the flag variety U(3)/U(1)3U(3)/U(1)^3 by U(2)U(2)-equivariant symplectic surgery. This implies that Tolman's space has a ``transversal multiplicity-free'' action of $…

1995-06-22abs ↗pdf ↗

Functor connects Lie groupoid algebras to bornological structures.

problem Establishing a functorial relationship between Lie groupoid convolution algebras and bornological structures.
method Developed a monoidal functor from differentiable stacks to Morita 2-category of complete bornological algebras.
result Convolution algebras are self-induced and convolution modules are smooth.

Post-hoc explanations improve CNNs by replacing final linear layer with k-means classifier.

problem CNNs lack accurate data representation in their built-in prototypes.
method Introduces k-means-based post-hoc explanations for CNNs, leveraging spatial consistency of convolutional receptive fields.
result Using shallower, less compressed feature activations improves semantic fidelity at the cost of slight predictive performance.

We classify compact, connected Hamiltonian and quasi-Hamiltonian manifolds of cohomogeneity one (which is the same as being multiplicity free of rank one). Here the group acting is a compact connected Lie group (simply connected in the quasi-Hamiltonian case). This work is a concretization of the more general classific…

2019-10-04abs ↗pdf ↗

In this paper, we develop a method for unsupervised clustering of two-way (matrix) data by combining two recent innovations from different fields: the Sparse Subspace Clustering (SSC) algorithm [10], which groups points coming from a union of subspaces into their respective subspaces, and the t-product [18], which was …

2014-12-22abs ↗pdf ↗

The study proves the existence of free boundary minimal disks in convex regions.

problem Proving the existence of free boundary minimal disks in convex regions.
method Based on a multiplicity-one theorem for the free boundary Simon-Smith min-max theory.
result Existence of at least three embedded free boundary minimal disks in strictly convex domains with nonnegative Ricci curvature.

Health professionals can use natural language processing (NLP) technologies when reviewing electronic health records (EHR). Machine learning free-text classifiers can help them identify problems and make critical decisions. We aim to develop deep learning neural network algorithms that identify EHR progress notes perta…

2018-09-16abs ↗pdf ↗

Deep learning has been widely applied and brought breakthroughs in speech recognition, computer vision, and many other domains. The involved deep neural network architectures and computational issues have been well studied in machine learning. But there lacks a theoretical foundation for understanding the approximation…

2018-05-28abs ↗pdf ↗

Proves multiplicity one for boundary minimal hypersurfaces in compact manifolds.

problem Proving multiplicity one for min-max free boundary minimal hypersurfaces in compact manifolds with boundary.
method Developed existence and regularity theory for free boundary hypersurfaces with prescribed mean curvature, including Morse index bounds.
result Proved multiplicity one theorem for min-max free boundary minimal hypersurfaces in compact manifolds with boundary.

Superior performance and ease of implementation have fostered the adoption of Convolutional Neural Networks (CNNs) for a wide array of inference and reconstruction tasks. CNNs implement three basic blocks: convolution, pooling and pointwise nonlinearity. Since the two first operations are well-defined only on regular-s…

2018-03-06abs ↗pdf ↗

Study multiplicity-free covering of graded manifolds, proving equivalence of categories.

problem Equivalence of categories of graded manifolds and symmetric vector bundles.
method Defined and computed multiplicity-free covering, showed deck transformation group isomorphic to SnS_n.
result Equivalence of categories of graded manifolds and symmetric nn-fold vector bundles.

Learning representation on graph plays a crucial role in numerous tasks of pattern recognition. Different from grid-shaped images/videos, on which local convolution kernels can be lattices, however, graphs are fully coordinate-free on vertices and edges. In this work, we propose a Gaussian-induced convolution (GIC) fra…

2018-11-11abs ↗pdf ↗

Deep neural network compression techniques such as pruning and weight tensor decomposition usually require fine-tuning to recover the prediction accuracy when the compression ratio is high. However, conventional fine-tuning suffers from the requirement of a large training set and the time-consuming training procedure. …

2018-12-05abs ↗pdf ↗