Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

59119178237 · Jun 202019922001200920172026
48 results for Orthogonal similarity

Study isotropy groups for complex orthogonal and skew-symmetric matrices.

problem Understanding isotropy subgroups of orthogonal similarity transformations.
method Analysis of group structure of nonsingular block matrices.
result Group structure of isotropy subgroups related to block Toeplitz matrices.

We construct a decomposition of the identity operator on a Riemannian manifold MM as a sum of smooth orthogonal projections subordinate to an open cover of MM. This extends a decomposition of the real line by smooth orthogonal projection due to Coifman, Meyer and Auscher, Weiss, Wickerhauser, and a similar decomposit…

2018-03-09abs ↗pdf ↗

Improves model predictability by mixing forecasts and orthogonalizing models.

problem Redundant models contaminate model space and degrade predictive performance.
method Principal Component Analysis for model orthogonalization in Bayesian forecast mixing.
result Better prediction accuracy and excellent uncertainty quantification.

New convergence guarantees for learning with unknown nuisance parameters.

problem Learning problems with unknown nuisance parameters.
method Stochastic gradient optimization with Neyman orthogonality and approximately orthogonalized updates.
result Stochastic gradient algorithms can converge under conditions of nuisance parameters.

Orthogonal deep models defend against black-box attacks by ensuring internal representations are nearly orthogonal.

problem Vulnerability of deep learning models to black-box adversarial attacks.
method Introduce a gradient regularization scheme to encourage deep models' internal representations to be orthogonal to another model's.
result Orthogonal deep models significantly boost robustness against transferable black-box adversarial attacks.

Different neural networks trained on the same dataset often learn similar input-output mappings with very different weights. Is there some correspondence between these neural network solutions? For linear networks, it has been shown that different instances of the same network architecture encode the same representatio…

2018-11-28abs ↗pdf ↗

Special orthogonal representations from octonions have geometric properties linked to binary cubics.

problem Understanding geometric properties of special orthogonal representations from octonions.
method Using octonions and their derivations, spinors, and covariants to show geometric properties.
result Covariants and Mathews identities of these representations are related to the Fano plane and (Z2)3(\mathbb{Z}_2)^3.

Muon optimizer simplifies matrix optimization with spectral orthogonalization.

problem Matrix optimization challenges, especially with large condition numbers.
method Simplified Muon optimizer using spectral orthogonalization of gradients.
result Simplified Muon converges linearly with independent scalar sequences, outperforming gradient descent and Adam.

Random convolutional networks can be fooled with adversarial examples.

problem Existence of adversarial examples for random convolutional networks.
method Utilizing isoperimetric inequalities on the special orthogonal group so(d)\mathbb{so}(d).
result Adversarial examples exist for various random convolutional networks.

State-of-the-art algorithms for sparse subspace clustering perform spectral clustering on a similarity matrix typically obtained by representing each data point as a sparse combination of other points using either basis pursuit (BP) or orthogonal matching pursuit (OMP). BP-based methods are often prohibitive in practic…

2017-10-31abs ↗pdf ↗

We consider the problem of sampling from posterior distributions for Bayesian models where some parameters are restricted to be orthogonal matrices. Such matrices are sometimes used in neural networks models for reasons of regularization and stabilization of training procedures, and also can parameterize matrices of bo…

2019-01-23abs ↗pdf ↗

We classify six-dimensional Lie groups which admit a left-invariant half-flat SU(3)-structure and which split in a direct product of three-dimensional factors. Moreover, a complete list of those direct products is obtained which admit a left-invariant half-flat SU(3)-structure such that the three-dimensional factors ar…

2009-12-17abs ↗pdf ↗

Improved Gaussian process models for interpretable predictions.

problem Complex responses require high-dimensional interaction terms in additive Gaussian processes.
method Orthogonal additive kernel (OAK) with orthogonality constraint on additive functions.
result OAK models achieve similar or better predictive performance with fewer terms, retaining interpretability.

Study asymptotics of extension and orthogonal Bergman kernels for high tensor powers of positive line bundles.

problem Asymptotic behavior of Bergman kernels for high tensor powers of positive line bundles.
method Analyzing the Schwartz kernel of the Ohsawa-Takegoshi extension operator and orthogonal Bergman projector, proving exponential estimates and asymptotic expansions.
result Explicit asymptotic expansions for the Ohsawa-Takegoshi extension operator and orthogonal Bergman projector.

Anti-transfer learning prevents misleading representations for speech tasks.

problem Misleading representations learned from orthogonal tasks in speech processing.
method Penalizes similarity between activations of a network and another trained on an orthogonal task.
result Improves classification accuracy and invariance to the orthogonal task.

Deterministic bounds for tensor singular values and vectors, differing from matrix cases.

problem Spectral learning of higher-order orthogonally decomposable tensors.
method Deterministic perturbation bounds for singular values and vectors of orthogonally decomposable tensors.
result Perturbation affects each essential singular value/vector in isolation, independent of multiplicity and distance from other singular values.

Defines a similarity measure for classification distributions.

problem Measuring similarity between classification distributions.
method Proposes task similarity, a novel measure quantifying performance of source distributions on target distributions.
result Empirical task similarity correlates with transfer efficiency and semantic similarity of source distributions.

OGD proves robustness to Catastrophic Forgetting in Continual Learning.

problem Catastrophic Forgetting in Continual Learning with deep neural networks.
method Theoretical framework based on Neural Tangent Kernel for OGD.
result First generalization bound for SGD and OGD in Continual Learning.

The paper shows how gradient flow on over-parametrized tensor decomposition behaves like deflation.

problem Understanding the training dynamics of gradient flow on tensor decomposition.
method Empirical observation and mathematical proof of gradient flow dynamics for orthogonally decomposable tensors.
result Gradient flow dynamics for orthogonally decomposable tensors follows a tensor deflation process, recovering all tensor components.

ORFit trains models on streaming data with one pass, minimizing memory and computational costs.

problem Training large models on a stream of data without retraining on previous data.
method Orthogonal Recursive Fitting (ORFit) using orthogonal gradient descent and recursive least-squares.
result ORFit updates parameters orthogonally to past gradients, leading to efficient memory and computational usage.

Let RR be an infinite commutative ring with identity and n2n\geq 2 be an integer. We prove that for each integer i=0,1,,n2,i=0,1,\cdots ,n-2, the L2L^{2}-Betti number bi(2)(G)=0,b_{i}^{(2)}(G)=0,  \ when G=GLn(R)G=\mathrm{GL}_{n}(R) the general linear group, SLn(R)\mathrm{SL}_{n}(R) the special linear group, % E_{n}(R) the group generated by…

2017-03-01abs ↗pdf ↗

New method improves reinforcement learning generalization.

problem Few environments lead to poor generalization in reinforcement learning.
method Integrates sequential structure into representation learning, using a policy similarity metric (PSM) and contrastive embeddings (PSEs).
result PSEs improve generalization across various benchmarks.

Overparameterized models improve performance in sequential learning tasks.

problem Catastrophic forgetting in overparameterized neural networks.
method Two-task linear regression problem with random orthogonal transformations.
result Overparameterization mitigates catastrophic forgetting in sequential learning tasks.

In this paper we explore the "vector semantics" problem from the perspective of "almost orthogonal" property of high-dimensional random vectors. We show that this intriguing property can be used to "memorize" random vectors by simply adding them, and we provide an efficient probabilistic solution to the set membership …

2018-02-23abs ↗pdf ↗

Racah matrices and higher jj-symbols are used in description of braiding properties of conformal blocks and in construction of knot polynomials. However, in complicated cases the logic is actually inverted: they are much better deduced from these applications than from the basic representation theory. Following the re…

2017-01-02abs ↗pdf ↗

Multi-head attention mechanism is capable of learning various representations from sequential data while paying attention to different subsequences, e.g., word-pieces or syllables in a spoken word. From the subsequences, it retrieves richer information than a single-head attention which only summarizes the whole sequen…

2019-10-10abs ↗pdf ↗

String structures have played an important role in algebraic topology, via elliptic genera and elliptic cohomology, in differential geometry, via the study of higher geometric structures, and in physics, via partition functions. We extend the description of String structures from connected covers of the definite-signat…

2015-04-08abs ↗pdf ↗

In this paper, we present new results on using orthogonal matching pursuit (OMP), to solve the sparse approximation problem over redundant dictionaries for complex cases (i.e., complex measurement vector, complex dictionary and complex additive white Gaussian noise (CAWGN)). A sufficient condition that OMP can recover …

2012-06-11abs ↗pdf ↗

The study extends Jacobi-orthogonality to indefinite scalar product spaces.

problem Generalizing Jacobi-orthogonality to indefinite scalar product spaces.
method Comparing principles, investigating tensor relations, proving properties.
result Every quasi-Clifford tensor is Jacobi-orthogonal; certain tensors are Jacobi-dual or Osserman.

OPT framework improves neural network generalization by learning an orthogonal transformation.

problem Improving neural network generalization.
method Orthogonal over-parameterized training (OPT) framework that minimizes hyperspherical energy.
result OPT framework provably minimizes hyperspherical energy and improves empirical generalization.

The paper studies surfaces in a bounded domain with orthogonal boundaries and proves curvature estimates.

problem Estimating the area of surfaces with orthogonal boundaries in a bounded domain.
method Weak formulation of orthogonality for curvature varifolds, classification of vanishing curvature varifolds.
result Existence of an orthogonal 2-varifold that minimizes L2L^2 curvature in the integer rectifiable class.

Orthogonal random features approximate a Bessel kernel, offering sharper bounds than random Fourier features.

problem Approximating Gaussian kernel efficiently for large datasets.
method Use of Haar orthogonal matrices to construct orthogonal random features and analyze their bias and variance.
result Orthogonal random features approximate a Bessel kernel, not the Gaussian kernel, with sharper bounds.