Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

3977116154 · Jun 202019922001200920172026
48 results for Matrix transpose

EDAs with matrix transpose improve Bayesian structure learning performance.

problem Improving Bayesian structure learning performance.
method Introducing a matrix transpose mutation operator for EDAs in Bayesian structure learning.
result EDAs with transpose mutation give markedly better performance than conventional EDAs.

Transposable data represents interactions among two sets of entities, and are typically represented as a matrix containing the known interaction values. Additional side information may consist of feature vectors specific to entities corresponding to the rows and/or columns of such a matrix. Further information may also…

2014-04-27abs ↗pdf ↗

Graph embedding learns low-dimensional representations for nodes in a graph and effectively preserves the graph structure. Recently, a significant amount of progress has been made toward this emerging research area. However, there are several fundamental problems that remain open. First, existing methods fail to preser…

2019-05-16abs ↗pdf ↗

Given a knot and an SL(n,C) representation of its group that is conjugate to its dual, the representation that replaces each matrix with its inverse-transpose, the associated twisted Reidemeister torsion is reciprocal. An example is given of a knot group and SL(3,Z) representation that is not conjugate to its dual for …

2009-05-15abs ↗pdf ↗

Introduces qq-transpose for qq-deformed modular group matrices.

problem Understanding qq-deformed rational numbers and their properties.
method Introduces qq-transpose and applies it to refine qq-deformed modular group actions.
result New proof and refinement of Leclere and Morier-Genoud's trace palindromicity theorem.

Combines gradient-based and competitive learning for unsupervised feature extraction.

problem Handling input data without supervision and replicating input manifold topology.
method Integrates gradient-based and competitive learning approaches to learn topological structures.
result The dual competitive layer outperforms the vanilla layer in high-dimensional datasets.

We introduce a guide to help deep learning practitioners understand and manipulate convolutional neural network architectures. The guide clarifies the relationship between various properties (input shape, kernel shape, zero padding, strides and output shape) of convolutional, pooling and transposed convolutional layers…

2016-03-23abs ↗pdf ↗

Kernel clustering algorithm improved for large datasets using incomplete Cholesky factorization.

problem Large memory usage in kernel-based clustering for large-scale datasets.
method Approximate the kernel matrix using incomplete Cholesky factorization and apply linear kk-means clustering.
result The proposed method achieves similar performance to kernel kk-means clustering but handles large-scale datasets efficiently.

Study explores K-means clustering of variables and its relation to PCA.

problem Exploring the relationship between K-means clustering of variables and PCA.
method Apply PCA to original data and K-means to transposed data, quantify variable contributions to principal components.
result Identifies how variable clusters contribute to principal components identified by PCA.

The Berglund-Hübsch rule connects Calabi-Yau orbifolds to Sasakian manifolds.

problem Connecting Calabi-Yau orbifolds to Sasakian manifolds.
method Applying the Berglund-Hübsch transpose rule to associate Sasaki manifolds.
result Four seven-dimensional Sasakian manifolds of positive Ricci curvature are associated with a K3 orbifold.

Variables in many massive high-dimensional data sets are structured, arising for example from measurements on a regular grid as in imaging and time series or from spatial-temporal measurements as in climate studies. Classical multivariate techniques ignore these structural relationships often resulting in poor performa…

2011-02-15abs ↗pdf ↗

This paper studies iteration convergence of Kronecker graphical lasso (KGLasso) algorithms for estimating the covariance of an i.i.d. Gaussian random sample under a sparse Kronecker-product covariance model and MSE convergence rates. The KGlasso model, originally called the transposable regularized covariance model by …

2012-04-03abs ↗pdf ↗

New algorithm estimates robust Gaussian covariance in nearly matrix multiplication time.

problem Estimating robust covariance from corrupted Gaussian samples.
method Developed a novel algorithm achieving near-optimal error in Mahalanobis norm with runtime nearly matrix multiplication time.
result Achieved the same statistical guarantees as previous work but with no dependence on ε in runtime.

Language models fail to process hallucinated responses, and this study diagnoses the failure.

problem Language models fail to process hallucinated responses, leading to over-concentration or diffuse attention.
method The study uses forced scoring of benchmark-labeled responses to compute attention shapes and analyze the symmetric component of the degree-normalized attention operator.
result The study proves that every transpose-invariant spectral diagnostic of the attention operator is orientation-blind and bounds the sensitivity of any Lipschitz diagnostic by the asymmetry coefficient \(G\).

We study the Low Rank Phase Retrieval (LRPR) problem defined as follows: recover an n×qn \times q matrix XX^* of rank rr from a different and independent set of mm phaseless (magnitude-only) linear projections of each of its columns. To be precise, we need to recover XX^* from yk:=Akxk,k=1,2,,qy_k := |A_k{}' x^*_k|, k=1,2,\dots, q

2019-02-13abs ↗pdf ↗

Tensor programs prove neural network limits for any architecture.

problem Understanding the limits of neural networks of any architecture.
method Prove convergence of neural network's Tangent Kernel (NTK) to a deterministic limit as network widths increase.
result Identify conditions for correct NTK limit calculation based on gradient independence assumption.

This work bridges competitive learning with gradient-based learning for faster feature extraction.

problem Lack of powerful feature extractors in competitive learning methods.
method Introduces gradient-based competitive layers for feature extraction.
result Demonstrates theoretical equivalence and faster convergence of gradient-based competitive layers.

We propose a deep factorization model for typographic analysis that disentangles content from style. Specifically, a variational inference procedure factors each training glyph into the combination of a character-specific content embedding and a latent font-specific style variable. The underlying generative model combi…

2019-10-02abs ↗pdf ↗

Most existing GANs architectures that generate images use transposed convolution or resize-convolution as their upsampling algorithm from lower to higher resolution feature maps in the generator. We argue that this kind of fixed operation is problematic for GANs to model objects that have very different visual appearan…

2018-01-25abs ↗pdf ↗

We study Nijenhuis structures on Courant algebroids in terms of the canonical Poisson bracket on their symplectic realizations. We prove that the Nijenhuis torsion of a skew-symmetric endomorphism N of a Courant algebroid is skew-symmetric if the square of N is proportional to the identity, and only in this case when t…

2011-02-07abs ↗pdf ↗

The paper extends Gray's result to quaternion-Kähler manifolds.

problem Understanding quaternion-Kähler manifolds with non-negative quaternionic sectional curvature.
method Introducing quaternionic sectional curvature, proving Wolf spaces have non-negative curvature, and using nearly Kähler twistor spaces.
result Every quaternion-Kähler manifold with non-negative quaternionic sectional curvature is a Wolf space.

A hybrid ASR system using conformer architecture improves word-error-rate and training speed.

problem Improving word-error-rate and training efficiency for hybrid ASR systems.
method Used conformer architecture, applied time downsampling, and transposed convolutions.
result Conformer-based hybrid model achieves competitive results and significantly outperforms BLSTM-based hybrid model.

In classical General Relativity, the way to exhibit the equations for the gravitational waves is based on two "tricks" allowing to transform the Einstein equations after linearizing them over the Minkowski metric. With specific notations used in the study of {\it Lie pseudogroups} of transformations of an nn-dimension…

2017-08-22abs ↗pdf ↗

Study on feature learning dynamics in infinite-depth neural networks, focusing on ResNets.

problem Understanding how features evolve during training in deep neural networks, especially in the large-depth limit.
method Conditional Gaussian representations and SDE system with decoupled backward weights.
result Depth-induced suppression of forward-backward coupling in infinite-depth networks, leading to a decoupled forward-backward SDE system.

Orthogonium offers unified, efficient layers for robust deep learning.

problem Fragmented and computationally demanding implementations of orthogonal and 1-Lipschitz layers.
method Unified, efficient PyTorch library providing orthogonal and 1-Lipschitz layers.
result Reduced overhead and standardized tools for robust experimentation.

This paper extends the evolution operator to contact mechanics, linking Lagrangian and Hamiltonian formulations.

problem Translating the evolution operator to contact mechanics for mechanical systems with dissipation.
method Using the evolution operator K to connect Lagrangian and Hamiltonian formalisms in contact mechanics.
result The evolution operator provides a geometric description of evolution equations and relates constraints.

Proposes a fixed smooth convolutional layer to reduce checkerboard artifacts in CNNs.

problem Checkerboard artifacts in CNNs during upsampling and strided convolution.
method Fixed convolutional layer with adjustable smoothness, applied to four CNNs and GANs.
result Significantly improves classification performance and image generation quality.

To date, attribute discretization is typically performed by replacing the original set of continuous features with a transposed set of discrete ones. This paper provides support for a new idea that discretized features should often be used in addition to existing features and as such, datasets should be extended, and n…

2018-02-09abs ↗pdf ↗

The concept of explainability is envisioned to satisfy society's demands for transparency on machine learning decisions. The concept is simple: like humans, algorithms should explain the rationale behind their decisions so that their fairness can be assessed. While this approach is promising in a local context (e.g. to…

2019-10-03abs ↗pdf ↗

Neural networks are commonly trained to make predictions through learning algorithms. Contrastive Hebbian learning, which is a powerful rule inspired by gradient backpropagation, is based on Hebb's rule and the contrastive divergence algorithm. It operates in two phases, the forward (or free) phase, where the data are …

2018-06-19abs ↗pdf ↗

Study local moduli of Sasaki-Einstein metrics on specific polynomial links.

problem Understanding the local moduli of Sasaki-Einstein metrics on links of invertible polynomials.
method Analyzing Sasaki-Einstein metrics on links of invertible polynomials of cycle type and Thom-Sebastiani sums.
result For polynomials of cycle type, local moduli spaces are zero-dimensional. For Thom-Sebastiani sums, dimensions are positive.

Proposes continuous convolution layers for flexible feature map resizing.

problem Fixed stride limitations in discrete convolution layers.
method Introduces Continuous Convolution (CC) layers that use learned continuous functions.
result Dynamic and consistent resizing of feature maps at any scale, non-integer and axis-dependent.