Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

2795588361,115 · Jun 202019922001200920172026
48 results for Data Transposition

Let Ng,nN_{g,n} be an nn--punctured non--orientable surface of genus gg with one boundary component. For g2g\geq 2 one of the generators of the mapping class group of Ng,nN_{g,n} is a crosscap transposition. We give explicit formulae for the action of crosscap transpositions and their inverses on the set of multicurves i…

2019-09-26abs ↗pdf ↗

Let NgN_{g} denote a closed nonorientable surface of genus gg. For g2g \geq 2 the mapping class group M(Ng)\mathcal{M}(N_{g}) is generated by Dehn twists and one crosscap slide (YY-homeomorphism) or by Dehn twists and a crosscap transposition. Margalit and Schleimer observed that Dehn twists have nontrivial roots. We gi…

2016-01-22abs ↗pdf ↗

New singularities and fibrations in non-orientable 4-manifolds.

problem Understanding singularities and fibrations in non-orientable 4-manifolds.
method Introducing MM-singularities and MM-fibrations, studying their handle decompositions and orientation double coverings.
result Relations among crosscap transpositions give rise to MM-fibrations on non-orientable 4-manifolds.

Classifies homomorphisms from braid groups to mapping class groups of nonorientable surfaces.

problem Classifying homomorphisms from braid groups to mapping class groups of nonorientable surfaces.
method Classifies homomorphisms based on the properties of Dehn twists and crosscap transpositions.
result Every homomorphism is either cyclic or maps generators to distinct Dehn twists or crosscap transpositions.

The study explores nonorientable 3-manifolds using open books and their monodromies.

problem Investigating open books for nonorientable 3-manifolds.
method Analyzing monodromies of open books for specific nonorientable 3-manifolds.
result Infinitely many nonisotopic genus two open books for P2imesS1P^2 imes S^1 and S2imes~S1S^2 \widetilde{ imes} S^1.

We show that the Lawrence--Krammer representation is unitary. We explicitly present the non-singular matrix representing the sesquilinear pairing invariant under the action. We show that reversing the orientation of a braid is equivalent to the transposition of its Lawrence--Krammer matrix followed by a certain conjuga…

2002-02-21abs ↗pdf ↗

We construct natural selfmaps of compact cohomgeneity one manifolds with finite Weyl group and compute their degrees and Lefschetz numbers. On manifolds with simple cohomology rings this yields in certain cases relations between the order of the Weyl group and the Euler characteristic of a principal orbit. We apply our…

2007-10-19abs ↗pdf ↗

A meander of order n is a simple closed curve in the plane which intersects a horizontal line transversely at 2n points. (Meanders which differ by an isotopy of the line and plane are considered equivalent.) Let Gamma_n be the Cayley graph of the symmetric group S_n as generated by all (n choose 2) transpositions. Let …

2006-06-08abs ↗pdf ↗

Study explores K-means clustering of variables and its relation to PCA.

problem Exploring the relationship between K-means clustering of variables and PCA.
method Apply PCA to original data and K-means to transposed data, quantify variable contributions to principal components.
result Identifies how variable clusters contribute to principal components identified by PCA.

Learning, taking into account full distribution of the data, referred to as generative, is not feasible with deep neural networks (DNNs) because they model only the conditional distribution of the outputs given the inputs. Current solutions are either based on joint probability models facing difficult estimation proble…

2017-09-25abs ↗pdf ↗

The classical Hurwitz numbers of degree n together with the Hurwitz numbers of the seamed surfaces of degree n give rise to the Klein topological field theory. We extend this construction to the Hurwitz numbers of all degrees at once. The corresponding Cardy-Frobenius algebra is induced by arbitrary Young diagrams and …

2012-12-10abs ↗pdf ↗

We investigate triangulations of the two-dimensional sphere and torus with the faces properly colored white and black. We focus on matchings between white triangles and incident vertices. On the torus our objects are perfect pairings, whereas on the sphere this is only true after removing one triangle and its vertices.…

2018-08-18abs ↗pdf ↗

This paper gives a new interpretation of the virtual braid group in terms of a strict monoidal category SC that is freely generated by one object and three morphisms, two of the morphisms corresponding to basic pure virtual braids and one morphism corresponding to a transposition in the symmetric group. The key to this…

2011-03-16abs ↗pdf ↗

Kernel-based clustering algorithm can identify and capture the non-linear structure in datasets, and thereby it can achieve better performance than linear clustering. However, computing and storing the entire kernel matrix occupy so large memory that it is difficult for kernel-based clustering to deal with large-scale …

2020-02-07abs ↗pdf ↗

A few years ago Kramer and Laubenbacher introduced a discrete notion of homotopy for simplicial complexes. In this paper, we compute the discrete fundamental group of the order complex of the Boolean lattice. As it turns out, it is equivalent to computing the discrete homotopy group of the 1-skeleton of the permutahedr…

2007-11-06abs ↗pdf ↗

The paper constructs new algebraic structures from Lie algebras and ternary Nambu-Lie algebras, leading to Yang-Baxter operators.

problem Constructing new algebraic structures from Lie algebras and ternary Nambu-Lie algebras.
method Using compositions of binary Lie algebras, 3-Lie algebras, and ternary Nambu-Lie algebras, the paper constructs ternary self-distributive objects and Yang-Baxter operators.
result The constructed Yang-Baxter operators are not gauge equivalent to the transposition operator and can be deformed to new solutions.

The study explores congruence subgroups of braid groups and their quotients.

problem Understanding the structure of congruence subgroups of braid groups.
method Utilizing the integral Burau representation and results from integral matrices, the study examines quotients of these subgroups.
result Findings of quotients that are not isomorphic to symmetric groups.

Paper presents a method to summarize HMC samples for neural networks, providing meaningful uncertainty estimates.

problem Lack of interpretable summary statistics for HMC samples in neural networks due to permutation symmetry.
method Introducing a transpositions metric to quantify permutations and using rebasin method to summarize HMC samples.
result Compact representation of HMC samples provides meaningful uncertainty estimates for each weight in a neural network.

Big data sets must be carefully partitioned into statistically similar data subsets that can be used as representative samples for big data analysis tasks. In this paper, we propose the random sample partition (RSP) data model to represent a big data set as a set of non-overlapping data subsets, called RSP data blocks,…

2017-12-12abs ↗pdf ↗

Prevents sensitive data generation in diffusion models using labeled and unlabeled data.

problem Generating sensitive data in diffusion models using unlabeled data.
method Positive-Unlabeled Diffusion Models, approximating ELBO with labeled and unlabeled data.
result Prevents the generation of sensitive data without compromising image quality.

Study reveals Data Shapley's inconsistent performance in data selection tasks.

problem Inconsistency of Data Shapley's performance in data selection across different settings.
method Hypothesis testing framework and identification of utility functions.
result Data Shapley's performance is no better than random selection without specific constraints.

PRRO generates synthetic tabular data that improves SL performance and class distribution.

problem Low SL utility of synthetic data due to class imbalance and overlooked data relationships.
method Data pruning and column reordering to optimize SL utility.
result Synthetic data generated with PRRO enhances predictive performance and class distribution.

Defines data science as a natural ecosystem with challenges and missions.

problem Challenges and missions in data science due to 5D complexities and data life cycle phases.
method Systemic and data-centric view of data science as a fusion of data universe and its challenges, formalizing a general-purpose architecture.
result Essential data science as a natural ecosystem integrating specific disciplines and high-impact applications.

Differential privacy allows quantifying privacy loss resulting from accessing sensitive personal data. Repeated accesses to underlying data incur increasing loss. Releasing data as privacy-preserving synthetic data would avoid this limitation, but would leave open the problem of designing what kind of synthetic data. W…

2019-12-10abs ↗pdf ↗

Paper creates fair synthetic data ensuring equal predictions across sensitive attributes.

problem Ensuring fair predictions across sensitive attributes in synthetic data.
method Equalizing target probability distributions across sensitive attributes in synthetic data generation.
result Synthetic data provides strong fair predictions, equal across all thresholds.

Efficient synthetic data generation improves model performance on tabular data.

problem Improving model robustness and performance with scarce or low-quality data.
method Hardness characterization to identify high-value training points, generating synthetic data only from these points.
result Synthetic data generated from hardest points outperforms non-targeted methods on tabular datasets.

For most problems in science and engineering we can obtain data sets that describe the observed system from various perspectives and record the behavior of its individual components. Heterogeneous data sets can be collectively mined by data fusion. Fusion can focus on a specific target relation and exploit directly ass…

2013-07-02abs ↗pdf ↗

DAERNN models censored data using neural networks with data augmentation.

problem Handling censored data in expectile regression.
method Data augmentation based Expectile Regression Neural Networks (ERNNs).
result DAERNN outperforms existing censored ERNNs methods and achieves comparable predictive performance to fully observed data.

Data preprocessing techniques are devoted to correct or alleviate errors in data. Discretization and feature selection are two of the most extended data preprocessing techniques. Although we can find many proposals for static Big Data preprocessing, there is little research devoted to the continuous Big Data problem. A…

2018-10-14abs ↗pdf ↗

Data stream classification methods demonstrate promising performance on a single data stream by exploring the cohesion in the data stream. However, multiple data streams that involve several correlated data streams are common in many practical scenarios, which can be viewed as multi-task data streams. Instead of handli…

2019-08-15abs ↗pdf ↗

This paper quantifies uncertainty in Data Shapley using statistical inference.

problem Uncertainty in data valuation due to dynamic data distribution.
method Established relationship with U-statistics and quantified uncertainty using statistical inference.
result Confidence intervals for Data Shapley estimations are provided.

DCoM uses deep neural networks to detect semantic data types from raw column values.

problem Detecting semantic data types from dirty and unseen data.
method DCoM employs multi-input NLP-based deep neural networks trained on 686,765 data columns.
result DCoM outperforms existing methods significantly on 78 different semantic data types.

Data mining is about obtaining new knowledge from existing datasets. However, the data in the existing datasets can be scattered, noisy, and even incomplete. Although lots of effort is spent on developing or fine-tuning data mining models to make them more robust to the noise of the input data, their qualities still st…

2019-06-20abs ↗pdf ↗