Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

295988117 · Jun 202019922001200920172026
48 results for topology-based clustering

CCP clusters correlated features and projects them to 1D for efficient dimensionality reduction.

problem Efficiency in handling large datasets with high intrinsic dimensions.
method CCP partitions features into correlated clusters and projects them to 1D based on sample correlations.
result CCP achieves efficient dimensionality reduction without matrix diagonalization.

We propose a new graph kernel for graph classification and comparison using Ollivier Ricci curvature. The Ricci curvature of an edge in a graph describes the connectivity in the local neighborhood. An edge in a densely connected neighborhood has positive curvature and an edge serving as a local bridge has negative curv…

2019-07-15abs ↗pdf ↗

This work concludes a series of four papers on the foundational theory of orbifolds and stacks. We apply the abstract theory, developed in its predecessors, to orbifolds derived from manifolds. Specifically, we show how the very concrete topological base spaces associated to such orbifolds can be described and manipula…

1995-03-14abs ↗pdf ↗

Paper introduces stable vectorization for multiparameter PH using signed barcodes.

problem Lack of stable vectorization methods for multiparameter persistent homology.
method Signed barcodes as measures for stable vectorization of MPH.
result Stable feature vectors from signed barcodes improve performance in data science.

Skein modules are the main objects of an algebraic topology based on knots (or position). In the same spirit as Leibniz we would call our approach "algebra situs." When looking at the panorama of skein modules we see, past the rolling hills of homologies and homotopies, distant mountains - the Kauffman bracket skein mo…

1998-09-21abs ↗pdf ↗

This paper is base on talks which I gave in May, 2010 at Workshop in Trieste (ICTP). In the first part we present an introduction to knots and knot theory from an historical perspective, starting from Summerian knots and ending on Fox 3-coloring. We show also a relation between 3-colorings and the Jones polynomial. In …

2011-05-11abs ↗pdf ↗

Proposes SimPool for graph pooling using structural similarity features.

problem Challenges in graph pooling due to lack of spatial locality.
method Integrates structural similarity features with a revised pooling layer to propose SimPool.
result SimPool produces node cluster assignments resembling CNN's locality preserving pooling.

We construct the Weil functor TAT^A corresponding to a general Weil algebra A=KNA = K \oplus N: this is a functor from the category of manifolds over a general topological base field or ring KK (of arbitrary characteristic) to the category of manifolds over AA. This result simultaneously generalizes results known for o…

2011-11-10abs ↗pdf ↗

The objective of the present paper (the second in a series of four) is to give a theory of multivector and extensor fields on a smooth manifold M of arbitrary topology based on the powerful geometric algebra of multivectors and extensors. Our approach does not suffer the problems of earlier attempts which are restricte…

2005-01-31abs ↗pdf ↗

In his recent investigation of a super Teichmüller space, Sachse (2007), based on work of Molotkov (1984), has proposed a theory of Banach supermanifolds using the `functor of points' approach of Bernstein and Schwarz. We prove that the the category of Berezin-Kostant-Leites supermanifolds is equivalent to the category…

2009-10-28abs ↗pdf ↗

Knot Theory is currently a very broad field. Even a long survey can only cover a narrow area. Here we concentrate on the path from Goeritz matrices to quasi-alternating links. On the way, we often stray from the main road and tell related stories, especially if they allow as to place the main topic in a historical cont…

2009-09-06abs ↗pdf ↗

This paper proposes a new method for automatically selecting the optimal kernel bandwidth in density estimation.

problem The challenge of selecting the optimal kernel bandwidth in unsupervised density estimation.
method The approach uses a topology-based loss function for automated bandwidth selection.
result Demonstrates the potential of the topology-based approach across different dimensions.

We describe in this chapter (Chapter IX) the idea of building an algebraic topology based on knots (or more generally on the position of embedded objects). That is, our basic building blocks are considered up to ambient isotopy (not homotopy or homology). For example, one should start from knots in 3-manifolds, surface…

2006-02-13abs ↗pdf ↗

In colored graphs, node classes are often associated with either their neighbors class or with information not incorporated in the graph associated with each node. We here propose that node classes are also associated with topological features of the nodes. We use this association to improve Graph machine learning in g…

2019-10-26abs ↗pdf ↗

The survey we are presenting is over 22 years old but it has still some ideas which where never published (except in Polish). This survey is the base of the third Chapter of my book: KNOTS: From combinatorics of knot diagrams to combinatorial topology based on knots, which is still in preparation (but compare http://ar…

2008-10-23abs ↗pdf ↗

Unreduced PDs can perform similarly to reduced PDs in machine learning tasks.

problem Ignoring much of the information in persistence diagrams in machine learning pipelines.
method Developed methods to generate topological feature vectors from unreduced boundary matrices.
result Unreduced PDs can perform on par with, and sometimes outperform, fully-reduced PDs in machine learning tasks.

Mode clustering is a nonparametric method for clustering that defines clusters using the basins of attraction of a density estimator's modes. We provide several enhancements to mode clustering: (i) a soft variant of cluster assignment, (ii) a measure of connectivity between clusters, (iii) a technique for choosing the …

2014-06-06abs ↗pdf ↗

Clustering is an essential data mining tool that aims to discover inherent cluster structure in data. For most applications, applying clustering is only appropriate when cluster structure is present. As such, the study of clusterability, which evaluates whether data possesses such structure, is an integral part of clus…

2018-08-24abs ↗pdf ↗

Clustering ensemble, or consensus clustering, has emerged as a powerful tool for improving both the robustness and the stability of results from individual clustering methods. Weighted clustering ensemble arises naturally from clustering ensemble. One of the arguments for weighted clustering ensemble is that elements (…

2019-10-06abs ↗pdf ↗

In many practical applications of clustering, the objects to be clustered evolve over time, and a clustering result is desired at each time step. In such applications, evolutionary clustering typically outperforms traditional static clustering by producing clustering results that reflect long-term trends while being ro…

2011-04-11abs ↗pdf ↗

Study examines how cluster number affects short-text clustering, introducing a stability metric.

problem Challenges in finding meaningful clusters in short-text data.
method Introduces a stability metric to determine cluster robustness and visualizes cluster subdivisions.
result Choosing a cluster number involves balancing informativeness and complexity, not seeking a single 'optimal' solution.

Convex clustering, a convex relaxation of k-means clustering and hierarchical clustering, has drawn recent attentions since it nicely addresses the instability issue of traditional nonconvex clustering methods. Although its computational and statistical properties have been recently studied, the performance of convex c…

2016-01-18abs ↗pdf ↗

Clustering is a central approach for unsupervised learning. After clustering is applied, the most fundamental analysis is to quantitatively compare clusterings. Such comparisons are crucial for the evaluation of clustering methods as well as other tasks such as consensus clustering. It is often argued that, in order to…

2017-01-23abs ↗pdf ↗