Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

3571106141 · Jun 202019922001200920182026
48 results for Cluster Categories

Survey on decorated marked surfaces for Calabi-Yau categories.

problem Understanding Calabi-Yau categories and related structures.
method Introducing decorations on marked surfaces to study various categories.
result Exploration of Calabi-Yau-2 and 3 categories, braid groups, quadratic differentials, and stability conditions.

Unified clustering model handles both pairwise and cardinality constraints for better performance.

problem Clustering with specific constraints (pairwise and cardinality) to improve clustering quality.
method Unified integer programming formulation, binary and quadratic constraints, reformulated as continuous constraints, solved using ADMM.
result Unified model outperforms single category constraints and achieves better clustering performance.

We study the cluster categories arising from marked surfaces (with punctures and non-empty boundaries). By constructing skewed-gentle algebras, we show that there is a bijection between tagged curves and string objects. Applications include interpreting dimensions of Ext1\operatorname{Ext}^1 as intersection numbers of ta…

2013-10-31abs ↗pdf ↗

This paper proposes a method to reduce complexity in GLMs with categorical predictors.

problem Wasteful, hard-to-interpret, and prone to overfitting of traditional one-hot encoding for high-cardinality categorical predictors.
method Clustering categories of categorical predictors through a numerical method that preserves or improves accuracy while reducing the number of coefficients.
result Clustering categories of categorical predictors reduces complexity substantially without harming accuracy.

This work draws inspiration from three important sources of research on dissimilarity-based clustering and intertwines those three threads into a consistent principled functorial theory of clustering. Those three are the overlapping clustering of Jardine and Sibson, the functorial approach of Carlsson and Mémoli to par…

2016-09-08abs ↗pdf ↗

Constructs exotic Lagrangian tori in Grassmannians using cluster algebra.

problem Finding non-displaceable and non-isotopic Lagrangian tori in Grassmannians.
method Iterative construction based on cluster algebra structure of a mirror Landau-Ginzburg model.
result Examples of exotic Lagrangian tori that support nonzero objects in different summands of the Fukaya category.

An unsupervised method clusters patient incident reports for content analysis.

problem Lack of methods to extract interpretable content from electronic healthcare records.
method Combines text-embedding with paragraph vectors and graph-theoretical multiscale community detection.
result Extracts high-intrinsic-consistency groups of patient incident reports.

We give a geometric model for a tube category in terms of homotopy classes of oriented arcs in an annulus with marked points on its boundary. In particular, we interpret the dimensions of extension groups of degree 1 between indecomposable objects in terms of negative geometric intersection numbers between correspondin…

2010-11-02abs ↗pdf ↗

We formalize the arithmetic topology, i.e. a relationship between knots and primes. Namely, using the notion of a cluster C*-algebra we construct a functor from the category of 3-dimensional manifolds M to a category of algebraic number fields K, such that the prime ideals (ideals, resp.) in the ring of integers of K c…

2017-06-20abs ↗pdf ↗

HCRL learns hierarchical embeddings from deep embeddings of hierarchy components.

problem Flat clustering limits cohesive instance relations in hierarchical data.
method Simultaneously optimizes representation learning and hierarchical clustering in the embedding space.
result HCRL achieves best hierarchical clustering and data reconstruction.

This paper explains spectral clustering and its equivalence to PCA, breaking it into fully connected and multi-connected cases.

problem Understanding the mathematics behind spectral clustering and its equivalence to PCA.
method Dividing spectral clustering into two categories based on graph connectivity and proving the equivalence to PCA.
result Spectral clustering and PCA are equivalent, with specific proofs for fully connected and multi-connected graphs.

The determination of cluster centers generally depends on the scale that we use to analyze the data to be clustered. Inappropriate scale usually leads to unreasonable cluster centers and thus unreasonable results. In this study, we first consider the similarity of elements in the data as the connectivity of nodes in an…

2016-10-19abs ↗pdf ↗

Proposes CRG_IMSC for better clustering of multi-view data.

problem Lack of effective connectivity in clustering results.
method Directly obtains clustering result with nonnegative constraint; constructs connectivity matrix based on spectral clustering result; uses multiplicative update algorithm.
result Improves clustering performance on benchmark datasets.

One basic requirement of many studies is the necessity of classifying data. Clustering is a proposed method for summarizing networks. Clustering methods can be divided into two categories named model-based approaches and algorithmic approaches. Since the most of clustering methods depend on their input parameters, it i…

2013-02-16abs ↗pdf ↗

The study finds that the export shares of machinery and food/crude materials are significantly correlated with GDP.

problem Understanding the relationship between export shares and GDP across different commodity sectors.
method Analysis of GDP and international trade data using the SITC classification from 1962 to 2000.
result The export shares of machinery and food/crude materials are significantly correlated with GDP, following a power-law relationship.

Adversarial MoE learns category-specific models for product search.

problem Variations in product features and importance across categories.
method Mixture of Experts with adversarial regularization and soft gating constraints.
result Improved clustering of gate output vectors and shared experts among similar categories.

Nowadays, financial data analysis is becoming increasingly important in the business market. As companies collect more and more data from daily operations, they expect to extract useful knowledge from existing collected data to help make reasonable decisions for new customer requests, e.g. user credit category, confide…

2016-09-04abs ↗pdf ↗

We give a geometric realization, the tagged rotation, of the AR-translation on the generalized cluster category associated to a surface S\mathbf{S} with marked points and non-empty boundary, which generalizes Brüstle-Zhang's result for the puncture free case. As an application, we show that the intersection of the shi…

2012-11-30abs ↗pdf ↗

Solves complex clustering and rotation synchronization problem.

problem Challenges in classifying and synchronizing rotated objects into multiple categories.
method Semidefinite programming relaxations to solve the joint problem of community detection and synchronization.
result Exact recovery of community detection and synchronization when extending stochastic block model.

In this study, we establish a network structure of the Korean stock market, one of the emerging markets, with its minimum spanning tree through the correlation matrix. Base on this analysis, it is found that the Korean stock market doesn't form the clusters of the business sectors or of the industry categories. When th…

2005-04-01abs ↗pdf ↗

The contraction inequality for Rademacher averages is extended to Lipschitz functions with vector-valued domains, and it is also shown that in the bounding expression the Rademacher variables can be replaced by arbitrary iid symmetric and sub-gaussian variables. Example applications are given for multi-category learnin…

2016-05-01abs ↗pdf ↗

In many applications that involve processing high-dimensional data, it is important to identify a small set of entities that account for a significant fraction of detections. Rather than formalize this as a clustering problem, in which all detections must be grouped into hard or soft categories, we formalize it as an i…

2018-05-08abs ↗pdf ↗

Develops a measure for checking cluster quality using multinomial distribution.

problem Determining the true number of clusters in a sample.
method Applying multinomial distribution to distances of data members from their cluster representatives.
result Demonstrates the capability to identify if a sample has inherent clusters.

New clustering method for uncertain data using Wasserstein barycenters.

problem Clustering uncertain and structured data with observational/experimental error.
method Wasserstein barycenters and geodesic criterion for optimal clustering.
result Effective clustering of complex data in astronomy, biology, and remote sensing.

New method tackles instance segmentation on 3D point clouds with improved metrics.

problem Evaluation metrics are affected by small regions containing few instances.
method Proposes a new method with O(Np) space complexity that learns embeddings for clusters of instances.
result Achieves state-of-the-art performance using both existing and proposed metrics.