Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

232463695926 · Jun 202019922001200920172026
48 results for set similarity

Lipschitz equivalence of self-similar sets is an important area in the study of fractal geometry. It is known that two dust-like self-similar sets with the same contraction ratios are always Lipschitz equivalent. However, when self-similar sets have touching structures the problem of Lipschitz equivalence becomes much …

2012-07-28abs ↗pdf ↗

Skeleton is a new notion designed for constructing space-filling curves of self-similar sets. It is shown in [Dai, Rao and Zhang, Space-filling curves of self-similar sets (II): Edge-to-trail substitution rule,https://doi.org/10.1088/1361-6544/ab1275] that for a connected self-similar set, space-filling curves can be c…

2018-04-27abs ↗pdf ↗

In this paper, we introduce the notion of asymptotic self-similar sets on general doubling metric spaces by extending the notion of self-similar sets, and determine their Hausdorff dimensions, which gives an extension of Balogh and Rohner 's result. This is carried out by introducing the notions of almost similarity ma…

2017-10-02abs ↗pdf ↗

In [9] Kaimanovich introduced the concept of augmented tree on the symbolic space of a self-similar set. It is hyperbolic in the sense of Gromov, and it was shown in [13] that under the open set condition, a self-similar set can be identified with the hyperbolic boundary of the tree. In the paper, we investigate in det…

2012-05-16abs ↗pdf ↗

Paper explains contrastive learning using cosine similarity and proposes mitigations for batch size effects.

problem Understanding and improving contrastive learning through batch size effects.
method Unified framework of cosine similarity, theoretical insights, and auxiliary loss.
result Performance improvement in small-batch settings through proposed auxiliary loss.

New stability measures for similar features improve feature selection accuracy.

problem Existing stability measures fail to distinguish similar features in highly correlated datasets.
method Introduce new adjusted stability measures that consider feature similarities.
result One new stability measure considers highly similar features as interchangeable.

Bayesian methods detect significant IIA violations in similarity choice data.

problem Detecting IIA violations in similarity choice data complicates classical models.
method Proposed two statistical methods: classical goodness-of-fit test and Bayesian PPC.
result Significant IIA violations confirmed in both datasets, driven by context effects.

Given only information in the form of similarity triplets "Object A is more similar to object B than to object C" about a data set, we propose two ways of defining a kernel function on the data set. While previous approaches construct a low-dimensional Euclidean embedding of the data set that reflects the given similar…

2016-07-28abs ↗pdf ↗

The two main theorems of this paper provide a characterization of hyperbolic affine iterated function systems defined on Rm. Atsushi Kameyama (Distances on Topological Self-Similar Sets, Proceedings of Symposia in Pure Mathematics, Volume 72.1, 2004) asked the following fundamental question: given a topological self-si…

2009-08-10abs ↗pdf ↗

We introduce GSimCNN (Graph Similarity Computation via Convolutional Neural Networks) for predicting the similarity score between two graphs. As the core operation of graph similarity search, pairwise graph similarity computation is a challenging problem due to the NP-hard nature of computing many graph distance/simila…

2018-10-23abs ↗pdf ↗

Proposes a new stability measure for model fitting on similar feature data sets.

problem Model fitting on data sets with similar features is challenging.
method Tuning hyperparameters in a multi-criteria fashion with predictive accuracy and feature selection stability.
result Our approach achieves similar or better predictive performance than single-criteria and stability selection approaches.

The study explores homeomorphism groups of self-similar 2-manifolds, including the 2-sphere and Cantor set.

problem Understanding the structure and properties of homeomorphism groups of self-similar 2-manifolds.
method Survey of recent results, exposition of classical results, treatment of stable sets, and proof of new theorems.
result Characterization of homeomorphisms of perfectly self-similar 2-manifolds and extensions of existing results.

In a context of document co-clustering, we define a new similarity measure which iteratively computes similarity while combining fuzzy sets in a three-partite graph. The fuzzy triadic similarity (FT-Sim) model can deal with uncertainty offers by the fuzzy sets. Moreover, with the development of the Web and the high ava…

2013-12-21abs ↗pdf ↗

Study properties of self-similar continua with finite intersection property.

problem Characterize self-similar continua with finite intersection property.
method Prove intersection graph criterion, finite order theorem, and parameter matching theorem.
result All Jordan arcs starting from a intersection point in such continuum on a plane should have the same slope parameter at that point.

Fractal Lipschitz-Killing curvature measures C^f_k(F,.), k = 0, ..., d, are determined for a large class of self-similar sets F in R^d. They arise as weak limits of the appropriately rescaled classical Lipschitz-Killing curvature measures C_k(F_r,.) from geometric measure theory of parallel sets F_r for small distances…

2010-07-05abs ↗pdf ↗

Study on self-similar sets on Riemannian manifolds with new separation conditions.

problem Analyzing self-similar sets on Riemannian manifolds with new separation conditions.
method Formulated weak separation and finite type conditions for conformal iterated function systems on Riemannian manifolds.
result Obtained formulas for Hausdorff dimensions of self-similar and graph self-similar sets.

Excessive reuse of test data has become commonplace in today's machine learning workflows. Popular benchmarks, competitions, industrial scale tuning, among other applications, all involve test data reuse beyond guidance by statistical confidence bounds. Nonetheless, recent replication studies give evidence that popular…

2019-05-29abs ↗pdf ↗

The problem of hierarchical clustering items from pairwise similarities is found across various scientific disciplines, from biology to networking. Often, applications of clustering techniques are limited by the cost of obtaining similarities between pairs of items. While prior work has been developed to reconstruct cl…

2012-07-19abs ↗pdf ↗

Study on self-similar solutions of supercritical Fujita equation, proving entropy and energy gap.

problem Characterization and stability of solutions to supercritical Fujita equation.
method Introduction of FF-functional, FF-stability, and entropy; use of mean curvature flows.
result Constant solution has lowest entropy among bounded positive self-similar solutions.

In this article we verify an orbifold version of a conjecture of Nimershiem from 1998. Namely, for every flat nn-manifold MM, we show that the set of similarity classes of flat metrics on MM which occur as a cusp cross-section of a hyperbolic (n+1)(n+1)-orbifold is dense in the space of similarity classes of flat metri…

2006-06-20abs ↗pdf ↗

Study uses trajectory embedding to measure place function similarity at fine spatial granularity.

problem Measuring place function similarity at fine spatial granularity.
method Trajectory embedding to reduce dimensions and measure similarity of place functions.
result Embedding similarity can be a metric proxy for place functions at fine spatial granularity.

Algorithm aggregates rewards from multiple players to learn related tasks in online bandit learning.

problem Learning related but slightly different tasks in an online setting with heterogeneous feedback.
method RobustAgg(ε)(ε) algorithm that aggregates rewards from different players.
result Achieves instance-dependent regret guarantees and nearly matching lower bounds.

There are plenty of problems where the data available is scarce and expensive. We propose a generator of semi-artificial data with similar properties to the original data which enables development and testing of different data mining algorithms and optimization of their parameters. The generated data allow a large scal…

2014-03-28abs ↗pdf ↗

Multi-label classification is an important learning problem with many applications. In this work, we propose a principled similarity-based approach for multi-label learning called SML. We also introduce a similarity-based approach for predicting the label set size. The experimental results demonstrate the effectiveness…

2017-10-27abs ↗pdf ↗

For many analytical problems the challenge is to handle huge amounts of available data. However, there are data science application areas where collecting information is difficult and costly, e.g., in the study of geological phenomena, rare diseases, faults in complex systems, insurance frauds, etc. In many such cases,…

2019-09-12abs ↗pdf ↗

CLS measures dataset similarity through decision rule performance.

problem Measuring dataset similarity in machine learning, especially for transfer learning and domain adaptation.
method Cross-Learning Score (CLS) measures similarity through bidirectional generalization performance of decision rules, linking to cosine similarity under canonical linear models.
result CLS effectively measures dataset similarity and transferability, validated on synthetic and real-world datasets.