A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
We refine and generalize several interpolation inequalities bounding the Lp norm of a probability density with respect to the reference measure μ by its Sobolev norm and the Kantorovich distance to μ on a smooth weighted Riemannian manifold satisfying CD(0,∞) condition.
Improved anti-cancer drug sensitivity prediction using REFINED CNN ensemble learning.
problem Challenges in predicting anti-cancer drug sensitivity for individual cell lines.
method Using REFINED CNN, which represents high-dimensional vectors as compact 2D images with spatial correlations, and building ensembles of these models.
result Ensemble approaches significantly improve drug sensitivity prediction performance compared to single models.
This paper investigates methods for quantifying similarity between audio signals, specifically for the task of of cover song detection. We consider an information-theoretic approach, where we compute pairwise measures of predictability between time series. We compare discrete-valued approaches operating on quantised au…
Regularity properties of intrinsic objects for a large class of Stein Manifolds, namely of Monge-Ampère exhaustions and Kobayashi distance, is interpreted in terms of modular data. The results lead to a construction of an infinite dimensional family of convex domains with squared Kobayashi distance of prescribed regula…
This paper introduces a machine for sampling approximate model-X knockoffs for arbitrary and unspecified data distributions using deep generative models. The main idea is to iteratively refine a knockoff sampling mechanism until a criterion measuring the validity of the produced knockoffs is optimized; this criterion i…
Let P,Q be Heegaard surfaces of a closed orientable 3-manifold. In this paper, we introduce a method for giving an upper bound of Hempel distance of P by using the Reeb graph derived from a certain horizontal arc in the ambient space [0,1]×[0,1] of the Rubinstein-Scharlemann graphic derived from P and Q…
The concept of refinement from probability elicitation is considered for proper scoring rules. Taking directions from the axioms of probability, refinement is further clarified using a Hilbert space interpretation and reformulated into the underlying data distribution setting where connections to maximal marginal diver…
Study non-orientable link cobordisms using Floer homologies to prove inequalities.
problem Prove inequalities involving Euler characteristic and local maxima in non-orientable cobordisms.
method Use unoriented instanton and knot Floer homology to introduce unoriented versions of band unknotting number and refined cobordism distance.
result Show that the difference between unoriented refined cobordism distance of a knot from the unknot and non-orientable slice genus can be arbitrarily large.
We consider various notions of strains; quantitative measures for the deviation of a linear transformation from an isometry. The main approach, which is motivated by physical applications and follows the work of Patrizio Neff and co-workers , is to select a Riemannian metric on GLn, and use its induced geodes…
Measuring conditional independence is one of the important tasks in statistical inference and is fundamental in causal discovery, feature selection, dimensionality reduction, Bayesian network learning, and others. In this work, we explore the connection between conditional independence measures induced by distances on …
We propose three measures of mutual dependence between multiple random vectors. All the measures are zero if and only if the random vectors are mutually independent. The first measure generalizes distance covariance from pairwise dependence to mutual dependence, while the other two measures are sums of squared distance…
The detection of interesting patterns in large high-dimensional datasets is difficult because of their dimensionality and pattern complexity. Therefore, analysts require automated support for the extraction of relevant patterns. In this paper, we present FDive, a visual active learning system that helps to create visua…
We initiate the rigorous study of classification in quasi-metric spaces. These are point sets endowed with a distance function that is non-negative and also satisfies the triangle inequality, but is asymmetric. We develop and refine a learning algorithm for quasi-metrics based on sample compression and nearest neighbor…
Understanding proper distance measures between distributions is at the core of several learning tasks such as generative models, domain adaptation, clustering, etc. In this work, we focus on mixture distributions that arise naturally in several application domains where the data contains different sub-populations. For …