Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,786 papers · 148 categories

Trend · papers per month

76152228304 · Jun 202019922001200920172026
48 results for Explainable Similarity

SPINEX improves clustering with explainable neighbors, outperforming other methods.

problem Improving clustering performance and explainability.
method Leverages similarity and higher-order interactions for clustering.
result SPINEX outperforms 13 clustering algorithms across various datasets.

SPINEX improves time series forecasting with explainable neighbors.

problem Enhancing time series forecasting accuracy and interpretability.
method Leverages similarity and higher-order temporal interactions across multiple scales.
result SPINEX consistently ranks among top performers in forecasting precision.

SX-GeoTree improves spatially coherent explanations in geospatial regression trees.

problem Capturing spatial dependence and producing robust explanations in tabular prediction models.
method Integrates three objectives: impurity reduction, spatial residual control, and explanation robustness via modularity maximization on a consensus similarity network.
result Improves residual spatial evenness and doubles attribution consensus (modularity: Fujian 0.19 vs 0.09; Seattle 0.10 vs 0.05).

This paper explores explaining GBDT2NN predictions by its teacher model, improving distillation performance.

problem Explaining GBDT2NN predictions when the models have different structures.
method Empirical study on new approach to explain GBDT2NN predictions and use it as an auxiliary learning task.
result Proposed methods achieve better performance on both explanations and predictions.

We give explicit formulae for fringe lengths of the Calegari-Walker Ziggurats -- i.e. graphs of extremal rotation numbers associated to positive words in free groups. These formulae reveal (partial) integral projective self-similarity in ziggurat fringes, which are low-dimensional projections of characteristic polyhedr…

2015-03-13abs ↗pdf ↗

Local surrogate explainers vary in objectives, leading to incomparable explanations.

problem Variability in objectives among local surrogate explainers.
method Review of multiple local surrogate explainers, focusing on extracted information.
result Diverse explanations from similar methods due to differing objectives.

Proposes Coherent Gradients to explain and reduce overfitting in neural networks.

problem Why neural networks generalize well despite fitting random data.
method Hypothesis about gradient dynamics and a modification to gradient descent.
result Supports hypothesis with heuristic arguments and perturbative experiments.

The aim of this paper is to analyze the processes of polarization and agglomeration, to explain the mechanisms and causes of these phenomena in order to identify similarities and differences. As the main implication of this study should be noted that both process pretend to explain the concentration of economic activit…

2011-10-25abs ↗pdf ↗

We describe first integrals of geostrophic equations, which are similar to the enstrophy invariants of the Euler equation for an ideal incompressible fluid. We explain the geometry behind this similarity, give several equivalent definitions of the Poisson structure on the space of smooth densities on a symplectic manif…

2008-02-29abs ↗pdf ↗

TX-Ray analyzes and quantifies model knowledge transfer in NLP.

problem Insufficient methods for explaining and quantifying model knowledge transfer in NLP.
method Modified computer vision explainability principle to NLP, visualizing feature preference distributions.
result TX-Ray reveals how self-supervised models learn linguistic abstractions and improves generalization.

RELAX provides first attribution-based explanations for representations.

problem Lack of methods to explain what influences learned representations.
method RELAX, a first approach for attribution-based explanations of representations, measuring similarities in representation space.
result Significantly outperforms gradient-based baseline and models uncertainty in explanations.

This paper explains spectral clustering and its equivalence to PCA, breaking it into fully connected and multi-connected cases.

problem Understanding the mathematics behind spectral clustering and its equivalence to PCA.
method Dividing spectral clustering into two categories based on graph connectivity and proving the equivalence to PCA.
result Spectral clustering and PCA are equivalent, with specific proofs for fully connected and multi-connected graphs.

Paper proposes a method to make image model explanations robust to distortions.

problem Ensuring robustness of explanations for images under distortions.
method Embedding perceptual distances in surrogate explainers to evaluate and improve robustness.
result Surrogate explanations become more coherent and robust to distortions.

RATIO improves neural network robustness and explainability.

problem Neural networks' lack of robustness to adversarial changes and uncertainty on out-distribution samples.
method RATIO: Adversarial Training on In- and Out-distribution.
result RATIO leads to robust models with reliable confidence estimates on out-distribution samples.

The Global Vectors for word representation (GloVe), introduced by Jeffrey Pennington et al. is reported to be an efficient and effective method for learning vector representations of words. State-of-the-art performance is also provided by skip-gram with negative-sampling (SGNS) implemented in the word2vec tool. In this…

2014-11-20abs ↗pdf ↗

The paper deals with the problem of identifying the internal dependencies and similarities among a large number of random processes. Linear models are considered to describe the relations among the time series and the energy associated to the corresponding modeling error is the criterion adopted to quantify their simil…

2008-01-19abs ↗pdf ↗

Paper explains contrastive learning using cosine similarity and proposes mitigations for batch size effects.

problem Understanding and improving contrastive learning through batch size effects.
method Unified framework of cosine similarity, theoretical insights, and auxiliary loss.
result Performance improvement in small-batch settings through proposed auxiliary loss.

To study how mental object representations are related to behavior, we estimated sparse, non-negative representations of objects using human behavioral judgments on images representative of 1,854 object categories. These representations predicted a latent similarity structure between objects, which captured most of the…

2019-01-09abs ↗pdf ↗

We introduce a new method to explain Gaussian processes using Shapley values.

problem Explaining the uncertainty in Gaussian process models.
method Extending Shapley values to stochastic cooperative games for Gaussian processes.
result Our method generates explanations that are random variables and satisfy favorable axioms.

We introduce self-dual manifolds and show that they can be used to encode mirror symmetry for affine-Kähler manifolds and for elliptic curves. Their geometric properties, especially the link with special lagrangian fibrations and the existence of a transformation similar to the Fourier-Mukai functor, suggest that this …

2002-02-02abs ↗pdf ↗

The paper studies harmonic map heat flow stability and decay rates.

problem Analyzing stability and decay rates of harmonic map heat flow solutions.
method Use of homogeneous Besov space B˙p,dp(Rd)\dot{B}^{\frac{d}{p}}_{p,\infty}(\mathbb{R}^d) for small initial data and self-similar decay assumption.
result Decay rates for solutions of the harmonic map flow of the form ablau(t)L(Rd)Ct12\| abla u(t) \|_{L^\infty(\mathbb{R}^d)}\leq Ct^{-\frac12} and self-similar decay under stronger initial conditions.

The paper proposes a new algorithm to select subsets of training data for better accuracy and explainability.

problem Tackles the challenge of balancing accuracy and explainability in pattern recognition.
method Identifies multiple subsets with simple local patterns by clustering similar instances.
result The sub-setting algorithm outperformed traditional decision trees by 15% on the international stroke dataset.

To every oriented link LL, we associate a topologically defined biquandle B^L\widehat{\mathcal{B}}_{L}, which we call the topological biquandle of LL. The construction of B^L\widehat{\mathcal{B}}_{L} is similar to the topological description of the fundamental quandle given by Matveev. We find a presentation of the top…

2018-03-12abs ↗pdf ↗

A conformal map from a Riemann surface to a Euclidean space of dimension greater than or equal to three is explained by using the Clifford algebra, in a similar fashion to quaternionic holomorphic geometry of surfaces in the Euclidean three- or four-space. The Weierstrass representation, the spin transform, the Darboux…

2017-07-23abs ↗pdf ↗

Pantypes improve prototypical models by capturing diverse input distributions.

problem Prototypical models lack sufficient data representation in low density regions.
method Introducing pantypes, a sparse set of diverse objects to represent the full diversity of input distribution.
result Pantypes empower prototypical models to foster high diversity, interpretability, and fairness.

RSM provides insights into deep survival models' decision-making.

problem Ensuring trust in deep survival models' predictions for healthcare applications.
method Reverse survival model (RSM) framework that explains deep survival models' decisions.
result RSM extracts relevant features for deep survival models' predictions.

A new score function improves explainability and reliability of AI systems.

problem Designing AI systems that are explainable, robust, and trustworthy.
method Integrates conformal prediction with explainable machine learning using a novel score function.
result The method achieves improved performance on target classes and satisfies conformal guarantees.

Predictive modeling applications increasingly use data representing people's behavior, opinions, and interactions. Fine-grained behavior data often has different structure from traditional data, being very high-dimensional and sparse. Models built from these data are quite difficult to interpret, since they contain man…

2016-07-21abs ↗pdf ↗

This paper explores using SSIM for better image generation in generative models.

problem Improving perceptual quality in generated images using 2\ell_2 norm.
method Theoretical discussion and practical implementation of SSIM in generative models and autoencoders.
result SSIM can be used in generative models and autoencoders to generate better images.