Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

12.5%25.0%37.5%50.0% · Jan 199419922001200920172026
48 results for Gini importance

The aim of this paper is to introduce a risk measure that extends the Gini-type measures of risk and variability, the Extended Gini Shortfall, by taking risk aversion into consideration. Our risk measure is coherent and catches variability, an important concept for risk management. The analysis is made under the Choque…

2017-07-23abs ↗pdf ↗

The article compares predictor importance in classification problems with categorical outcomes.

problem Comparing predictor importance in classification problems with categorical response variables.
method The approach is based on the categorical Gini correlation (CGC) and tests differences in CGCs across predictor groups.
result The proposed methodology accommodates predictors of arbitrary and unequal dimensions and allows for dependence between predictor groups.

We study the problems related to the estimation of the Gini index in presence of a fat-tailed data generating process, i.e. one in the stable distribution class with finite mean but infinite variance (i.e. with tail index α(1,2)α\in(1,2)). We show that, in such a case, the Gini coefficient cannot be reliably estimated usin…

2017-07-05abs ↗pdf ↗

This paper fills in local bounds for Spearman's footrule and Gini's gamma measures of association.

problem Local bounds for bivariate copulas with respect to Spearman's footrule and Gini's gamma measures.
method Computing quasi-copulas that are not copulas for certain values of the measures.
result Presented local bounds for Spearman's footrule and Gini's gamma measures.

Identifying statistical dependence between the features and the label is a fundamental problem in supervised learning. This paper presents a framework for estimating dependence between numerical features and a categorical label using generalized Gini distance, an energy distance in reproducing kernel Hilbert spaces (RK…

2019-06-05abs ↗pdf ↗

The paper examines the stability of binary choice models using Gini index and scoring indicators.

problem Stability and discriminatory power of binary choice models.
method Derives the real Gini index and incorporates PSI and KS statistics into the model.
result The real Gini index should be less than the calculated Gini index when the population distribution changes.

Paper proves convergence of Gini index to equilibrium in Wasserstein distance.

problem Proving convergence of Gini index to equilibrium in Wasserstein distance.
method Analyzes Gini index as Lyapunov functional and proves convergence in Wasserstein distance.
result Proves convergence of Gini index to equilibrium in Wasserstein distance.

Analytical results bound the approach to oligarchy in a modified asset exchange model.

problem Analyzing economic inequality in a modified asset exchange model.
method Analytical results using Gini coefficient and differential inequalities.
result The Gini coefficient is bounded by a first-order differential inequality.

Unified view of improving tree model interpretability and debiasing feature importance.

problem Improving interpretability and debiasing feature importance in tree-based models.
method Demonstrates a common thread among bias correction methods and local explanations for trees.
result Points out a bias in explainable AI for trees algorithms due to inbag data inclusion.

The paper analyzes worst-case distortion risk metrics and weighted entropy under partial information.

problem Analyzing worst-case distortion risk metrics and weighted entropy with limited information.
method General distributions, partial information (mean and variance), various entropies and risk measures.
result Provides worst-case results for distortion risk metrics and weighted entropy.

Direct measurements of Gini coefficients by conventional arithmetic calculations are a poor estimator, even if paradoxically, they include the entire population, as because of super-additivity they cannot lend themselves to comparisons between units of different size, and intertemporal analyses are vitiated by the popu…

2015-10-16abs ↗pdf ↗

Hollow-tree Super resolves feature importance in large datasets.

problem Lack of effective scaling for large feature numbers in boosted tree models.
method Hollow-tree Super (HOTS) methodology for feature importance visualization.
result HOTS effectively resolves feature importance and directionality in high-dimensional neuroscientific data.

Two regularization techniques improve GCNN explainability and preference from chemists.

problem Difficulty in rationalizing molecular graph neural network predictions.
method Batch Representation Orthonormalization (BRO) and Gini regularization applied during GCNN training.
result Regularization improves GCNN attribution methods and preference from chemists.

The standard deviation and Gini mean difference order based on tail behavior.

problem Ordering between standard deviation and Gini mean difference for real-valued risks.
method Analysis of the mean excess function of the pairwise difference XX|X - X'|.
result Dominance regimes of SD and GMD are determined by tail behavior of the distribution.

Wealth redistribution through Fokker-Planck equation controls preserves Gini coefficient.

problem Preserving Gini coefficient through proportional wealth tax.
method Formulating optimal redistribution as a control problem for Fokker-Planck equation.
result Progressive taxes redistribute within policy-relevant timescales.

Socio-economic inequality is characterized from data using various indices. The Gini (gg) index, giving the overall inequality is the most common one, while the recently introduced Kolkata (kk) index gives a measure of 1k1-k fraction of population who possess top kk fraction of wealth in the society. Here, we show t…

2016-06-10abs ↗pdf ↗

The paper calculates bounds for risk metrics and entropies under partial information constraints.

problem Analyzing risk metrics and entropies for unimodal, symmetric distributions with limited information.
method Develops lower and upper bounds for worst-case distortion riskmetrics and weighted entropy for unimodal, symmetric distributions with known mean and variance.
result Sharp upper bounds for distortion riskmetrics and weighted entropy for symmetric distributions.

Using tax and census data, we demonstrate that the distribution of individual income in the USA is exponential. Our calculated Lorenz curve without fitting parameters and Gini coefficient 1/2 agree well with the data. From the individual income distribution, we derive the distribution function of income for families wi…

2000-08-21abs ↗pdf ↗

Framework monitors insurance pricing models for drift and recalibration.

problem Maintaining predictive performance of pricing models in evolving insurance portfolios.
method Formalizes deviance loss and Murphy's score, studies Gini score, develops monitoring framework.
result Framework guides decisions on refitting or recalibrating pricing models.

Sparse molecular representations improve interpretability in graph neural networks.

problem Difficulty in understanding which molecular graph aspects drive deep learning predictions.
method Constrain weights in a graph convolutional neural network using the Gini index to maximize representation inequality.
result The Gini-constrained approach does not degrade evaluation metrics and allows for interpretable representation combination.

New models reduce regional inequality by adjusting exchange range and asset distribution bias.

problem Reduction of regional inequality in economic systems.
method Proposed new asset exchange models with spatial exchange range and local support bias to adjust asset distribution and circulation rates.
result Achieved asset distribution from over-concentration to exponential and eventually normal, reducing Gini coefficient.

We study the effect of the social stratification on the wealth distribution on a system of interacting economic agents that are constrained to interact only within their own economic class. The economical mobility of the agents is related to its success in exchange transactions. Different wealth distributions are obtai…

2005-05-23abs ↗pdf ↗

We collect well known and less known facts about the bivariate normal distribution and translate them into copula language. In addition, we prove a very general formula for the bivariate normal copula, we compute Gini's gamma, and we provide improved bounds and approximations on the diagonal.

2009-12-15abs ↗pdf ↗

Study uses Perelman and Ricci flow methods to analyze economic inequality.

problem Impact of socio-economic challenges and technological progress on economic inequality.
method Perelman model and Ricci flow methods.
result Technological innovations and social protection programs reduce inequality.

Study uses random forest to detect unlawful insider trading in financial data.

problem Detecting and identifying unlawful insider trading in complex financial data.
method Integrates PCA-RF and standalone RF models with semi-manually labeled transactions.
result 96.43% accurate classification of transactions, 95.47% lawful, 98.00% unlawful.

Group Shapley evaluates feature groups in business data, improving explainability in AI.

problem Evaluating the importance of feature groups in business and economic data.
method Developed Group Shapley and a significance testing procedure based on chi-square approximation.
result Market-related variables are identified as the most influential feature group.

The construction of efficient and effective decision trees remains a key topic in machine learning because of their simplicity and flexibility. A lot of heuristic algorithms have been proposed to construct near-optimal decision trees. ID3, C4.5 and CART are classical decision tree algorithms and the split criteria they…

2015-11-25abs ↗pdf ↗