New framework controls statistical dispersion for high-stakes applications.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Introduces data ethics for mathematicians, covering background, open data, and privacy.
When society maintains a competitive system to promote an abstract goal, competition by necessity relies on imperfect proxy measures. For instance profit is used to measure value to consumers, patient volumes to measure hospital performance, or the Journal Impact Factor to measure scientific value. Here we note that \t…
Machine learning is used extensively in recommender systems deployed in products. The decisions made by these systems can influence user beliefs and preferences which in turn affect the feedback the learning system receives - thus creating a feedback loop. This phenomenon can give rise to the so-called "echo chambers" …
Study finds 'happiness' search data predicts stock returns, suggesting utility needs impact firm performance.
PCL framework optimizes climate risk management across three clusters.
New fair regression method improves fairness in chronic kidney disease classification.
Framework generates precise synthetic populations for scalable modeling.
The potential for learned models to amplify existing societal biases has been broadly recognized. Fairness-aware classifier constraints, which apply equality metrics of performance across subgroups defined on sensitive attributes such as race and gender, seek to rectify inequity but can yield non-uniform degradation in…
Proposes a new fairness definition based on equity for machine learning classification.
Survey of technologies for trustworthy machine learning systems.
Improves fairness in machine learning by adding underrepresented group data.
Increasingly, discrimination by algorithms is perceived as a societal and legal problem. As a response, a number of criteria for implementing algorithmic fairness in machine learning have been developed in the literature. This paper proposes the Continuous Fairness Algorithm (CFA) which enables a continuous interpol…
New fairness criteria for algorithmic recourse actions that consider causal relationships.
Proposes CSRN for better news recommendation by integrating RNN and UserCF.
Survey on biases in image analysis for industrial safety.
New bounds show multicalibration error is close to prediction error.
The last decades have not only been characterized by an explosive growth of data, but also an increasing appreciation of data as a valuable resource. Their value comes with the ability to extract meaningful patterns that are of economic, societal or scientific relevance. A particular challenge is the identification of …
Study finds telemetric data not effective for predicting truck accident risk.
Economies and societal structures in general are complex stochastic systems which may not lend themselves well to algebraic analysis. An addition of subjective value criteria to the mechanics of interacting agents will further complicate analysis. The purpose of this short study is to demonstrate capabilities of agent-…
We propose a novel formulation of group fairness with biased feedback in the contextual multi-armed bandit (CMAB) setting. In the CMAB setting, a sequential decision maker must, at each time step, choose an arm to pull from a finite set of arms after observing some context for each of the potential arm pulls. In our mo…
Text corpora are widely used resources for measuring societal biases and stereotypes. The common approach to measuring such biases using a corpus is by calculating the similarities between the embedding vector of a word (like nurse) and the vectors of the representative words of the concepts of interest (such as gender…
Notions of "fair classification" that have arisen in computer science generally revolve around equalizing certain statistics across protected groups. This approach has been criticized as ignoring societal issues, including how errors can hurt certain groups disproportionately. We pose a modification of one of the fairn…
Study reveals DNNs prefer easy-to-learn cues over essential ones in image recognition.
Paper investigates differentiable fuzzy implications and their suitability for learning.
Adversarial attacks can manipulate ML-aided visualizations, tricking analysts.
This paper defines less discriminatory algorithms and explores their feasibility.
Societal bias towards certain communities is a big problem that affects a lot of machine learning systems. This work aims at addressing the racial bias present in many modern gender recognition systems. We learn race invariant representations of human faces with an adversarially trained autoencoder model. We show that …
The paper analyzes frameworks for integrating sustainability into investment decisions.
TransINT embeds KGs by preserving implication rules, outperforming existing methods.
A popular approach of achieving fairness in optimization problems is by constraining the solution space to "fair" solutions, which unfortunately typically reduces solution quality. In practice, the ultimate goal is often an aggregate of sub-goals without a unique or best way of combining them or which is otherwise only…
This thesis tackles bias in AI decision-making in banking.
Machine learning (ML) is increasingly being used in high-stakes applications impacting society. Therefore, it is of critical importance that ML models do not propagate discrimination. Collecting accurate labeled data in societal applications is challenging and costly. Active learning is a promising approach to build an…
Study aims to measure and mitigate biases in motor insurance pricing.
Bayesian model forecasts hospital resource use during pandemic.
Actuaries tackle loss of earning capacity in Denmark, balancing public benefits and private insurance.
Relational data in its most basic form is a static collection of known facts. However, by learning to infer and deduct additional information and structure, we can massively increase the usefulness of the underlying data. One common form of inferential reasoning in knowledge bases is implication discovery. Here, by lea…
AI boosts study of rare weather extremes with lower costs.
Fair active learning selects data points to balance model accuracy and fairness.
A textbook on machine learning explaining patterns, predictions, and actions.
The study of linguistic typology is rooted in the implications we find between linguistic features, such as the fact that languages with object-verb word ordering tend to have post-positions. Uncovering such implications typically amounts to time-consuming manual processing by trained and experienced linguists, which p…
Mitigates gender bias amplification in model predictions.
This study conducts a comprehensive analysis of time series segmentation on the Japanese stock prices listed on the first section of the Tokyo Stock Exchange during the period from 4 January 2000 to 30 January 2012. A recursive segmentation procedure is used under the assumption of a Gaussian mixture. The daily number …
It has been suggested in 1999 that a certain volume growth condition for geodesically complete Riemannian manifolds might imply that the manifold is stochastically complete. This is motivated by a large class of examples and by a known analogous criterion for recurrence of Brownian motion. We show that the suggested im…
Interacting systems are prevalent in nature, from dynamical systems in physics to complex societal dynamics. The interplay of components can give rise to complex behavior, which can often be explained using a simple model of the system's constituent parts. In this work, we introduce the neural relational inference (NRI…
Deep Neural Networks (DNNs) are universal function approximators providing state-of- the-art solutions on wide range of applications. Common perceptual tasks such as speech recognition, image classification, and object tracking are now commonly tackled via DNNs. Some fundamental problems remain: (1) the lack of a mathe…
Data is one of the most important assets of the information age, and its societal impact is undisputed. Yet, rigorous methods of assessing the quality of data are lacking. In this paper, we propose a formal definition for the quality of a given dataset. We assess a dataset's quality by a quantity we call the expected d…
The study examines how social biases are reinforced in machine learning models used for credit scoring.