This dissertation tackles challenges in reliable machine learning measurement.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Standard agglomerative clustering suggests establishing a new reliable linkage at every step. However, in order to provide adaptive, density-consistent and flexible solutions, we study extracting all the reliable linkages at each step, instead of the smallest one. Such a strategy can be applied with all common criteria…
Bayesian optimization has been proposed as a practical and efficient tool through which to tune parameters in many difficult settings. Recently, such techniques have been combined with real-time fMRI to propose a novel framework which turns on its head the conventional functional neuroimaging approach. This closed-loop…
XAI methods fail to explain ML models reliably.
This paper combines and improves probabilistic forecasts of wind speeds using advanced statistical methods.
Active learning has shown to reduce the number of experiments needed to obtain high-confidence drug-target predictions. However, in order to actually save experiments using active learning, it is crucial to have a method to evaluate the quality of the current prediction and decide when to stop the experimentation proce…
The majority of traditional classification ru les minimizing the expected probability of error (0-1 loss) are inappropriate if the class probability distributions are ill-defined or impossible to estimate. We argue that in such cases class domains should be used instead of class distributions or densities to construct …
Survey and framework for efficient active learning in structural reliability.
This tutorial evaluates machine learning for healthcare applications.
The paper tackles causal rule discovery from observational data.
Throughout the past five years, the susceptibility of neural networks to minimal adversarial perturbations has moved from a peculiar phenomenon to a core issue in Deep Learning. Despite much attention, however, progress towards more robust models is significantly impaired by the difficulty of evaluating the robustness …
New method combines neural networks with Monte Carlo for complex system reliability.
Deep learning detects microsleep episodes in EEG data.
Improves reliability of BBVI optimization methods.
IndiSeek learns disentangled representations by balancing independence and completeness.
As researchers and practitioners of applied machine learning, we are given a set of requirements on the problem to be solved, the plausibly obtainable data, and the computational resources available. We aim to find (within those bounds) reliably useful combinations of problem, data, and algorithm. An emphasis on algori…
Bayesian active learning improves holistic educational assessments.
Paper tackles ESG rating disagreement in sustainable investing portfolios.
In order to obtain a reasonable and reliable forecast method for crude oil price volatility, this paper evaluates the forecast performance of single-regime GARCH models (including the standard linear GARCH model and the nonlinear GJR-GARCH and EGARCH models) and the two-regime Markov Regime Switching GARCH (MRS-GARCH) …
A new method improves robustness and efficiency of Bayesian LOO-CV.
New method tests tree models without causing computational pressure.
Mathematical advances needed for Digital Twins, differing from traditional models.
We consider the problem of estimating the set of all inputs that leads a system to some particular behavior. The system is modeled by an expensive-to-evaluate function, such as a computer experiment, and we are interested in its excursion set, i.e. the set of points where the function takes values above or below some p…
Paper refines royalty determination using Bayesian methods.
Missing data and noisy observations pose significant challenges for reliably predicting events from irregularly sampled multivariate time series (longitudinal) data. Imputation methods, which are typically used for completing the data prior to event prediction, lack a principled mechanism to account for the uncertainty…
We give some general criteria of being a homeomorphism for continuous mappings of topological manifolds, as well as criteria of being a diffeomorphism for smooth mappings of smooth manifolds. As an illustration, we apply these criteria to the problems arising in two- and three-dimensional grid generation.
The study reveals flaws in pruning criteria and proposes a new assumption for better filter selection.
New criteria for Heegaard splittings ensure strong irreducibility and finite Goeritz groups.
Multi-criteria recommender systems have been increasingly valuable for helping consumers identify the most relevant items based on different dimensions of user experiences. However, previously proposed multi-criteria models did not take into account latent embeddings generated from user reviews, which capture latent se…
Develops scenario theory for multi-criteria decision making.
We consider the problem of identifying patterns in a data set that exhibit anomalous behavior, often referred to as anomaly detection. In most anomaly detection algorithms, the dissimilarity between data samples is calculated by a single criterion, such as Euclidean distance. However, in many cases there may not exist …
The paper evaluates criteria for selecting cryptocurrencies based on historical data.
Small Medium-sized Enterprises (SMEs) face many obstacles when they try to access credit market. These obstacles are increased if the SMEs are innovative. In this case, financial data are insufficient or even not reliable. Thus, when building a judgemental rating model, mainly based on qualitative criteria (soft inform…
Classification algorithms aim to predict an unknown label (e.g., a quality class) for a new instance (e.g., a product). Therefore, training samples (instances and labels) are used to deduct classification hypotheses. Often, it is relatively easy to capture instances but the acquisition of the corresponding labels remai…
Study evaluates five LLMs for financial report analysis, revealing performance differences and variability.
Model selection for time series forecasting can be biased by the distribution of scores.
We present criteria for establishing a triangulation of a manifold. Given a manifold M, a simplicial complex A, and a map H from the underlying space of A to M, our criteria are presented in local coordinate charts for M, and ensure that H is a homeomorphism. These criteria do not require a differentiable structure, or…
When sufficient labeled data are available, classical criteria based on Receiver Operating Characteristic (ROC) or Precision-Recall (PR) curves can be used to compare the performance of un-supervised anomaly detection algorithms. However , in many situations, few or no data are labeled. This calls for alternative crite…
The paper analyzes performance criteria for competing fund managers in Ito-diffusion markets.
A new AI framework reduces costs and improves performance.
We shall give useful criteria of lips, beaks and swallowtail singularities of smooth map from the plane into the plane. As an application of criteria, we will discuss the singularities of Cauchy problem of single conservation law.
A new method for multi-criteria recommender systems using graph attention networks.
New framework for resilient bi-criteria optimization under noisy feedback.
New algorithms minimize risk in MNL bandits, achieving near-optimal performance.
Regression models are increasingly built using datasets which do not follow a design of experiment. Instead, the data is e.g. gathered by an automated monitoring of a technical system. As a consequence, already the input data represents phenomena of the system and violates statistical assumptions of distributions. The …
A game-theoretic approach to multi-criteria ranking from ordinal data.
Stress, edge crossings, and crossing angles play an important role in the quality and readability of graph drawings. Most standard graph drawing algorithms optimize one of these criteria which may lead to layouts that are deficient in other criteria. We introduce an optimization framework, Stress-Plus-X (SPX), that sim…
Recent work on fairness in machine learning has focused on various statistical discrimination criteria and how they trade off. Most of these criteria are observational: They depend only on the joint distribution of predictor, protected attribute, features, and outcome. While convenient to work with, observational crite…