SUMER updates deployed models with new data, improving performance.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
A brief historical perspective is first given concerning financial crashes, - from the 17th till the 20th century. In modern times, it seems that log periodic oscillations are found before crashes in several financial indices. The same is found in sand pile avalanches on Sierpinski gaskets. A discussion pertains to the…
REMEDI improves neural entropy estimation across various tasks.
Paper presents a probabilistic diagnostic model for identifying and treating supervised learning degradation issues.
Nested Cavity Classifier (NCC) is a classification rule that pursues partitioning the feature space, in parallel coordinates, into convex hulls to build decision regions. It is claimed in some literatures that this geometric-based classifier is superior to many others, particularly in higher dimensions. First, we give …
We use automatic speech recognition to assess spoken English learner pronunciation based on the authentic intelligibility of the learners' spoken responses determined from support vector machine (SVM) classifier or deep learning neural network model predictions of transcription correctness. Using numeric features produ…
Study on W2S generalization with spurious correlations, proposing remedies.
This paper has been withdrawn by the authors, because it has been made obsolete by the detailed expositions in our papers in arXiv:0812.4885 (the mathematics part) and arXiv:0812.4737 (the economics part).
We develop new techniques in the theory of convex surfaces to prove complete classification results for tight contact structures on lens spaces, solid tori, and T^2 X I. Erratum: In this note we seek to remedy errors which appeared in version 2 and were propagated in subsequent papers.
Sigmoid-type networks avoid vanishing gradients with regularization and rescaling.
Deep Neural Network has proved its potential in various perception tasks and hence become an appealing option for interpretation and data processing in security sensitive systems. However, security-sensitive systems demand not only high perception performance, but also design robustness under various circumstances. Unl…
The problem of inhomogeneous cluster densities has been a long-standing issue for distance-based and density-based algorithms in clustering and anomaly detection. These algorithms implicitly assume that all clusters have approximately the same density. As a result, they often exhibit a bias towards dense clusters in th…
Every smooth fiber bundle admits a complete (Ehresmann) connection. This result appears in several references, with a proof on which we have found a gap, that does not seem possible to remedy. In this note we provide a definite proof for this fact, explain the problem with the previous one, and illustrate with examples…
The vector of periodic, compound returns of a typical investment portfolio is almost never a convex combination of the return vectors of the securities in the portfolio. As a result the ex post version of Harry Markowitz's "standard mean-variance portfolio selection model" does not apply to compound return data. We pro…
Transformers without skip connections collapse token representations to a single direction.
Unified AI system for data quality control and governance in regulated environments.
Despite their impressive performance in many tasks, deep neural networks often struggle at relational reasoning. This has recently been remedied with the introduction of a plug-in relational module that considers relations between pairs of objects. Unfortunately, this is combinatorially expensive. In this extended abst…
New framework identifies hidden risks and optionality in American options.
In retail, there are predictable yet dramatic time-dependent patterns in customer behavior, such as periodic changes in the number of visitors, or increases in customers just before major holidays. The current paradigm of multi-armed bandit analysis does not take these known patterns into account. This means that for a…
We study the properties of Expected Shortfall from the point of view of financial risk management. This measure --- which emerges as a natural remedy in some cases where Value at Risk (VaR) is not able to distinguish portfolios which bear different levels of risk --- is indeed shown to have much better properties than …
Existing feature selection methods fail to properly account for interactions between features when evaluating feature subsets. In this paper, we attempt to remedy this issue by using orthogonal variance decomposition to evaluate features. The orthogonality of the decomposition allows us to directly calculate the total …
Fibonacci anyons are attractive for use in topological quantum computation because any unitary transformation of their state space can be approximated arbitrarily accurately by braiding. However there is no known braid that entangles two qubits without leaving the space spanned by the two qubits. In other words, there …
The paper discusses methods to compute Green's function on algebraic surfaces using Schottky uniformization.
We investigate the difference between using an penalty versus an constraint in generalized eigenvalue problems, such as principal component analysis and discriminant analysis. Our main finding is that an penalty may fail to provide very sparse solutions; a severe disadvantage for variable sel…
This paper establishes that so-called instrumental variables enable the identification and the estimation of a fully nonparametric regression model with Berkson-type measurement error in the regressors. An estimator is proposed and proven to be consistent. Its practical performance and feasibility are investigated via …
This article is a response to the recent Worrying Trends in Econophysics critique written by four respected theoretical economists. Two of the four have written books and papers that provide very useful critical analyses of the shortcomings of the standard textbook economic model, neo-classical economic theory and have…
Develops algorithm to reduce real-world inequality.
In recent years, Deep Learning has become the go-to solution for a broad range of applications, often outperforming state-of-the-art. However, it is important, for both theoreticians and practitioners, to gain a deeper understanding of the difficulties and limitations associated with common approaches and algorithms. W…
Added examples of S^1-manifolds with finite 2nd homotopy group and non-zero A-genus.
It is important to develop mathematically tractable models than can interpret knowledge extracted from the data and provide reasonable predictions. In this paper, we present a Linear Distillation Learning, a simple remedy to improve the performance of linear neural networks. Our approach is based on using a linear func…
The amount of information in the form of features and variables avail- able to machine learning algorithms is ever increasing. This can lead to classifiers that are prone to overfitting in high dimensions, high di- mensional models do not lend themselves to interpretable results, and the CPU and memory resources necess…
Normalization methods are a central building block in the deep learning toolbox. They accelerate and stabilize training, while decreasing the dependence on manually tuned learning rate schedules. When learning from multi-modal distributions, the effectiveness of batch normalization (BN), arguably the most prominent nor…
Node2vec embeddings are unstable and unstable with parameter choices.
Much of the recent work on learning molecular representations has been based on Graph Convolution Networks (GCN). These models rely on local aggregation operations and can therefore miss higher-order graph properties. To remedy this, we propose Path-Augmented Graph Transformer Networks (PAGTN) that are explicitly built…
Generalized Stacey-Roberts lemma for Banach manifolds.
Classical (Itô diffusions) stochastic volatility models are not able to capture the steepness of small-maturity implied volatility smiles. Jumps, in particular exponential Lévy and affine models, which exhibit small-maturity exploding smiles, have historically been proposed to remedy this (see \cite{Tank} for an overvi…
Supervised training of deep learning models requires large labeled datasets. There is a growing interest in obtaining such datasets for medical image analysis applications. However, the impact of label noise has not received sufficient attention. Recent studies have shown that label noise can significantly impact the p…
The Fisher information approximation (FIA) is an implementation of the minimum description length principle for model selection. Unlike information criteria such as AIC or BIC, it has the advantage of taking the functional form of a model into account. Unfortunately, FIA can be misleading in finite samples, resulting i…
The risks and perils of overfitting in machine learning are well known. However most of the treatment of this, including diagnostic tools and remedies, was developed for the supervised learning case. In this work, we aim to offer new perspectives on the characterization and prevention of overfitting in deep Reinforceme…
Optimal biomarker combinations for treatment-selection can be derived by minimizing total burden to the population caused by the targeted disease and its treatment. However, when multiple biomarkers are present, including all in the model can be expensive and hurt model performance. To remedy this, we consider feature …
Explainable Artificial Intelligence (XAI)has received a great deal of attention recently. Explainability is being presented as a remedy for the distrust of complex and opaque models. Model agnostic methods such as LIME, SHAP, or Break Down promise instance-level interpretability for any complex machine learning model. …
GrowNet uses shallow neural networks for gradient boosting, outperforming existing methods.
Paper introduces a new gradient statistic to improve deep learning convergence.
Dealing with uncertainty is essential for efficient reinforcement learning. There is a growing literature on uncertainty estimation for deep learning from fixed datasets, but many of the most popular approaches are poorly-suited to sequential decision problems. Other methods, such as bootstrap sampling, have no mechani…
We demonstrate both analytically and numerically that the existing methods for measuring tail dependence in copulas may sometimes underestimate the extent of extreme co-movements of dependent risks and, therefore, may not always comply with the new paradigm of prudent risk management. This phenomenon holds in the conte…
The study aims to prevent unfair content presentation in recommender systems.
Proposes efficient calibration for indoor localization models.
I sketch a program for a microeconomic theory of the main component of the business cycle as a recurring disequilibrium, driven by incompleteness of the financial market and by information asymmetries between borrowers and lenders. This proposal seeks to incorporate five distinct but connected processes that have been …