Bayesian deep learning predicts satellite collisions.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Space debris warnings follow a predictable pattern, allowing timely satellite maneuvers.
We consider a Black-Scholes market in which a number of stocks and an index are traded. The simplified Capital Asset Pricing Model is the conjunction of the usual Capital Asset Pricing Model, or CAPM, and the statement that the appreciation rate of the index is equal to its squared volatility plus the interest rate. (T…
A Bayesian network is a graphical model that encodes probabilistic relationships among variables of interest. When used in conjunction with statistical techniques, the graphical model has several advantages for data analysis. One, because the model encodes dependencies among all variables, it readily handles situations…
Machine learning competition predicts spacecraft collision risks.
One of the objectives of designing feature selection learning algorithms is to obtain classifiers that depend on a small number of attributes and have verifiable future performance guarantees. There are few, if any, approaches that successfully address the two goals simultaneously. Performance guarantees become crucial…
Develops a new GLM framework for claims reserving with adaptive estimation.
The paper introduces closed-form expressions for interpreting Tsetlin Machines.
Study on list learning with noisy data, showing limits and some learnable cases.
New estimator improves off-policy evaluation for large action spaces.
There is a need for the development of models that are able to account for discreteness in data, along with its time series properties and correlation. Our focus falls on INteger-valued AutoRegressive (INAR) type models. The INAR type models can be used in conjunction with existing model-based clustering techniques to …
In this article, the logic rule ensembles approach to supervised learning is applied to the unsupervised or semi-supervised clustering. Logic rules which were obtained by combining simple conjunctive rules are used to partition the input space and an ensemble of these rules is used to define a similarity matrix. Simila…
We obtain an expression for the curvature of the Lie group SDiff and use it to derive Lukatskii's formula for the case where is locally Euclidean. We discuss qualitatively some previous findings for SDiff in conjunction with our result.
We prove that there are examples of finitely generated groups G together with group ring elements Q \in \bbQ G for which the von Neumann dimension \dim_{LG}\ker Q is irrational, so (in conjunction with other known results) answering a question of Atiyah.
Better use of unlabelled data improves Bayesian active learning models.
Our team won the second prize of the Safe Aging with SPHERE Challenge organized by SPHERE, in conjunction with ECML-PKDD and Driven Data. The goal of the competition was to recognize activities performed by humans, using sensor data. This paper presents our solution. It is based on a rich pre-processing and state of th…
A new clustering algorithm tracks satellite hotspot data for bushfire tracking.
We propose the Gaussian attention model for content-based neural memory access. With the proposed attention model, a neural network has the additional degree of freedom to control the focus of its attention from a laser sharp attention to a broad attention. It is applicable whenever we can assume that the distance in t…
We develop an online learning method for prediction, which is important in problems with large and/or streaming data sets. We formulate the learning approach using a covariance-fitting methodology, and show that the resulting predictor has desirable computational and distribution-free properties: It is implemented onli…
Ranked data appear in many different applications, including voting and consumer surveys. There often exhibits a situation in which data are partially ranked. Partially ranked data is thought of as missing data. This paper addresses parameter estimation for partially ranked data under a (possibly) non-ignorable missing…
This is a survey of some of the work of Tom Farrell and Lowell Jones. This is the lead article of a special issue of the Pure and Applied Mathematics Quarterly. This issue is published in conjunction with the conference "Geometry,Topology, and their Interactions" held in Morelia, Mexico.
Proposes a robust similarity measure for sparse time series data.
In this paper, we propose and study random maxout features, which are constructed by first projecting the input data onto sets of randomly generated vectors with Gaussian elements, and then outputing the maximum projection value for each set. We show that the resulting random feature map, when used in conjunction with …
We propose a general-purpose approach to discovering active learning (AL) strategies from data. These strategies are transferable from one domain to another and can be used in conjunction with many machine learning models. To this end, we formalize the annotation process as a Markov decision process, design universal s…
In this paper, we use fully convolutional neural networks for the semantic segmentation of eye tracking data. We also use these networks for reconstruction, and in conjunction with a variational auto-encoder to generate eye movement data. The first improvement of our approach is that no input window is necessary, due t…
This paper, to be regularly updated, lists those prime knots with the fewest possible number of crossings for which values of basic knot invariants, such as the unknotting number or the smooth 4-genus, are unknown. This list is being developed in conjunction with "KnotInfo" (www.indiana.edu/~knotinfo), a web-based tabl…
We prove the equality of the analytic torsion and the value at zero of a Ruelle dynamical zeta function associated with an acyclic unitarily flat vector bundle on a closed locally symmetric reductive manifold. This solves a conjecture of Fried. This article should be read in conjunction with an earlier paper by Moscovi…
We construct two knot invariants. The first knot invariant is a sum constructed using linking numbers. The second is an invariant of flat knots and is a formal sum of flat knots obtained by smoothing pairs of crossings. This invariant can be used in conjunction with other flat invariants, forming a family of invariants…
This paper presents a general notion of Mahalanobis distance for functional data that extends the classical multivariate concept to situations where the observed data are points belonging to curves generated by a stochastic process. More precisely, a new semi-distance for functional observations that generalize the usu…
Recent years have seen rapid advances in the data-driven analysis of dynamical systems based on Koopman operator theory and related approaches. On the other hand, low-rank tensor product approximations -- in particular the tensor train (TT) format -- have become a valuable tool for the solution of large-scale problems …
Ricci flow preserves positive sectional curvature on homogeneous spheres
Framework generates causal probabilities from observational data.
This paper introduces the combinatorial Boolean model (CBM), which is defined as the class of linear combinations of conjunctions of Boolean attributes. This paper addresses the issue of learning CBM from labeled data. CBM is of high knowledge interpretability but naïve learning of it requires exponentially large compu…
Anomaly detection is of great interest in fields where abnormalities need to be identified and corrected (e.g., medicine and finance). Deep learning methods for this task often rely on autoencoder reconstruction error, sometimes in conjunction with other errors. We show that this approach exhibits intrinsic biases that…
Can textual data be compressed intelligently without losing accuracy in evaluating sentiment? In this study, we propose a novel evolutionary compression algorithm, PARSEC (PARts-of-Speech for sEntiment Compression), which makes use of Parts-of-Speech tags to compress text in a way that sacrifices minimal classification…
We note an area-charge inequality orignially due to Gibbons: if the outermost horizon in an asymptotically flat electrovacuum initial data set is connected then , where is the total charge and is the area radius of . A consequence of this inequality is that for connected black hole…
Two sets of high quality income data are analysed in detail, one set from the UK, one from the USA. It is firstly demonstrated that both a log-normal distribution and a Boltzmann distribution can give very accurate fits to both these data sets. The absence of a power tail in the US data set is then discussed. Taken in …
We establish continuous maximal regularity results for parabolic differential operators acting on sections of tensor bundles on Riemannian manifolds. As an application, we show that solutions to the Yamabe flow instantaneously regularize and become real analytic in space and time. The regularity result is obtained by i…
We apply machine learning to the problem of finding numerical Calabi-Yau metrics. Building on Donaldson's algorithm for calculating balanced metrics on Kähler manifolds, we combine conventional curve fitting and machine-learning techniques to numerically approximate Ricci-flat metrics. We show that machine learning is …
It is becoming increasingly important to understand the vulnerability of machine learning models to adversarial attacks. In this paper we study the feasibility of robust learning from the perspective of computational learning theory, considering both sample and computational complexity. In particular, our definition of…
Survey on heat equation estimates on manifolds.
For certain problems involving vector fields, it is possible to find an associated imaginary field that, in conjunction with the first, forms a complex field for which the equation can be solved. This result is generalized to arbitrary Clifford algebras, followed by quaternionic vectors as a special case. All results a…
FML uses neural networks to model unknown systems accurately.
Applying machine learning techniques to the quickly growing data in science and industry requires highly-scalable algorithms. Large datasets are most commonly processed "data parallel" distributed across many nodes. Each node's contribution to the overall gradient is summed using a global allreduce. This allreduce is t…
Trivial links are unique up to number of link components, but they can be hard to recognize from arbitrary diagrams. We define a new measure of the complexity of a link embedding, the crumple, and show how this may be used to measure progress toward a trivial embedding. In conjunction with a modified form of arc presen…
Develops theory for data-driven methods in dynamical systems.
Dynamic jumps in the price and volatility of an asset are modelled using a joint Hawkes process in conjunction with a bivariate jump diffusion. A state space representation is used to link observed returns, plus nonparametric measures of integrated volatility and price jumps, to the specified model components; with Bay…
Cross-sectional signatures of market panic were recently discussed on daily time scales in [1], extended here to a study of cross-sectional properties of stocks on intra-day time scales. We confirm specific intra-day patterns of dispersion and kurtosis, and find that the correlation across stocks increases in times of …