When response variables are nominal and populations are cross-classified with respect to multiple polytomies, questions often arise about the degree of association of the responses with explanatory variables. When populations are known, we introduce a nominal association vector and matrix to evaluate the dependence of …
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We construct models for the pricing and risk management of inflation-linked derivatives. The models are rational in the sense that linear payoffs written on the consumer price index have prices that are rational functions of the state variables. The nominal pricing kernel is constructed in a multiplicative manner that …
Generalizes causal inference to high-dimensional outcomes.
Proposes a new framework for uncertainty evaluation in ML classification models.
We develop a novel probabilistic generative model based on the variational autoencoder approach. Notable aspects of our architecture are: a novel way of specifying the latent variables prior, and the introduction of an ordinality enforcing unit. We describe how to do supervised, unsupervised and semi-supervised learnin…
Suppose that one particular block in a stochastic block model is of interest, but block labels are only observed for a few of the vertices in the network. Utilizing a graph realized from the model and the observed block labels, the vertex nomination task is to order the vertices with unobserved block labels into a rank…
Given a vertex of interest in a network , the vertex nomination problem seeks to find the corresponding vertex of interest (if it exists) in a second network . A vertex nomination scheme produces a list of the vertices in , ranked according to how likely they are judged to be the corresponding vertex of …
In this paper, a novel approach for coding nominal data is proposed. For the given nominal data, a rank in a form of complex number is assigned. The proposed method does not lose any information about the attribute and brings other properties previously unknown. The approach based on these knew properties can been used…
The paper introduces subgraph nomination for finding similar subgraphs in networks.
Suppose that a graph is realized from a stochastic block model where one of the blocks is of interest, but many or all of the vertices' block labels are unobserved. The task is to order the vertices with unobserved block labels into a ``nomination list'' such that, with high probability, vertices from the interesting b…
In this paper we develop a method for learning nonlinear systems with multiple outputs and inputs. We begin by modelling the errors of a nominal predictor of the system using a latent variable framework. Then using the maximum likelihood principle we derive a criterion for learning the model. The resulting optimization…
Given a graph in which a few vertices are deemed interesting a priori, the vertex nomination task is to order the remaining vertices into a nomination list such that there is a concentration of interesting vertices at the top of the list. Previous work has yielded several approaches to this problem, with theoretical re…
We present the Mixed Likelihood Gaussian process latent variable model (GP-LVM), capable of modeling data with attributes of different types. The standard formulation of GP-LVM assumes that each observation is drawn from a Gaussian distribution, which makes the model unsuited for data with e.g. categorical or nominal a…
The paper explores how to find relevant vertices in one graph using another graph's attributes and structure.
Researchers developed a generic model to account for structural variability in SHM.
PANDA augments data to regularize GLM estimation and inference.
In the framework of prediction with expert advice, we consider a recently introduced kind of regret bounds: the bounds that depend on the effective instead of nominal number of experts. In contrast to the Normal- Hedge bound, which mainly depends on the effective number of experts but also weakly depends on the nominal…
Two-stage recommender systems show better performance when components interact rather than operate independently.
Develops FSC for maxima nominated samples, improving classification in rare-event data.
Paper proposes a new daily benchmark for post-GFC government bond CIP deviations.
New method reduces bias in estimating causal effects from discretized variables.
This paper presents a distributionally robust Q-Learning algorithm (DrQ) which leverages Wasserstein ambiguity sets to provide idealistic probabilistic out-of-sample safety guarantees during online learning. First, we follow past work by separating the constraint functions from the principal objective to create a hiera…
A new method clusters mixed-type data efficiently.
TabSODA improves imputation of surveys with skips and ordinal data.
We demonstrate the existence of an empirical linkage between the nominal financial networks and the underlying economic fundamentals across countries. We construct the nominal return correlation networks from daily data to encapsulate sector-level dynamics and figure the relative importance of the sectors in the nomina…
Responds to critiques on tests for causal parameter confidence intervals.
A fundamental problem arising in many areas of machine learning is the evaluation of the likelihood of a given observation under different nominal distributions. Frequently, these nominal distributions are themselves estimated from data, which makes them susceptible to estimation errors. We thus propose to replace each…
Enhances deep learning models for anomaly detection in time series data.
Timely detection of abrupt anomalies is crucial for real-time monitoring and security of modern systems producing high-dimensional data. With this goal, we propose effective and scalable algorithms. Proposed algorithms are nonparametric as both the nominal and anomalous multivariate data distributions are assumed unkno…
A new framework estimates causal effects for ordinal variables.
Paper develops PGMM framework for debiased inference on nonparametric IV estimators.
This work models variability in composite blades' vibrations.
This paper fortifies the recently introduced hierarchical-optimization recursive least squares (HO-RLS) against outliers which contaminate infrequently linear-regression models. Outliers are modeled as nuisance variables and are estimated together with the linear filter/system variables via a sparsity-inducing (non-)co…
New RL algorithm learns robust policies without knowing nominal model.
Two-stage recommender systems struggle with exploration, leading to linear regret.
We propose an AdaPtive Noise Augmentation (PANDA) technique to regularize the estimation and construction of undirected graphical models. PANDA iteratively optimizes the objective function given the noise augmented data until convergence to achieve regularization on model parameters. The augmented noises can be designe…
Information regarding the location of power distribution grid can be extracted from the power signature embedded in the multimedia signals (e.g., audio, video data) recorded near electrical activities. This implicit mechanism of identifying the origin-of-recording can be a very promising tool for multimedia forensics a…
In this paper a new method for heat load prediction in district energy systems is proposed. The method uses a nominal model for the prediction of the outdoor temperature dependent space heating load, and a data driven latent variable model to predict the time dependent residual heat load. The residual heat load arises …
A method for ranking items using distance-based learning from positive and unlabeled data.
Given a pair of graphs and and a vertex set of interest in , the vertex nomination (VN) problem seeks to find the corresponding vertices of interest in (if they exist) and produce a rank list of the vertices in , with the corresponding vertices of interest in concentrating, ideally, at…
Resolving abstract anaphora is an important, but difficult task for text understanding. Yet, with recent advances in representation learning this task becomes a more tangible aim. A central property of abstract anaphora is that it establishes a relation between the anaphor embedded in the anaphoric sentence and its (ty…
New model detects communities in network data from edge nominations.
Investigates optimal life insurance and annuity decisions in inflationary economies.
We propose a method for estimation in high-dimensional linear models with nominal categorical data. Our estimator, called SCOPE, fuses levels together by making their corresponding coefficients exactly equal. This is achieved using the minimax concave penalty on differences between the order statistics of the coefficie…
Adversarial robustness improved by abstaining from decisions.
In this paper we introduce a class of information-based models for the pricing of fixed-income securities. We consider a set of continuous- time information processes that describe the flow of information about market factors in a monetary economy. The nominal pricing kernel is at any given time assumed to be given by …
Exact distribution of split conformal prediction coverage found.
Robust subgroup discovery finds non-redundant, statistically significant subgroups.