Paper estimates spectral risk measures for insurance data with truncated and censored data.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Bayesian method estimates LTLL distribution parameters for time-to-event data.
New method estimates treatment effects over time for survival data, improving accuracy and smoothness.
This paper gives quantitative global estimates between a time dependent flow on a Riemannian manifold and the flow of a vector field constructed by truncating the formal Magnus expansion for the logarithm of the flow. As a corollary, we also find quantitative estimates between the composition of the …
Unified framework for mean testing under truncation bias.
A new method prices time-to-event cash flows using survival analysis.
The paper estimates common mean of entangled Gaussians with bounded variances.
In a previous analysis the problem of "zero-inflated" time data (caused by high frequency trading in the electronic order book) was handled by left-truncating the inter-arrival times. We demonstrated, using rigorous statistical methods, that the Weibull distribution describes the corresponding stochastic dynamics for a…
We show that given an estimate that is close to a general high-rank positive semi-definite (PSD) matrix in spectral norm (i.e., ), the simple truncated SVD of produces a multiplicative approximation of in Frobenius norm. This observation leads to many inte…
The paper efficiently estimates parameters from truncated Gaussian and linear models.
Paper establishes tight lower bounds for minimizing certain smooth and convex functions.
Non-negative matrix factorization (NMF) minimizes the Euclidean distance between the data matrix and its low rank approximation, and it fails when applied to corrupted data because the loss function is sensitive to outliers. In this paper, we propose a Truncated CauchyNMF loss that handle outliers by truncating large e…
New DP framework using data truncation for efficient estimation.
New method for constructing truncated vine copulas.
Score matching method improves density estimation for truncated data on manifolds.
Proposes a method to handle sparse multiway count data with false zeros using zero-truncated Poisson regression.
Truncated densities are probability density functions defined on truncated domains. They share the same parametric form with their non-truncated counterparts up to a normalizing constant. Since the computation of their normalizing constants is usually infeasible, Maximum Likelihood Estimation cannot be easily applied t…
Typically, operational risk losses are reported above some threshold. This paper studies the impact of ignoring data truncation on the 0.999 quantile of the annual loss distribution for operational risk for a broad range of distribution parameters and truncation levels. Loss frequency and severity are modelled by the P…
The paper investigates the convergence of Vendi scores under finite samples and introduces a truncated version for better performance.
Paper proposes robust estimators for heavy-tailed data with infinite variance.
Truncated backpropagation through time (TBPTT) is a popular method for learning in recurrent neural networks (RNNs) that saves computation and memory at the cost of bias by truncating backpropagation after a fixed number of lags. In practice, choosing the optimal truncation length is difficult: TBPTT will not converge …
Estimates domain truncation error for option pricing PDEs.
The problem of an arbitrary truncated Levy flight description using the method of cumulant approach has been solved. The set of cumulants of the truncated Levy distribution given the assumption of arbitrary truncation has been found. The influence of truncation shape on the truncated Levy flight properties in the Gauss…
We study inference and learning based on a sparse coding model with `spike-and-slab' prior. As in standard sparse coding, the model used assumes independent latent sources that linearly combine to generate data points. However, instead of using a standard sparse prior such as a Laplace distribution, we study the applic…
Deep neural RDEs improve portfolio optimization accuracy and risk sensitivity.
Efficiently estimate Boolean product distribution parameters from truncated samples.
This article describes an implementation of a nonparametric Bayesian approach to solving binary classification problems on graphs. We consider a hierarchical Bayesian approach with a prior that is constructed by truncating a series expansion of the soft label function using the graph Laplacian eigenfunctions as basis f…
UDN adapts depth to data complexity, outperforming standard neural networks.
In the paper "On Truncated Variation of Brownian Motion with Drift" (Bull. Pol. Acad. Sci. Math. 56 (2008), no.4, 267 - 281) we defined truncated variation of Brownian motion with drift, where is a standard Brownian motion. Truncated variation differs from regular variation by neglect…
We consider the problem of online linear regression on arbitrary deterministic sequences when the ambient dimension d can be much larger than the number of time rounds T. We introduce the notion of sparsity regret bound, which is a deterministic online counterpart of recent risk bounds derived in the stochastic setting…
As in standard linear regression, in truncated linear regression, we are given access to observations whose dependent variable equals , where is some fixed unknown vector of interest and is independent noise; except we are only given an observation if its dep…
Learning with a {\it convex loss} function has been a dominating paradigm for many years. It remains an interesting question how non-convex loss functions help improve the generalization of learning with broad applicability. In this paper, we study a family of objective functions formed by truncating traditional loss f…
Optimal algorithm learns Gaussian under halfspace truncation with minimal samples.
PMT uses public data moments to make DP feasible for unbounded data.
Stochastic gradient descent (SGD) is commonly used for optimization in large-scale machine learning problems. Langford et al. (2009) introduce a sparse online learning method to induce sparsity via truncated gradient. With high-dimensional sparse data, however, the method suffers from slow convergence and high variance…
Combines ML and DA to infer unresolved scale parametrisation from noisy data.
Truncated SGD with heavy-tailed noise eliminates sharp local minima.
We develop a scale-invariant truncated Lévy (STL) process to describe physical systems characterized by correlated stochastic variables. The STL process exhibits Lévy stability for the probability density, and hence shows scaling properties (as observed in empirical data); it has the advantage that all moments are fini…
Paper proposes approximate Stein classes for efficient truncated density estimation.
Paper defines new risk measures for elliptical distributions.
Derives equations for deep learning biases and weights, showing data complexity reduction.
Kernel ridge regression (KRR) is a well-known and popular nonparametric regression approach with many desirable properties, including minimax rate-optimality in estimating functions that belong to common reproducing kernel Hilbert spaces (RKHS). The approach, however, is computationally intensive for large data sets, d…
In recent studies the truncated Levy process (TLP) has been shown to be very promising for the modeling of financial dynamics. In contrast to the Levy process, the TLP has finite moments and can account for both the previously observed excess kurtosis at short timescales, along with the slow convergence to Gaussian at …
The paper analyzes how random perturbations affect RSVD and its applications.
New kernels capture both local and non-local interactions efficiently.
Study identifies and analyzes three types of errors in learning Fourier operators.
The method approximates stationary distributions of Markov models by truncating irrelevant states.
Faster diffusion-based models generate data with fewer steps.