Paper proposes approximate Stein classes for efficient truncated density estimation.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Truncated densities are probability density functions defined on truncated domains. They share the same parametric form with their non-truncated counterparts up to a normalizing constant. Since the computation of their normalizing constants is usually infeasible, Maximum Likelihood Estimation cannot be easily applied t…
Stochastic gradient descent (SGD) is commonly used for optimization in large-scale machine learning problems. Langford et al. (2009) introduce a sparse online learning method to induce sparsity via truncated gradient. With high-dimensional sparse data, however, the method suffers from slow convergence and high variance…
Derives equations for deep learning biases and weights, showing data complexity reduction.
New method for constructing truncated vine copulas.
UDN adapts depth to data complexity, outperforming standard neural networks.
We present a probabilistic framework for nonlinearities, based on doubly truncated Gaussian distributions. By setting the truncation points appropriately, we are able to generate various types of nonlinearities within a unified framework, including sigmoid, tanh and ReLU, the most commonly used nonlinearities in neural…
We present trellis networks, a new architecture for sequence modeling. On the one hand, a trellis network is a temporal convolutional network with special structure, characterized by weight tying across depth and direct injection of the input into deep layers. On the other hand, we show that truncated recurrent network…
In the past years, Deep convolution neural network has achieved great success in many artificial intelligence applications. However, its enormous model size and massive computation cost have become the main obstacle for deployment of such powerful algorithm in the low power and resource-limited mobile systems. As the c…
A new method balances model quality and Byzantine robustness in Federated Learning.
New method improves sampling from logconcave distributions truncated on polytopes.
New algorithm improves hypergraph clustering for unbalanced communities.
Constructs classifiers for neural networks with specific data configurations.
COS method convergence conditions expanded for heavy-tailed distributions.
Deep learning has shown that learned functions can dramatically outperform hand-designed functions on perceptual tasks. Analogously, this suggests that learned optimizers may similarly outperform current hand-designed optimizers, especially for specific problems. However, learned optimizers are notoriously difficult to…
For any symmetric collection of natural numbers h^{p,q} with p+q=k, we construct a smooth complex projective variety whose weight k Hodge structure has these Hodge numbers; if k=2m is even, then we have to impose that h^{m,m} is bigger than some quadratic bound in m. Combining these results for different weights, we so…
Paper tackles unbounded density ratio estimation for covariate shift adaptation.
The problem of an arbitrary truncated Levy flight description using the method of cumulant approach has been solved. The set of cumulants of the truncated Levy distribution given the assumption of arbitrary truncation has been found. The influence of truncation shape on the truncated Levy flight properties in the Gauss…
A new method reduces variance in PG methods for RL, improving efficiency and convergence.
Efficiently estimate Boolean product distribution parameters from truncated samples.
In the paper "On Truncated Variation of Brownian Motion with Drift" (Bull. Pol. Acad. Sci. Math. 56 (2008), no.4, 267 - 281) we defined truncated variation of Brownian motion with drift, where is a standard Brownian motion. Truncated variation differs from regular variation by neglect…
Dropout-based regularization methods can be regarded as injecting random noise with pre-defined magnitude to different parts of the neural network during training. It was recently shown that Bayesian dropout procedure not only improves generalization but also leads to extremely sparse neural architectures by automatica…
Efficiently estimates covariance for sub-Weibull vectors with sub-Gaussian rate.
Optimal algorithm learns Gaussian under halfspace truncation with minimal samples.
Non-negative matrix factorization (NMF) minimizes the Euclidean distance between the data matrix and its low rank approximation, and it fails when applied to corrupted data because the loss function is sensitive to outliers. In this paper, we propose a Truncated CauchyNMF loss that handle outliers by truncating large e…
Paper defines new risk measures for elliptical distributions.
Improved regret bounds for adversarial linear contextual bandits.
Paper develops methods for analyzing forms with synchronized singularities.
New DP framework using data truncation for efficient estimation.
Unified framework for mean testing under truncation bias.
Score matching method improves density estimation for truncated data on manifolds.
New sampling method for Heston model reduces complexity.
Recommender systems are widely used to recommend the most appealing items to users. These recommendations can be generated by applying collaborative filtering methods. The low-rank matrix completion method is the state-of-the-art collaborative filtering method. In this work, we show that the skewed distribution of rati…
It is well known that Convolutional Neural Networks (CNNs) have significant redundancy in their filter weights. Various methods have been proposed in the literature to compress trained CNNs. These include techniques like pruning weights, filter quantization and representing filters in terms of a basis functions. Our ap…
Truncated backpropagation through time (TBPTT) is a popular method for learning in recurrent neural networks (RNNs) that saves computation and memory at the cost of bias by truncating backpropagation after a fixed number of lags. In practice, choosing the optimal truncation length is difficult: TBPTT will not converge …
The method approximates stationary distributions of Markov models by truncating irrelevant states.
Paper tackles overestimation bias in continuous control, improving performance by 25%.
Marginal structural models (MSMs) estimate the causal effect of a time-varying treatment in the presence of time-dependent confounding via weighted regression. The standard approach of using inverse probability of treatment weighting (IPTW) can lead to high-variance estimates due to extreme weights and be sensitive to …
We consider an appoximation of a catenoid constructed from "odd" truncated cones that maintains minimality in a certain sense. Thorough this procedure, we obtain a discrete curve approximating a catenary by exploiting the fact that it is the function that generates a catenoid. In this investigation, the theory of the G…
Estimates domain truncation error for option pricing PDEs.
Machine learning approximates Calabi-Yau Hodge numbers from weight systems.
Choppy optimizes ranked list truncation using Transformer architecture.
New COS method formula improves option pricing accuracy.
The generalized correlation approach, which has been successfully used in statistical radio physics to describe non-Gaussian random processes, is proposed to describe stochastic financial processes. The generalized correlation approach has been used to describe a non-Gaussian random walk with independent, identically d…
Lower bound shows super-polynomial gap for estimating truncated Gaussian means.
As in standard linear regression, in truncated linear regression, we are given access to observations whose dependent variable equals , where is some fixed unknown vector of interest and is independent noise; except we are only given an observation if its dep…
The paper analyzes and mitigates biases in scalable Gaussian Process methods.
We show that generalised geometry gives a unified description of maximally supersymmetric consistent truncations of ten- and eleven-dimensional supergravity. In all cases the reduction manifold admits a "generalised parallelisation" with a frame algebra with constant coefficients. The consistent truncation then arises …