Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

111222333444 · Jun 202019922001200920182026
48 results for Bernoulli random variable

Rule-based classifiers quantify uncertainty using Bernoulli random variables.

problem Quantifying the uncertainty of precision estimates for rule-based text classifiers.
method Treat partitions of sub-strings as Bernoulli random variables, compare means using statistical tests, and combine classifiers using Dempster-Shafer theory.
result The approach can be used to combine binary classifiers into a multi-label classifier.

Active covariance estimation using random sub-sampling of variable subsets.

problem Estimating covariance matrices for partially observed random vectors.
method Unbiased covariance estimator under a model of partially observed variables and active learning framework.
result Derivation of error bounds revealing relations between sub-sampling probabilities and covariance matrix entries.

In this paper, we consider the multivariate Bernoulli distribution as a model to estimate the structure of graphs with binary nodes. This distribution is discussed in the framework of the exponential family, and its statistical properties regarding independence of the nodes are demonstrated. Importantly the model can e…

2012-06-08abs ↗pdf ↗

Paper compares largest claim amounts from two interdependent portfolios.

problem Comparing claim amounts from two sets of interdependent portfolios.
method Stochastic comparisons using dependent non-negative random variables and Bernoulli variables.
result Stochastic order results for largest claim amounts.

Paper compares smallest claim amounts from two interdependent portfolios.

problem Comparing smallest claim amounts from two sets of interdependent portfolios.
method Dependent non-negative random variables with survival copula, Bernoulli random variables, likelihood ratio order.
result Bounds for survival function of the smallest claim amount in a portfolio.

Paper proposes a new estimator for generic discrete distributions.

problem Estimating gradients for stochastic nodes in deep generative models.
method Generalized Gumbel-Softmax estimator using truncation, Gumbel-Softmax trick, and linear transformation.
result Efficacy and practical value demonstrated in synthetic examples and topic models.

The paper proves ML estimators are strongly consistent for identifying edge weights in BAR models.

problem Identifying edge weights in Bernoulli Autoregressive (BAR) models.
method Maximum Likelihood (ML) estimation for two variants of BAR models.
result ML estimators are strongly consistent for edge weight identification.

Characterizes symmetric Bernoulli distributions with minimal convex sums.

problem Understanding minimal dependence among Bernoulli random vectors.
method Geometric and algebraic representations of multivariate symmetric Bernoulli distributions.
result Characterizes extremal negative dependence and builds minimal dependence copulas.

We introduce an infectious default and recovery model for N obligors. Obligors are assumed to be exchangeable and their states are described by N Bernoulli random variables S_{i} (i=1,...,N). They are expressed by multiplying independent Bernoulli variables X_{i},Y_{ij},Y'_{ij}, and default and recovery infections are …

2006-10-31abs ↗pdf ↗

Improved regret bounds for DP-KLUCB and DP-IMED in Bernoulli bandits.

problem Minimizing regret in stochastic bandits under ε-global Differential Privacy.
method Developed DP versions of KLUCB and IMED, proving tighter lower bounds and matching upper bounds.
result DP-KLUCB and DP-IMED achieve asymptotically optimal regret under ε-global DP.

A new method for efficient nonlinear process monitoring using random Bernoulli features.

problem High computational demands and real-time responsiveness in online monitoring systems.
method Random Bernoulli principal component analysis to capture nonlinear patterns efficiently.
result The proposed methods offer excellent scalability and reduced computational complexity.

The structure of a Bayesian network encodes most of the information about the probability distribution of the data, which is uniquely identified given some general distributional assumptions. Therefore it's important to study the variability of its network structure, which can be used to compare the performance of diff…

2009-09-09abs ↗pdf ↗

The paper cleans label noise in supervised classification using Bernoulli sampling.

problem Label noise degrades supervised classifier performance.
method Proposes a label noise cleaning method based on Bernoulli random sampling.
result The method separates clean and noisy observations without prior label information.

We present two alternative ways to apply PAC-Bayesian analysis to sequences of dependent random variables. The first is based on a new lemma that enables to bound expectations of convex functions of certain dependent random variables by expectations of the same functions of independent Bernoulli random variables. This …

2011-05-12abs ↗pdf ↗

The structure of a Bayesian network includes a great deal of information about the probability distribution of the data, which is uniquely identified given some general distributional assumptions. Therefore it's important to study its variability, which can be used to compare the performance of different learning algor…

2010-05-23abs ↗pdf ↗

Dropout improves matrix factorization by acting as a low-rank regularizer.

problem Improving matrix factorization performance through regularization.
method Using Bernoulli random variables to drop columns of factors, demonstrating equivalence to a deterministic model with sum of squared Euclidean norms.
result Dropout achieves the global minimum of a convex approximation problem with squared nuclear norm regularization.

Optimal adaptive algorithm estimates coin mixture fractions with tight sample complexity bounds.

problem Estimating the fraction of positive coins in a mixture with unknown biases.
method Fully-adaptive algorithm with tight sample complexity bounds of Θ(ρ/ε²Δ² log(1/δ)).
result Upper and lower bounds of Θ(ρ/ε²Δ² log(1/δ)) samples for 1-δ probability of success.

Hybrid continuous-discrete models naturally represent many real-world applications in robotics, finance, and environmental engineering. Inference with large-scale models is challenging because relational structures deteriorate rapidly during inference with observations. The main contribution of this paper is an efficie…

2012-10-16abs ↗pdf ↗

Adaptive learning method identifies and corrects corrupted data.

problem Robust learning from corrupted training sets.
method Identifies corrupted and non-corrupted samples with latent Bernoulli variables, formulates as likelihood maximization with marginalized latent variables, solved via variational inference and Expectation-Maximization.
result Improves over state-of-the-art by automatically inferring corruption level with minimal overhead.

LCBM model improves image classification without human supervision.

problem Improving interpretability and generalization of unsupervised concept-based models.
method LCBM models concepts as random variables in a Bernoulli latent space, reducing the number of concepts without sacrificing performance.
result LCBM outperforms existing models in generalization and interpretability.

Develops methods to construct exchangeable sequences of random multisets.

problem Creating models for random multisets with unknown base measures.
method Uses exchangeable sequences of point processes and conditional-i.i.d. negative binomial processes.
result Provides constructions for negative binomial processes with random base measures.

Proposes a new method for feature selection in non-linear functions.

problem Feature selection for non-linear functions in high-dimensional data.
method Continuous relaxation of Bernoulli distributions to learn feature selection indicators via gradient descent.
result Demonstrates the effectiveness of the approach on synthetic and real-life applications.

We introduce BAR processes for modeling binary interactions and show they mix rapidly.

problem Modeling binary interactions in various graphical structures.
method Introduce Bernoulli Autoregressive Processes (BAR) with autoregressive dynamics and efficient structure learning.
result BAR processes mix rapidly with a mixing time of O(logp)O(\log p).

Paper justifies ST estimator using pWGF and proposes an improved variant.

problem Theoretical justification for ST estimator for discrete variables.
method Interpreted ST as pWGF simulation and proposed an improved estimator.
result Established theoretical foundation for ST estimator and improved variant.

Study explores geometric structure and prior for beta-logistic distribution.

problem Understanding the geometric structure and prior distributions of the beta-logistic distribution.
method Exploring dual geometric structure and uncovering α\alpha-parallel prior.
result The beta-logistic distribution admits an α\alpha-parallel prior for any real number α\alpha.

Optimizing betting frequency in dynamic games with Kelly criterion.

problem Finding the optimal betting frequency in a dynamic game setting.
method Using Kelly's expected logarithmic growth criterion, the study analyzes the performance of high-frequency and low-frequency bettors.
result The optimal performance gn* changes with n, and the high-frequency case does not always lead to the best performance.

Optimal transport is #P-hard when components are independent, even with approximate solutions.

problem Computational complexity of optimal transport with independent marginals.
method Proved #P-hardness and developed a pseudo-polynomial time approximation algorithm.
result Optimal transport is #P-hard even with independent components and approximate solutions.

This research explores various sampling methods and probability distributions for hard alignment in sequence-to-sequence TTS synthesis.

problem Improving alignment accuracy in sequence-to-sequence text-to-speech synthesis.
method Investigated various sampling methods (greedy, beam, random) and probability distributions (Bernoulli, Concrete) for hard alignment.
result Deterministic search is more preferable than stochastic search for natural alignment transition.

We study in detail and explicitly solve the version of Kyle's model introduced in a specific case in \cite{BB}, where the trading horizon is given by an exponentially distributed random time. The first part of the paper is devoted to the analysis of time-homogeneous equilibria using tools from the theory of one-dimensi…

2016-03-29abs ↗pdf ↗

This paper describes a recursive estimation procedure for multivariate binary densities (probability distributions of vectors of Bernoulli random variables) using orthogonal expansions. For dd covariates, there are 2d2^d basis coefficients to estimate, which renders conventional approaches computationally prohibitive …

2011-12-07abs ↗pdf ↗

Paper analyzes conditions for clustering BMMs with unknown clusters.

problem Clustering Bernoulli Mixture Models (BMMs) with unknown number of clusters.
method Theoretical analysis of sample complexity and dimensionality for PAC-clusterability.
result First non-asymptotic bounds on sample complexity for learning or clustering BMMs.

The beta-Bernoulli process provides a Bayesian nonparametric prior for models involving collections of binary-valued features. A draw from the beta process yields an infinite collection of probabilities in the unit interval, and a draw from the Bernoulli process turns these into binary-valued features. Recent work has …

2011-06-03abs ↗pdf ↗

Probabilistic learning for binary classification with categorical variables.

problem Binary classification with categorical covariates.
method Probabilistic analysis and two algorithms for learning boolean functions.
result Effective learning of boolean functions from binary data.

A new method prunes neural networks efficiently without losing effectiveness.

problem Efficient pruning of neural networks without sacrificing performance.
method Deterministic approximation of binary gates and L0L_0 regularization.
result Pruning neural networks significantly without loss in effectiveness.