MIC consistently estimates dependence in large datasets.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
High-dimensional, large-sample astrophysical databases of galaxy clusters, such as the Chandra Deep Field South COMBO-17 database, provide measurements on many variables for thousands of galaxies and a range of redshifts. Current understanding of galaxy formation and evolution rests sensitively on relationships between…
A measure of dependence is said to be equitable if it gives similar scores to equally noisy relationships of different types. Equitability is important in data exploration when the goal is to identify a relatively small set of strongest associations within a dataset as opposed to finding as many non-zero associations a…
Reshef et al. recently proposed a new statistical measure, the "maximal information coefficient" (MIC), for quantifying arbitrary dependencies between pairs of stochastic quantities. MIC is based on mutual information, a fundamental quantity in information theory that is widely understood to serve this need. MIC, howev…
Proposes an adversarial algorithm to learn unbiased representations via HGR coefficient.
Reshef & Reshef recently published a paper in which they present a method called the Maximal Information Coefficient (MIC) that can detect all forms of statistical dependence between pairs of variables as sample size goes to infinity. While this method has been praised by some, it has also been criticized for its lack …
A new method for multi-objective Bayesian optimization using entropy search and variational lower bound maximization.
The principle of absence of arbitrage opportunities allows obtaining the distribution of stock price fluctuations by maximizing its information entropy. This leads to a physical description of the underlying dynamics as a random walk characterized by a stochastic diffusion coefficient and constrained to a given value o…
Study S-shaped utility maximization with VaR constraint and unobservable drift.
Sharp bounds for eigenfunctions on product spaces restricted to submanifolds.
The maximal information coefficient (MIC) is a tool for finding the strongest pairwise relationships in a data set with many variables (Reshef et al., 2011). MIC is useful because it gives similar scores to equally noisy relationships of different types. This property, called {\em equitability}, is important for analyz…
New examples show high twisting doesn't guarantee open book maximality.
This paper studies the question of filtering and maximizing terminal wealth from expected utility in a partially information stochastic volatility models. The special features is that the only information available to the investor is the one generated by the asset prices, and the unobservable processes will be modeled …
A new model detects complex network communities using node attributes.
The paper studies the robust maximization of utility of terminal wealth in the diffusion financial market model. The underlying model consists with risky tradable asset, whose price is described by diffusion process with misspecified trend and volatility coefficients, and non-tradable asset with a known parameter. The …
We investigate the ergodic problem of growth-rate maximization under a class of risk constraints in the context of incomplete, Itô-process models of financial markets with random ergodic coefficients. Including {\em value-at-risk} (VaR), {\em tail-value-at-risk} (TVaR), and {\em limited expected loss} (LEL), these cons…
Sparse linear (or generalized linear) models combine a standard likelihood function with a sparse prior on the unknown coefficients. These priors can conveniently be expressed as a maximization over zero-mean Gaussians with different variance hyperparameters. Standard MAP estimation (Type I) involves maximizing over bo…
We consider the problem of recovering block-sparse signals whose structures are unknown \emph{a priori}. Block-sparse signals with nonzero coefficients occurring in clusters arise naturally in many practical scenarios. However, the knowledge of the block structure is usually unavailable in practice. In this paper, we d…
Framework uses deep learning and statistical models to solve PDEs with discontinuous coefficients.
Abstract: Determines thermoelastic coefficients from boundary data.
Modelling highly multi-modal data is a challenging problem in machine learning. Most algorithms are based on maximizing the likelihood, which corresponds to the M(oment)-projection of the data distribution to the model distribution. The M-projection forces the model to average over modes it cannot represent. In contras…
New algorithm solves utility maximization with deep learning for constrained problems.
Influence maximization (IM) is one of the most important problems in social network analysis. Its objective is to find a given number of seed nodes that maximize the spread of information through a social network. Since it is an NP-hard problem, many approximate/heuristic methods have been developed, and a number of th…
Investigates the use of Information Coefficient as a stock selection model performance measure.
The paper extends physics-based information maximization to complex bandit problems.
New quantum states capture more information, enabling advanced processing tasks.
New bandit algorithm maximizes information gain.
New index formula for hypoelliptic operators on manifolds.
In this work, we consider a manufactory process which can be described by a multiple-instance logistic regression model. In order to compute the maximum likelihood estimation of the unknown coefficient, an expectation-maximization algorithm is proposed, and the proposed modeling approach can be extended to identify the…
This paper concerns the recursive utility maximization problem. We assume that the coefficients of the wealth equation and the recursive utility are concave. Then some interesting and important cases with nonlinear and nonsmooth coefficients satisfy our assumption. After given an equivalent backward formulation of our …
Proposes a new method to enhance neural learning by maximizing information gain.
A new method estimates the learning coefficient using empirical loss.
The speed of convergence of the Expectation Maximization (EM) algorithm for Gaussian mixture model fitting is known to be dependent on the amount of overlap among the mixture components. In this paper, we study the impact of mixing coefficients on the convergence of EM. We show that when the mixture components exhibit …
We analyze the impact of the sampling interval on the estimation of Kramers-Moyal coefficients. We obtain the finite-time expressions of these coefficients for several standard processes. We also analyze extreme situations such as the independence and no-fluctuation limits that constitute useful references. Our results…
The effectiveness of utility-maximization techniques for portfolio management relies on our ability to estimate correctly the parameters of the dynamics of the underlying financial assets. In the setting of complete or incomplete financial markets, we investigate whether small perturbations of the market coefficient pr…
We study the global probability distribution of energy consumption per capita around the world using data from the U.S. Energy Information Administration (EIA) for 1980-2010. We find that the Lorenz curves have moved up during this time period, and the Gini coefficient G has decreased from 0.66 in 1980 to 0.55 in 2010,…
The study uses DCC for financial market analysis, revealing hidden correlations.
Improves functional linear regression with shape transfer learning.
On a real analytic 5-dimensional CR-generic submanifold M^5 in C^4 of codimension 3, hence of CR dimension 1, which enjoys the generically satisfied nondegeneracy condition that Lie brackets up to length 3 of T^{1,0}M generate CTM, a canonical Cartan connection is constructed after reduction to a certain partially expl…
The paper analyzes generalization of noisy, iterative algorithms using maximal leakage.
Proposes a framework to maximize mutual information in VAE models for better latent code representation.
Study finds Stäckel equivalence for superintegrable systems via invariant quadrics.
Generalizes underlap coefficient for multivariate group separation.
Kernel methods linked to feature subspaces and maximal correlation kernels.
Recently, crowdsourcing has emerged as an effective paradigm for human-powered large scale problem solving in various domains. However, task requester usually has a limited amount of budget, thus it is desirable to have a policy to wisely allocate the budget to achieve better quality. In this paper, we study the princi…
Information-maximization clustering learns a probabilistic classifier in an unsupervised manner so that mutual information between feature vectors and cluster assignments is maximized. A notable advantage of this approach is that it only involves continuous optimization of model parameters, which is substantially easie…
The paper provides formulas for Hadamard coefficients using Green's operators.
Investors pay for additional asset information based on utility maximization.