Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

3672108144 · Jun 202019922001200920172026
48 results for weak correlations

Weak correlations explain linear dynamics in deep learning models.

problem Understanding the linear structure in gradient-based learning algorithms.
method Characterization of weak correlations between derivatives and parameters.
result Weak correlations are the underlying principle for linearization in deep learning models.

Weak diffusion priors can still perform well in inverse problems.

problem Using mismatched or low-fidelity diffusion priors in inverse problems.
method Extensive experiments and theoretical analysis combining Bayesian-consistency theory and local-correlation analysis.
result Weak priors succeed when measurements are highly informative, and they fail in other regimes.

Since manually labeling training data is slow and expensive, recent industrial and scientific research efforts have turned to weaker or noisier forms of supervision sources. However, existing weak supervision approaches fail to model multi-resolution sources for sequential data, like video, that can assign labels to in…

2019-10-21abs ↗pdf ↗

We introduce a mean-reverting SDE whose solution is naturally defined on the space of correlation matrices. This SDE can be seen as an extension of the well-known Wright-Fisher diffusion. We provide conditions that ensure weak and strong uniqueness of the SDE, and describe its ergodic limit. We also shed light on a use…

2011-08-26abs ↗pdf ↗

Study on W2S generalization with spurious correlations, proposing remedies.

problem Understanding and improving W2S generalization with spurious correlations.
method Theoretical analysis and algorithmic remedies for W2S fine-tuning.
result W2S always happens with sufficient pseudolabels when group fractions match, but may fail otherwise.

We analyze the Standard & Poor's 500 stock market index from the last 22 years. The probability density function of price returns exhibits two well-distinguished regimes with self-similar structure: the first one displays strong super-diffusion together with short-time correlations, and the second one corresponds to we…

2019-02-11abs ↗pdf ↗

CAKD framework optimizes knowledge transfer by focusing on influential components of distillation.

problem Balancing and optimizing knowledge transfer in distillation models.
method Decouple KL divergence into BCD, SCD, and WCD; prioritize influential components.
result CAKD framework consistently outperforms baseline across diverse models and datasets.

We analyze correlations among stock returns via a series of widely adopted parameters which we refer to as explanatory variables. We subsequently exploit the results to propose a long only quantitative adaptive technique to construct a profitable portfolio of assets which exhibits minor drawdowns and higher recoveries …

2018-06-13abs ↗pdf ↗

WeLa-VAE learns interpretable disentangled representations with weak supervision.

problem Learning disentangled representations without strong supervision.
method Variational inference framework with shared latent variables and modified variational lower bound.
result WeLa-VAE learns alternative disentangled representations (polar) from weak labels (distance and angle) without refined supervision.

We consider the problem of providing nonparametric confidence guarantees for undirected graphs under weak assumptions. In particular, we do not assume sparsity, incoherence or Normality. We allow the dimension DD to increase with the sample size nn. First, we prove lower bounds that show that if we want accurate infe…

2013-09-26abs ↗pdf ↗

New model shows weak teachers can help strong students learn even with imperfect labels.

problem Improving strong student's performance with weak teacher's imperfect pseudolabels.
method Stylized overparameterized spiked covariance model with Gaussian covariates, proving two phases of generalization.
result Provable successful and random guessing phases of strong student's generalization.

To investigate the universal structure of interactions in financial dynamics, we analyze the cross-correlation matrix C of price returns of the Chinese stock market, in comparison with those of the American and Indian stock markets. As an important emerging market, the Chinese market exhibits much stronger correlations…

2012-02-02abs ↗pdf ↗

We examine several recently suggested methods for the detection of long-range correlations in data series based on similar ideas as the well-established Detrended Fluctuation Analysis (DFA). In particular, we present a detailed comparison between the regular DFA and two recently suggested methods: the Centered Moving A…

2008-04-25abs ↗pdf ↗

Study shows disentanglement models learn correlations from data, impacting fairness.

problem Disentanglement models learn correlations in real-world data, affecting downstream applications.
method Empirical study on 4260 models, analyzing correlations in latent representations.
result Systematically induced correlations are learned by disentanglement models, impacting fairness.

Study improves weak error estimates for rough volatility models.

problem Efficient numerical schemes for non-Markovian stochastic processes with rough volatility.
method Analyzes weak rates for a class of stochastic processes with rough stochastic volatility.
result Weak rate is of order min{3H+0.5, 1} for a large class of test functions.

The weak variance-alpha-gamma process is a multivariate Lévy process constructed by weakly subordinating Brownian motion, possibly with correlated components with an alpha-gamma subordinator. It generalises the variance-alpha-gamma process of Semeraro constructed by traditional subordination. We compare three calibrati…

2018-01-26abs ↗pdf ↗

Study on error rates for approximating rough volatility models.

problem Simulation of rough volatility models with fractional Brownian motion.
method Analysis of weak error rates for numerical schemes, focusing on fBm and cubic test functions.
result Convergence rates for approximations are (3H+12)1(3H+ \frac{1}{2}) \wedge 1 for exact left-point discretization and H+12H+\frac{1}{2} for hybrid schemes.

New tuning rules for Metropolis algorithms derived from Bayesian large-sample asymptotics.

problem Optimal scaling in random-walk Metropolis algorithms under realistic assumptions.
method Large-sample asymptotics to derive weak convergence results and tuning guidelines.
result Tuning guidelines consistent with previous ones when target density is product form, accounting for correlation structure.

In the presence of weak overall correlation, it may be useful to investigate if the correlation is significantly and substantially more pronounced over a subpopulation. Two different testing procedures are compared. Both are based on the rankings of the values of two variables from a data set with a large number n of o…

2015-04-21abs ↗pdf ↗

Study shows superdiffusive behavior in geodesic flows on curved surfaces.

problem Understanding the statistical behavior of geodesic flows on curved surfaces.
method Proved nonstandard central limit theorem with superdiffusive normalisation (tlogt)1/2(t\log t)^{1/2} for geodesic flows on nonpositively curved surfaces.
result Geodesic flows exhibit superdiffusive behavior with correlations decaying at rate t1t^{-1}.

During times of extreme market turmoil, it is acknowledged that there is a tendency towards "flight to safety". A strong (weak) safe haven is defined as an asset that has a significant positive (negative) return in periods where another asset is in distress, while hedge has to be negatively correlated (uncorrelated) on…

2017-03-01abs ↗pdf ↗

This paper treats the problem of screening for variables with high correlations in high dimensional data in which there can be many fewer samples than variables. We focus on threshold-based correlation screening methods for three related applications: screening for variables with large correlations within a single trea…

2011-02-06abs ↗pdf ↗

We analyze the fluctuation of the loss from default around its large portfolio limit in a class of reduced-form models of correlated firm-by-firm default timing. We prove a weak convergence result for the fluctuation process and use it for developing a conditionally Gaussian approximation to the loss distribution. Nume…

2013-04-04abs ↗pdf ↗

As machine learning models continue to increase in complexity, collecting large hand-labeled training sets has become one of the biggest roadblocks in practice. Instead, weaker forms of supervision that provide noisier but cheaper labels are often used. However, these weak supervision sources have diverse and unknown a…

2018-10-05abs ↗pdf ↗

In high-dimensional data, structured noise caused by observed and unobserved factors affecting multiple target variables simultaneously, imposes a serious challenge for modeling, by masking the often weak signal. Therefore, (1) explaining away the structured noise in multiple-output regression is of paramount importanc…

2014-10-27abs ↗pdf ↗

In the last few years, many different performance measures have been introduced to overcome the weakness of the most natural metric, the Accuracy. Among them, Matthews Correlation Coefficient has recently gained popularity among researchers not only in machine learning but also in several application fields such as bio…

2010-08-17abs ↗pdf ↗

Develops a new cluster validity index to find multiple optimal cluster numbers.

problem Finding the optimal number of clusters in real-world data with varying densities, sizes, and shapes.
method A new correlation-based cluster validity index that yields multiple local peaks.
result The new index finds multiple optimal cluster numbers in various scenarios.

Model simulates correlation emergence in two coupled limit order books.

problem Modeling correlation emergence in coupled limit order books.
method Simulated two coupled diffusive limit order books using random walks in the fluid limit, with trader interactions.
result Demonstrated the recovery of an Epps effect from the model.

This short note suggests a heuristic method for detecting the dependence of random time series that can be used in the case when this dependence is relatively weak and such that the traditional methods are not effective. The method requires to compare some special functionals on the sample characteristic functions with…

2010-10-13abs ↗pdf ↗

How can graph theory be applied to investing in the stock market? The answer may help investors realize the true risks of their investments, help prevent recessions like that of 2008, and increase financial literacy amongst students. Using several original Python programs, we take a correlation matrix with correlations…

2019-02-02abs ↗pdf ↗

FABLE incorporates instance features into PWS label models for improved performance.

problem Lack of instance features in existing label models limits their performance.
method FABLE uses a mixture of Bayesian label models and a Gaussian Process classifier to incorporate instance features.
result FABLE achieves the highest averaged performance across nine baselines on benchmark datasets.

Study shows how a strong model can learn a task's feature while retaining other capabilities.

problem How to align superhuman AI systems using weak-to-strong generalization.
method Two-layer neural networks, reward-model learning, multi-step SGD, feature learning.
result The strong model efficiently learns task features while retaining general capabilities.

Study reveals efficient recovery of multi-modal signals via Bayesian methods and sequential learning.

problem Recovering multiple high-dimensional signals from correlated modalities.
method Bayesian Approximate Message Passing and Sequential Curriculum Learning.
result Sequential learning strategy optimally recovers weak signals in multi-modal settings.

Canonical Correlation Analysis (CCA) is widely used for multimodal data analysis and, more recently, for discriminative tasks such as multi-view learning; however, it makes no use of class labels. Recent CCA methods have started to address this weakness but are limited in that they do not simultaneously optimize the CC…

2019-07-17abs ↗pdf ↗

A persistent challenge in practical classification tasks is that labeled training sets are not always available. In particle physics, this challenge is surmounted by the use of simulations. These simulations accurately reproduce most features of data, but cannot be trusted to capture all of the complex correlations exp…

2018-01-30abs ↗pdf ↗

In phase retrieval we want to recover an unknown signal xCd\boldsymbol x\in\mathbb C^d from nn quadratic measurements of the form yi=ai,x2+wiy_i = |\langle{\boldsymbol a}_i,{\boldsymbol x}\rangle|^2+w_i where aiCd\boldsymbol a_i\in \mathbb C^d are known sensing vectors and wiw_i is measurement noise. We ask the following weak rec…

2017-08-20abs ↗pdf ↗

New method handles correlated responses and interaction effects in multi-response regression.

problem Handling correlated responses and interaction effects in multi-response regression.
method MADMMplasso, an ADMM-based approach for multi-response regression with overlapping groups and interaction effects.
result The proposed method outperforms in prediction and variable selection for correlated responses and interaction effects.

We consider support recovery in the quadratic logistic regression setting - where the target depends on both p linear terms xix_i and up to p2p^2 quadratic terms xixjx_i x_j. Quadratic terms enable prediction/modeling of higher-order effects between features and the target, but when incorporated naively may involve solvi…

2017-03-08abs ↗pdf ↗

Optimal spectral estimators and AMP combine for efficient weak recovery in orthogonally invariant GLMs.

problem Parameter estimation from generalized linear models with complex correlation structures.
method Spectral initialization and approximate message passing (AMP) algorithm.
result Established rigorous performance guarantees for spectral initialization and AMP.

In this study we consider relations between companies in Poland taking into account common branches they belong to. It is clear that companies belonging to the same branch compete for similar customers, so the market induces correlations between them. On the other hand two branches can be related by companies acting in…

2006-11-15abs ↗pdf ↗