Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

4182123164 · Jun 202019922001200920172026
48 results for correlation length

The paper studies the correlation of Hilbert lengths for convex projective surfaces.

problem Understanding the correlation of Hilbert lengths for convex projective surfaces.
method Asymptotic formula for free homotopy classes with renormalized Hilbert length.
result The correlation number is not uniformly bounded away from zero but can be larger than a uniform strictly positive constant.

This study examines how sequential correlations affect in-context learning in sequence models.

problem Understanding how in-context learning works with sequentially correlated data.
method Extended linear regression model to sequentially correlated data, tested on transformer architectures.
result Sequential correlations alter the effective context length and attention architecture effectiveness.

Study shows LLC correlates with neural network compressibility.

problem Evaluating limits of neural network compression.
method Extended minimum description length principle using singular learning theory.
result Complexity estimates based on LLC are linearly correlated with compressibility.

The paper describes correlations of spectra for higher rank Anosov representations.

problem Understanding correlations of spectra for Anosov representations of higher rank groups.
method Relates correlation problem to counting projections in truncated hypertubes.
result Extends previous work on rank one representations to higher rank.

LaRT models LLMs' response accuracy and CoT length to evaluate reasoning ability and speed.

problem Valid evaluation of Large Language Models (LLMs) via response accuracy and chain-of-thought length.
method Introduces Latency-Response Theory (LaRT) to jointly model response accuracy and CoT length using latent ability and latent speed.
result LaRT yields higher estimation accuracy and shorter confidence intervals for latent traits compared to IRT.

We examine Deep Canonically Correlated LSTMs as a way to learn nonlinear transformations of variable length sequences and embed them into a correlated, fixed dimensional space. We use LSTMs to transform multi-view time-series data non-linearly while learning temporal relationships within the data. We then perform corre…

2018-01-16abs ↗pdf ↗

New insights show embedding lengths correlate with semantic properties.

problem Contrastive embedding norms ignore embedding magnitudes but correlate with semantic properties.
method Formal theoretical framework and analysis of optimization dynamics.
result Embedding lengths encode semantic information as a byproduct of training.

We study finite sample properties of estimators of power-law cross-correlations -- detrended cross-correlation analysis (DCCA), height cross-correlation analysis (HXA) and detrending moving-average cross-correlation analysis (DMCA) -- with a special focus on short-term memory bias as well as power-law coherency. Presen…

2014-09-24abs ↗pdf ↗

We discuss algorithms for estimating the Shannon entropy h of finite symbol sequences with long range correlations. In particular, we consider algorithms which estimate h from the code lengths produced by some compression algorithm. Our interest is in describing their convergence with sequence length, assuming no limit…

2002-03-21abs ↗pdf ↗

ESS improves MCMC efficiency for correlated & multimodal distributions.

problem Slice Sampling's sensitivity to initial length scale and difficulty with correlated distributions.
method Adaptive tuning and parallel walkers for efficient sampling.
result ESS improves efficiency by more than an order of magnitude on correlated distributions.

Research examines correlations of complex logarithms of lattice points, showing level repulsion and Poissonian behavior.

problem Analyzing correlations of complex logarithms of lattice points.
method Proving existence of pair correlation functions and examining behavior at various scalings.
result Level repulsion observed at linear scaling, Poissonian behavior at sublinear scalings.

The study examines correlations of logarithms of integers at different scalings.

problem Analyzing pair correlations of logarithms of integers at various scalings.
method Examined correlations of logarithms of positive integers at different scalings, proving the existence of pair correlation functions.
result Level repulsion at linear scaling, total loss of mass at superlinear scalings, and Poissonian behavior at sublinear scalings.

Preformer improves Transformer for long-term time series forecasting.

problem Transformer's quadratic complexity and lack of context-awareness for long-term forecasting.
method Introduces Multi-Scale Segment-Correlation mechanism for efficient time series segmentation and context-aware attention.
result Preformer outperforms other Transformer-based methods in long-term time series forecasting.

This study uses local Gaussian correlation to analyze stock return tails, revealing more sensitive network properties.

problem Misleading results from Pearson correlation in financial networks.
method Local Gaussian correlation coefficient for capturing nonlinear dependence and heavy-tailed distributions.
result Local Gaussian correlation network among negative tails is more sensitive to stock market risks.

Fine-tuning improves information conveyance in language models by reorganizing uncertainty into more informative sequences.

problem Uncertainty reduction in large language models through fine-tuning is not fully understood, especially regarding output length.
method Proposed Canopy Entropy (CE\mathrm{CE}^\star) to measure uncertainty in both output length and sequence, capturing total Shannon entropy.
result Fine-tuned models exhibit stronger positive correlation between entropy rate and semantic diversity, indicating more informative and semantically meaningful generations.

Large bundles of myelinated axons, called white matter, anatomically connect disparate brain regions together and compose the structural core of the human connectome. We recently proposed a method of measuring the local integrity along the length of each white matter fascicle, termed the local connectome. If communicat…

2018-04-22abs ↗pdf ↗

We analyse the structure of the distribution of eigenvalues of the stock market correlation matrix with increasing length of the time series representing the price changes. We use 100 highly-capitalized stocks from the American market and relate result to the corresponding ensemble of Wishart random matrices. It turns …

2005-05-10abs ↗pdf ↗

VAE improves MCMC efficiency by generating diverse prior proposals.

problem Inefficient MCMC methods in Bayesian inverse problems, especially subsurface flow modeling.
method Uses Variational Autoencoder (VAE) to generate broader-spectrum prior proposals.
result VAE achieves comparable accuracy to Karhunen-Loève Expansion (KLE) and outperforms it when correlation length is unknown.

We examine the performance of six estimators of the power-law cross-correlations -- the detrended cross-correlation analysis, the detrending moving-average cross-correlation analysis, the height cross-correlation analysis, the averaged periodogram estimator, the cross-periodogram estimator and the local cross-Whittle e…

2016-02-17abs ↗pdf ↗

Study validates Lillo-Mike-Farmer model predicting financial market long-range correlations.

problem Quantifying long-range correlations in financial markets.
method Analyzed nine years of market data to classify traders as order-splitting or random, measured metaorder-length distributions, and compared to LMF model predictions.
result Agreement between LMF model predictions and actual data, validating the model.

Study correlations of spectral lengths and displacements in higher rank groups.

problem Analyzing correlations of spectral lengths and displacements in higher rank groups.
method Study Jordan and Cartan projections in tubes of Anosov subgroups of semisimple real algebraic groups.
result Prove existence of δ_ρ(\mathsf{v}) such that correlations of spectral lengths and displacements follow specific exponential growth patterns.

Shorter adversarial prompts help protect LLMs from jailbreak attacks.

problem Protecting large language models from jailbreak attacks with long adversarial suffixes.
method Adversarial training on shorter adversarial suffixes to defend against longer adversarial suffixes.
result Aligning LLMs on shorter adversarial suffixes can effectively defend against jailbreak attacks with longer suffixes.

Quantum field theory connects deep neural networks to criticality.

problem Understanding the criticality and training dynamics of deep neural networks.
method Constructing quantum field theory for deep neural networks, computing corrections to correlation functions.
result Found precise analogy with O(N)O(N) vector model, providing corrections to correlation length.
Agents Play Mix-gamephysics.soc-ph

In mix-game which is an extension of minority game, there are two groups of agents; group1 plays the majority game, but the group2 plays the minority game. This paper studies the change of the average winnings of agents and volatilities vs. the change of mixture of agents in mix-game model. It finds that the correlatio…

2005-05-17abs ↗pdf ↗

Standardizes weighted ranking correlation coefficients to maintain zero expected value.

problem Measuring correlation between weighted rankings of items.
method Develops a standardization function g(·) that transforms coefficients to zero expected value under randomness.
result A general standardization function g(Γ) that preserves the domain [-1,1] and reduces to the identity for coefficients already satisfying zero-expected-value property.

VBS improves sampling efficiency in cosmological data analysis.

problem High dimensionality of cosmological parameter space makes sampling computationally challenging.
method Developed a hybrid scheme combining variational self-boosted sampling with Hamiltonian Monte Carlo.
result VBS generates better quality samples and reduces auto-correlation length by a factor of 10-50.

Paper tackles variable-length, incomplete wearable sensor data to improve personalized insights.

problem Variable-length and incomplete time series data from wearable sensors.
method HeartSpace integrates a time series encoding module and pattern aggregation network, along with a Siamese-triplet network for representation learning.
result Empirical evaluation shows significant performance gains in personality prediction, demographics inference, and user identification.

Deep learning reveals lagged correlations in stock markets, showing accuracy decreases with shorter prediction horizons.

problem Capturing non-linear interactions in financial prediction problems using large-scale datasets.
method Applying deep learning to econometrically constructed gradients to learn and exploit lagged correlations among S&P 500 stocks.
result Model accuracies decrease with shorter prediction horizons, but remain significant in both stable and volatile markets.

Algorithm minimizes regret and converges to equilibria in Markov games.

problem Regret minimization and convergence to equilibria in general-sum Markov games under adversarial opponents.
method Decentralized algorithm that uses policy optimization and controls path length to achieve sublinear regret.
result Sublinear regret guarantees for convergence to correlated equilibrium in Markov games.

Proposes a new model for EHR data using time-dependent Gaussian processes.

problem Joint modeling of multiple clinical variables over time.
method Multivariate nonstationary Gaussian processes with time-varying parameters and posterior inference via HMC.
result The proposed model outperforms stationary models and reveals latent correlations predictive of patient risk.

One of the ubiquitous representation of long DNA sequence is dividing it into shorter k-mer components. Unfortunately, the straightforward vector encoding of k-mer as a one-hot vector is vulnerable to the curse of dimensionality. Worse yet, the distance between any pair of one-hot vectors is equidistant. This is partic…

2017-01-23abs ↗pdf ↗