Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

6.3%12.5%18.8%25.0% · Oct 199319922001200920172026
48 results for Persistent Contrastive Divergence

Learning algorithms for energy based Boltzmann architectures that rely on gradient descent are in general computationally prohibitive, typically due to the exponential number of terms involved in computing the partition function. In this way one has to resort to approximation schemes for the evaluation of the gradient.…

2018-01-08abs ↗pdf ↗

New method tackles incomplete data in RBM inverse Ising problems.

problem Computing data and model expectations in inverse Ising problems with missing observations.
method Combines mean-field approximation, persistent contrastive divergence, and spatial Monte Carlo integration.
result Effective and accurate tuning of model parameters compared to conventional methods.

Persistently trained EBMs generate images and estimate complex densities.

problem Challenges in ML learning for energy-based models, especially non-convergence of MCMC.
method Introduce diffusion data, learn a joint EBM through persistent training with enhanced sampling.
result First simultaneous achievement of stability, post-training image generation, and superior out-of-distribution detection for image data.

CRBMs improve financial regime detection with PCD and free energy analysis.

problem Detecting systemic risk regimes in financial time series.
method Extended RBM to CRBM with autoregressive conditioning and PCD. Decomposed free energy into magnitude and correlation components.
result CRBM's free energy metric distinguishes between magnitude shocks and market regimes.

RényiCL uses Rényi divergence for robust contrastive learning with stronger data augmentations.

problem Learning useful representations from multiple data views with hard augmentations.
method RényiCL employs Rényi divergence for contrastive learning, using a novel variational objective to manage hard negative sampling.
result RényiCL achieves better performance with stronger augmentations compared to other methods.

This paper studies the problem of parameter learning in probabilistic graphical models having latent variables, where the standard approach is the expectation maximization algorithm alternating expectation (E) and maximization (M) steps. However, both E and M steps are computationally intractable for high dimensional d…

2016-05-26abs ↗pdf ↗

Theoretical proof shows COMs are a type of contrastive divergence model with improved sampling.

problem Improving sampling quality in offline model-based optimization.
method Showed COMs are contrastive divergence models, proposed Langevin MCMC sampler, and decoupled model.
result Improved sampling quality achieved by decoupling model and using Langevin MCMC.

This paper proposes a more efficient training method for energy-based models.

problem The computational burden and validity trade-off in Contrastive Divergence training.
method Introducing Diffusion Contrastive Divergence (DCD) to replace Langevin dynamics with diffusion processes.
result The proposed DCDs are more computationally efficient and handle gradient terms better than Contrastive Divergence.

We develop a method to combine Markov chain Monte Carlo (MCMC) and variational inference (VI), leveraging the advantages of both inference approaches. Specifically, we improve the variational distribution by running a few MCMC steps. To make inference tractable, we introduce the variational contrastive divergence (VCD)…

2019-05-10abs ↗pdf ↗

The study reveals distinct patterns in retail investors' holding periods affecting stock returns.

problem Understanding the impact of retail investors' investment horizons on stock returns.
method Using self-reported holding periods from StockTwits, the study categorizes retail investors into long-horizon and short-horizon groups and analyzes their return patterns.
result Long-horizon retail investors exhibit underreaction to earnings announcements, while short-horizon investors show overreaction.

New statistical theory explains contrastive learning effectiveness.

problem Understanding why contrastive learning works well for representation extraction.
method Developed a new theoretical framework based on approximate sufficient statistics.
result Near-sufficient encoders derived from contrastive learning can be adapted for downstream tasks.

Simplicial persistence measures financial market dynamics, revealing long-term structure evolution.

problem Understanding the long-term structure evolution of financial markets.
method Simplicial persistence, null models, TMFG filtering, thresholding, generative process analysis.
result More liquid markets exhibit slower persistence decay, suggesting higher fragility to systemic shocks.

Paper proposes f-DPG for aligning language models with preferences.

problem Aligning language models with user preferences.
method Uses f-divergence to approximate target distributions and minimizes a forward KL from it using DPG.
result Jensen-Shannon divergence often outperforms forward KL divergence, leading to significant improvements.

Study shows similarities and differences in crypto and equity dynamics during pandemic.

problem Comparing cryptocurrency and equity market dynamics during the pandemic.
method New methodologies applied to study cryptocurrency and equity market dynamics, including recently introduced methods for trajectory and anomaly analysis.
result Cryptocurrencies exhibit stronger collective dynamics and correlation, while equities show greater persistence in anomalies over time.

The paper explores how Finsler manifolds differ from Riemannian ones in functional inequalities.

problem Analytic phenomena on Finsler manifolds differ from Riemannian ones.
method Comparative analysis of Finsler and Riemannian manifolds, focusing on Sobolev spaces, Hardy inequalities, and uncertainty principles.
result Functional inequalities (Hardy, uncertainty) break down on Finsler Cartan-Hadamard manifolds, while Caffarelli-Kohn-Nirenberg inequality exhibits a sharp threshold.

Self-supervised and supervised methods learn similar intermediate visual representations but diverge in final layers.

problem Comparing self-supervised and supervised methods for visual learning.
method Comparison of contrastive self-supervised and supervised methods on simple image data.
result Contrastive and supervised methods learn similar intermediate representations but diverge in final layers.

This study uses persistent homology to analyze complex transitional networks from time series data.

problem Lack of effective tools to summarize complex topology in transitional networks.
method Persistent homology from topological data analysis applied to coarse-grained state-space networks (CGSSN).
result CGSSN improves dynamic state detection and noise robustness compared to other methods.

Study forecasts U.S. bond index using deep learning, finding persistence is key.

problem Forecasting U.S. aggregate bond index with deep learning methods.
method Constructed a stationary but maximally persistent representation of the bond index, evaluated using MLPs and CNNs.
result Deep learning models outperform traditional methods in short-horizon forecasting of bond indices.

Contrastive divergence (CD) is a promising method of inference in high dimensional distributions with intractable normalizing constants, however, the theoretical foundations justifying its use are somewhat shaky. This document proposes a framework for understanding CD inference, how/when it works, and provides multiple…

2014-05-03abs ↗pdf ↗

New measure EC assesses node contributions in nonlinear, time-varying systems.

problem Existing node contribution measures assume linear, time-invariant dynamics, failing for complex, real-world systems.
method Defined 'emergent contribution (EC)' as a dynamical leverage measure from Jacobians of differentiable models.
result EC diverges from average controllability under persistent regime switching and sign reversal, identifying limits of local linearization.

SGDm with fixed step-size diverges under covariate shift, similar to a parametric oscillator.

problem SGDm with fixed step-size diverges under covariate shift.
method Approximated learning system as a time-varying system of ODEs and characterized divergence/convergence modes.
result SGDm with fixed step-size can diverge under covariate shift, similar to resonance in oscillators.

Co-TSFA improves time series forecasting by distinguishing between short-lived and persistent anomalies.

problem Standard forecasting models fail to distinguish between short-lived and persistent anomalies, leading to overreaction or underreaction.
method Co-TSFA learns to ignore forecast-irrelevant anomalies and respond to forecast-relevant ones through input-only and input-output augmentations and a latent-output alignment loss.
result Co-TSFA improves performance under anomalous conditions while maintaining accuracy on normal data.

By exploiting the property that the RBM log-likelihood function is the difference of convex functions, we formulate a stochastic variant of the difference of convex functions (DC) programming to minimize the negative log-likelihood. Interestingly, the traditional contrastive divergence algorithm is a special case of th…

2017-09-21abs ↗pdf ↗

Timely detection of abrupt anomalies is crucial for real-time monitoring and security of modern systems producing high-dimensional data. With this goal, we propose effective and scalable algorithms. Proposed algorithms are nonparametric as both the nominal and anomalous multivariate data distributions are assumed unkno…

2018-09-14abs ↗pdf ↗

CD learning is shown to be an adversarial game for fitting models.

problem Difficulty in understanding the convergence properties of CD learning.
method Presented an alternative derivation of CD without approximation, showing it as a time-reversal adversarial game.
result CD is an adversarial learning procedure where a discriminator tries to classify time-reversed Markov chains.