Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

4795142189 · May 202619922001200920172026
48 results for Log Evidence

The log-determinant of a kernel matrix appears in a variety of machine learning problems, ranging from determinantal point processes and generalized Markov random fields, through to the training of Gaussian processes. Exact calculation of this term is often intractable when the size of the kernel matrix exceeds a few t…

2017-04-05abs ↗pdf ↗

The presence of log-periodic structures before and after stock market crashes is considered to be an imprint of an intrinsic discrete scale invariance (DSI) in this complex system. The fractal framework of the theory leaves open the possibility of observing self-similar log-periodic structures at different time scales.…

2005-01-21abs ↗pdf ↗

Two sets of high quality income data are analysed in detail, one set from the UK, one from the USA. It is firstly demonstrated that both a log-normal distribution and a Boltzmann distribution can give very accurate fits to both these data sets. The absence of a power tail in the US data set is then discussed. Taken in …

2004-06-28abs ↗pdf ↗

New method corrects Laplace/BIC errors in singular models, revealing effective dimension.

problem Laplace/BIC errors in singular models due to incorrect effective dimension assumption.
method RLCT (real log canonical threshold) to correct effective dimension in linear models.
result Correct evidence slope and effective dimension estimation in linear settings.

Develops an anytime-valid framework for optimal policy identification from logged contextual bandit data.

problem Selecting the optimal policy from a candidate policy class while monitoring evidence continuously.
method Constructs a time-indexed set that retains the true optimal policy set uniformly over time.
result The procedure allows the analyst to monitor policy values, eliminate clearly suboptimal policies, and stop at data-dependent times without invalidating inference.

Expectation Maximization (EM) is among the most popular algorithms for maximum likelihood estimation, but it is generally only guaranteed to find its stationary points of the log-likelihood objective. The goal of this article is to present theoretical and empirical evidence that over-parameterization can help EM avoid …

2018-10-26abs ↗pdf ↗

Bitcoin volatility shows multifractal structure, contradicting rough volatility models.

problem Applying rough volatility models to Bitcoin volatility data.
method Normalised p-variation framework, multifractal Detrended Fluctuation Analysis, log-log moment scaling, wavelet leaders.
result Bitcoin volatility exhibits multifractal structure, violating rough volatility model assumptions.

ECLIPSE detects AI hallucinations in finance with high accuracy.

problem Hallucinations in AI-generated answers limit safe deployment in finance.
method Combines entropy estimation and perplexity decomposition to measure model evidence use.
result ECLIPSE achieves ROC AUC of 0.89 and average precision of 0.90 on financial QA dataset.

Stochastic variational inference (SVI) plays a key role in Bayesian deep learning. Recently various divergences have been proposed to design the surrogate loss for variational inference. We present a simple upper bound of the evidence as the surrogate loss. This evidence upper bound (EUBO) equals to the log marginal li…

2019-12-02abs ↗pdf ↗

Diffusion models adapt to data geometry through log-domain smoothing.

problem Understanding why diffusion models generalize well across diverse domains.
method Investigating the role of score matching and log-domain smoothing in diffusion models.
result Log-domain smoothing adapts the diffusion model to the data manifold.

Since August 2000, the stock market in the USA as well as most other western markets have depreciated almost in synchrony according to complex patterns of drops and local rebounds. In \cite{SZ02QF}, we have proposed to describe this phenomenon using the concept of a log-periodic power law (LPPL) antibubble, characteriz…

2003-10-05abs ↗pdf ↗

Develops a support-aware framework for reserve-policy selection in advertising markets.

problem Log-based reserve-price evaluation risks weak support and subgroup harm.
method Support-aware offline decision framework converting logged evidence into certified policies.
result Preserves the best gate-passing policy while eliminating only policies with certified regret.

Calculation of the log-normalizer is a major computational obstacle in applications of log-linear models with large output spaces. The problem of fast normalizer computation has therefore attracted significant attention in the theoretical and applied machine learning literature. In this paper, we analyze a recently pro…

2015-06-12abs ↗pdf ↗

Bayesian inference and superstatistics model financial volatility dynamics across different timescales.

problem Modeling correlated volatility in financial time series with heavy tails and long memory.
method Superstatistical dynamics, Bayesian Inference, Metropolis-Hasting sampling.
result The log-Normal model is reliable for short timescales, while inverse-Gamma is preferred for long timescales.

In this paper we review the concepts of Bayesian evidence and Bayes factors, also known as log odds ratios, and their application to model selection. The theory is presented along with a discussion of analytic, approximate and numerical techniques. Specific attention is paid to the Laplace approximation, variational Ba…

2014-11-11abs ↗pdf ↗

We apply two non-parametric methods to test further the hypothesis that log-periodicity characterizes the detrended price trajectory of large financial indices prior to financial crashes or strong corrections. The analysis using the so-called (H,q)-derivative is applied to seven time series ending with the October 1987…

2002-05-25abs ↗pdf ↗

Transformers for binary decisions are sensitive to evidence order, leading to unreliable outcomes.

problem Order sensitivity in Transformers for binary decisions leads to unreliable outcomes.
method Formalized an expectation-realization gap and developed QMV and EDFL bounds.
result Uniform permutation mixtures reduce dispersion and improve reliability.

A hypothesis that the financial log-periodicity, cascading self-similarity through various time scales, carries signatures of a law is pursued. It is shown that the most significant historical financial events can be classified amazingly well using a single and unique value of the preferred scaling factor lambda=2, whi…

2002-09-25abs ↗pdf ↗
Critical Crashescond-mat.stat-mech

We argue that the word ``critical'' in the title is not purely literary. Based on our and other previous work on nonlinear complex dynamical systems, we summarize present evidence, on the Oct. 1929, Oct. 1987, Oct. 1987 Hong-Kong, Aug. 1998 global market events and on the 1985 Forex event, for the hypothesis advanced f…

1999-01-06abs ↗pdf ↗

The study examines volatility models and finds decoupling of short- and long-term correlation structures.

problem Understanding the dynamic of volatility at different time scales.
method Developed a composite likelihood estimation framework for parametric continuous-time stationary Gaussian processes.
result The short- and long-term correlation structures of stochastic volatility are decoupled.

Normalizing flow regression approximates posterior distributions without additional sampling.

problem Bayesian inference with computationally expensive likelihood evaluations.
method Normalizing flow regression (NFR) for offline inference.
result NFR yields a tractable posterior approximation through regression on existing log-density evaluations.

We clarify the status of log-periodicity associated with speculative bubbles preceding financial crashes. In particular, we address Feigenbaum's [2001] criticism and show how it can be rebuked. Feigenbaum's main result is as follows: ``the hypothesis that the log-periodic component is present in the data cannot be reje…

2001-06-26abs ↗pdf ↗

Improved cutting plane method for convex optimization and games.

problem Efficiently finding points in convex sets or proving they do not contain balls.
method Optimal cutting plane algorithm using leverage scores and advanced data structures.
result Significant improvement in time complexity for convex optimization and games.

Dropout-based regularization methods can be regarded as injecting random noise with pre-defined magnitude to different parts of the neural network during training. It was recently shown that Bayesian dropout procedure not only improves generalization but also leads to extremely sparse neural architectures by automatica…

2017-05-20abs ↗pdf ↗

New score helps choose PIML model parameters, reducing ambiguity in model quality.

problem Ambiguity in measuring model quality in PIML due to multi-objective fitting.
method Introduces Physics-Informed Log Evidence (PILE) score in Gaussian process framework.
result PILE minimizes ambiguity in model selection, improving hyperparameter choices.

VAEs improve representation learning by inverting the data-generating process through self-consistency.

problem VAEs struggle to invert the data-generating process, yet often succeed in representation learning.
method Studied VAEs in the limit of near-deterministic decoders, proving self-consistency and showing ELBO convergence to a regularized log-likelihood.
result VAEs can perform independent mechanism analysis (IMA), recovering true latent factors under specific conditions.

Unified Bayesian model explains in-context learning and activation steering in LLMs.

problem Understanding and controlling the behavior of large language models (LLMs) through prompts and activations.
method Developed a Bayesian model to explain and predict the effects of in-context learning and activation steering.
result Unified model predicts distinct phases and sudden shifts in LLM behavior, explaining prior empirical phenomena.

In Bayesian statistics, the marginal likelihood, also known as the evidence, is used to evaluate model fit as it quantifies the joint probability of the data under the prior. In contrast, non-Bayesian models are typically compared using cross-validation on held-out data, either through kk-fold partitioning or leave-$p…

2019-05-21abs ↗pdf ↗

New bounds on SGD's final iterate convergence rate in constant dimension.

problem Characterize the convergence rate of SGD's final iterate in constant dimension.
method Proved lower bounds of Ω(logd/T)Ω(\log d/\sqrt{T}) and Ω(logd/T)Ω(\log d/T) for non-smooth Lipschitz convex and strongly convex functions respectively.
result First general dimension dependent lower bound on SGD's final iterate convergence rate.

Online method for state estimation and parameter learning in SSMs.

problem State estimation and parameter learning in state-space models.
method Stochastic gradient optimization of variational lower bound, using backward decompositions and Bellman recursions.
result Ability to operate online without revisiting historic observations.