Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

21436485 · May 202619922001200920172026
48 results for nonlinear TDC

This paper analyzes the sample complexity of two timescale reinforcement learning algorithms.

problem Analyzing the sample complexity of two timescale reinforcement learning algorithms.
method Non-asymptotic analysis of linear and nonlinear TDC and Greedy-GQ algorithms under Markovian sampling with constant stepsize.
result The paper provides non-asymptotic convergence results for two timescale linear and nonlinear TDC and Greedy-GQ algorithms.

We devise a distributional variant of gradient temporal-difference (TD) learning. Distributional reinforcement learning has been demonstrated to outperform the regular one in the recent study \citep{bellemare2017distributional}. In the policy evaluation setting, we design two new algorithms called distributional GTD2 a…

2018-05-20abs ↗pdf ↗

This paper improves tail dependence analysis by introducing a path-based approach.

problem The classical tail dependence coefficient fails to capture non-exchangeable features of tail dependence.
method The paper introduces a path-based maximal tail dependence approach to capture the most pronounced feature of dependence over all possible paths.
result The paper proves the existence and provides an explicit characterization of the path-based maximal TDC, improving analytical and computational tractability.

New measures capture tail dependence and non-exchangeability in financial data.

problem Underestimation of tail dependence and inability to capture non-exchangeable tail dependence.
method Tail copulas and novel tail dependence measures (MTCM, ATCM) are proposed.
result Captures non-exchangeable tail dependence and provides analytical forms for various copulas.

This study improves convergence of two-timescale SA under Markovian noise in reinforcement learning.

problem Stability and convergence of two-timescale stochastic approximations under Markovian noise.
method Introduced a new control strategy for the fast timescale parameter.
result Established almost sure convergence of TDC with eligibility traces under off-policy learning with linear function approximation.

The paper analyzes the sample complexities for policy evaluation with linear function approximation.

problem Policy evaluation with linear function approximation in discounted infinite horizon Markov decision processes.
method Investigates sample complexities for two policy evaluation algorithms: TD and TDC.
result Establishes high-probability sample complexity bounds for policy evaluation algorithms.

Paper analyzes Greedy-GQ for reinforcement learning with Markovian noise.

problem Analyzing Greedy-GQ for reinforcement learning with Markovian noise.
method Develops finite-sample analysis for Greedy-GQ with linear function approximation under Markovian noise.
result Provides theoretical justification for choosing stepsizes for faster convergence.

New TD algorithms stabilize RL tasks by reformulating updates into fixed point equations.

problem TD learning's sensitivity to step size specification.
method Implicit TD algorithms reformulate TD updates into fixed point equations.
result Implicit TD algorithms are more stable and less sensitive to step size.

Tabular in-context learners perform well on biomolecular tasks, but performance depends on the representation used.

problem Predicting biomolecular properties from limited labeled data.
method Evaluating tabular in-context learners on protein fitness regression and small-molecule classification tasks.
result Tabular in-context learners are competitive for protein fitness regression but not for small-molecule classification.

Study of weighted nonlinear flags in symplectic geometry.

problem Understanding the geometry of weighted nonlinear flags.
method Generalizing weighted nonlinear Grassmannians to Frechet manifolds and using them to describe coadjoint orbits.
result Description of coadjoint orbits of Hamiltonian diffeomorphisms using weighted isotropic nonlinear flags.

For a system of second order differential equations we determine a nonlinear connection that is compatible with a given generalized Lagrange metric. Using this nonlinear connection, we can find the whole family of metric nonlinear connections that can be associated with a system of SODE and a generalized Lagrange struc…

2004-12-06abs ↗pdf ↗

Proposes a new method for nonlinear Bayesian updates using ensemble kernel regression.

problem Nonlinear and non-Gaussian Bayesian updates for complex systems.
method Combines Kalman filtering for observed components and kernel density estimation for unobserved components, with subsampling and clustering.
result Reduces estimation errors in highly nonlinear scenarios compared to standard linear updates.

Study solves inverse problems for equations with fractional nonlinearities.

problem Solving inverse problems for semilinear elliptic equations with fractional power nonlinearities.
method Higher order linearization method adapted for fractional order.
result Results of previous studies remain valid for general power nonlinearities.

Sharp Lipschitz bounds and gradient estimates for fully nonlinear parabolic equations.

problem Understanding moduli of continuity for fully nonlinear parabolic equations.
method Proving moduli of continuity of viscosity solutions are subsolutions of one-dimensional parabolic equations.
result Sharp Lipschitz bounds and gradient estimates for fully nonlinear parabolic equations with bounded initial data.

The paper introduces a fast algorithm for learning and forecasting nonlinear dynamics from noisy time series data.

problem Challenges in capturing nonlinear dynamics from noisy time series data.
method A projected nonlinear state-space model with kernel functions applied to projected lines.
result The model effectively learns and forecasts complex nonlinear dynamics with computational efficiency.

AdaKoop efficiently models nonlinear dynamics from nonstationary data streams.

problem Capturing nonlinear dynamics in nonstationary data streams with computational efficiency.
method Koopman operator theory and probabilistic framework for streaming data.
result AdaKoop outperforms state-of-the-art methods in real-time forecasting accuracy and efficiency.

JULIA combines multi-linear and nonlinear models for tensor completion.

problem Complex patterns in real-world tensors require a unified model.
method JULIA unifies multi-linear and nonlinear models with flexible component assignment and efficient alternating optimization.
result JULIA outperforms existing methods in large-scale tensor completion.

Bayesian filtering approach identifies nonlinear restoring forces in dynamic systems.

problem Identification of nonlinear dynamic systems in engineering.
method Modeling the nonlinear restoring force as a Gaussian process, converting it to a state-space model, and inferring internal states and the nonlinear restoring force through filtering and smoothing.
result The approach effectively identifies nonlinear restoring forces in both simulated and experimental datasets.

Unified analysis for nonlinear parametric models in Bayesian optimization.

problem Limited theoretical guarantees for nonlinear parametric models in Bayesian optimization.
method Kernel-based framework for analyzing regularized nonlinear parametric models trained on adaptively collected data.
result Unified convergence guarantees for nonlinear acquisition and surrogate models.

The paper develops adaptive deep learning methods for nonlinear time series models.

problem Estimating mean functions of non-stationary and nonlinear time series models.
method Develops non-penalized and sparse-penalized DNN estimators for general non-stationary time series, derives minimax lower bounds, and shows the sparse-penalized DNN estimator is adaptive and optimal.
result Sparse-penalized DNN estimator achieves minimax optimal rates for many nonlinear AR models.

This paper tackles efficient optimization for nonlinear embeddings in similarity learning.

problem Learning similarity with nonlinear embeddings is challenging due to the large number of pairs.
method Detailed derivations and efficient optimization methods for nonlinear embeddings are developed.
result Efficient optimization methods for nonlinear embeddings are shown to be highly effective.

Constructs differential characters on nonlinear Graßmannians.

problem No specific problem stated; focuses on mathematical construction.
method Using a nonlinear version of the tautological bundle, a transgression map is constructed from MM to nonlinear Graßmannians of submanifolds of fixed type.
result Obtains prequantum circle bundles and central Lie group extensions.

Stock networks, constructed from stock price time series, are a well-established tool for the characterization of complex behavior in stock markets. Following Mantegna's seminal paper, the linear Pearson's correlation coefficient between pairs of stocks has been the usual way to determine network edges. Recently, possi…

2018-04-26abs ↗pdf ↗

The paper introduces a framework to assess nonlinear causality in financial markets.

problem Identifying and quantifying co-dependence between financial instruments.
method Transfer entropy and convergent cross-mapping methods to assess linear and nonlinear causality.
result Stock indices exhibit significant nonlinear causality, and correlation underestimates causality.

Paper connects contrastive learning to MI maximization and establishes robust methods for nonlinear ICA and subspace estimation.

problem Understanding and improving unsupervised representation learning and density ratio estimation.
method The paper connects contrastive learning to MI maximization, establishes new recovery conditions for nonlinear ICA, and proposes a practical outlier-robust method for nonlinear subspace estimation.
result The proposed methods can be seen as maximizing MI, performing nonlinear ICA, or estimating nonlinear subspaces, and are robust to outliers.

Non-Markovian point process shows power-law scaling, similar to nonlinear Markovian process.

problem Understanding the scaling behavior of non-Markovian point processes.
method Analyzed a confined fractional Brownian motion-driven point process and compared it to a nonlinear Markovian process.
result A nonlinear Markovian process can reproduce the power-law scaling behavior of a non-Markovian point process.

Paper accelerates nonlinear mapping in online systems with lower time complexity.

problem Speeding up nonlinear mapping in online systems.
method Integrates an acceleration module into Dendrite Net (DD) to reduce time complexity.
result DD with AC has lower time complexity while maintaining nonlinear mapping and system identification properties.