Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

174348522696 · Jun 202019922001200920172026
48 results for Transition Function

A new method uses deep learning to efficiently sample rare transitions for estimating committor functions.

problem Efficiently sampling rare transitions to estimate committor functions in high-dimensional problems.
method DASTR (Deep Adaptive Sampling on Transition Paths) method using deep generative models.
result Significantly improved accuracy in approximating committor functions through efficient sampling.

The study proves unique harmonic functions and combinatorial properties of vertex-transitive graphs.

problem Proving combinatorial properties of vertex-transitive graphs.
method Using harmonic functions and quasi-isometry to R\mathbb{R}, proving uniqueness and combinatorial results.
result Connective constant of non-degenerate vertex-transitive graphs is at least the golden mean.

We derive the exact solution of a one-dimensional Markov functional model with log-normally distributed interest rates in discrete time. The model is shown to have two distinct limiting states, corresponding to small and asymptotically large volatilities, respectively. These volatility regimes are separated by a phase …

2010-07-05abs ↗pdf ↗

New algorithm achieves data-dependent regret bounds in MDPs with unknown transitions.

problem Achieving best-of-both-worlds guarantees with data-dependent regret bounds in MDPs with unknown transitions.
method Optimistic follow-the-regularized-leader algorithm with new optimistic Q-function estimators and transition bonus.
result First-order, second-order, and path-length bounds with polylog(T) regret in the stochastic regime.

We focus on variational inference in dynamical systems where the discrete time transition function (or evolution rule) is modelled by a Gaussian process. The dominant approach so far has been to use a factorised posterior distribution, decoupling the transition function from the system states. This is not exact in gene…

2018-12-14abs ↗pdf ↗

New algorithms achieve no-regret learning even with adversarial transitions and losses.

problem No-regret learning impossible with adversarial transitions and losses.
method Developed algorithms for adversarial Markov Decision Processes with smooth regret increase.
result Achieved O~(T+CextsfP)\widetilde{O}(\sqrt{T} + C^{ extsf{P}}) regret, with CextsfPC^{ extsf{P}} measuring adversarial transition function.

The existence of nonconstant harmonic Dirichlet functions on a Cayley graph of a discrete group is equivalent to the nonvanishing of the first L2-cohomology of the given group. It was first proven by Cheeger and Gromov that such functions do not exists on the Cayley-graph of an amenable group. The result was extended u…

1998-06-23abs ↗pdf ↗

A multi-task GP model tracks time-varying transition probabilities between two states.

problem Tracking time-varying transition probabilities between 'moves' and 'pauses' states.
method Kernel-based multi-task Gaussian Process model with time-variability and constraints.
result Enforces constraints while learning transition probabilities.

Predicting labels of nodes in a network, such as community memberships or demographic variables, is an important problem with applications in social and biological networks. A recently-discovered phase transition puts fundamental limits on the accuracy of these predictions if we have access only to the network topology…

2014-04-30abs ↗pdf ↗

Proposes a method for approximating transition densities of SDEs driven by gamma processes.

problem Calculating transition densities for SDEs driven by gamma processes.
method Taylor-type approximation and conditional expectation of multiple stochastic integrals.
result Efficiency of the proposed method demonstrated through numerical tests.

New methods use machine learning to simulate rare transitions in molecular systems.

problem Simulating rare transitions between metastable states in molecular dynamics.
method Generative models and reinforcement learning for importance sampling.
result Efficiently generated transition paths linking metastable states.

FourNet approximates financial transition densities using Fourier transforms.

problem Approximating transition densities in finance with high accuracy.
method FourNet is a novel FFNN with Gaussian activation, learning from characteristic functions.
result FourNet can approximate transition densities arbitrarily well with finite neurons.

Deep reinforcement learning method finds rare events in complex systems.

problem Computing transition pathways in high-dimensional systems.
method Formulated as a cost minimization problem, solved using DDPG with physical properties.
result Efficiently samples and computes globally optimal transition pathways.

Study phase transitions with prescribed mean curvature in Riemannian manifolds.

problem Understanding phase transitions with prescribed mean curvature in geometric settings.
method Analyzing solutions to inhomogeneous semilinear elliptic PDEs, establishing bounds and asymptotics.
result Established upper and lower bounds for eigenvalues of phase transition problems.

We obtain explicit formulas for the trivialization functions of the SU(3){\rm SU}(3) principal bundle G2S6G_2 \to S^6 over two affine charts. We also calculate the explicit transition function of this fibration over the equator of the six-sphere. In this way we obtain a new proof of the known fact that this fibration corres…

2018-11-08abs ↗pdf ↗

PQR estimates reward functions from actions and states without assuming state-only rewards.

problem Estimating reward functions from actions and states without state-only assumptions.
method Deep learning approach that sequentially estimates policy, Q-function, and reward.
result PQR uniquely recovers true reward with known transitions and bounds error with unknown transitions.

Abstract: Investigates the role of activation functions in neural networks and their physical basis.

problem Understanding the role of activation functions in neural networks and their physical basis.
method Formalizes the use of activation functions in neural inference by relating them to phase transitions in statistical physics.
result Reveals the physical justification for the performance of typical activation functions in neural networks.

The paper develops ML algorithms for calibrating credit rating transition models for high and low default portfolios.

problem Calibration of credit rating transition models for high and low default portfolios.
method Developed Maximum likelihood (ML) algorithms, including Laplace approximation for high-default portfolios and particle filter with Gaussian process regression for low-default portfolios.
result Both algorithms produce accurate approximations of the likelihood function and ML estimates of model parameters.

New algorithms reduce regret in reinforcement learning with MNL approximations.

problem Efficient reinforcement learning with MNL function approximation for MDPs.
method Proposed randomized exploration algorithms with frequentist regret guarantees.
result Achieved improved regret bounds for MNL transition models.

ISOKANN learns collective variables and effective dynamics for metastable transitions.

problem Understanding metastable transitions in complex molecular systems.
method Integrates Koopman operators with neural networks to extract CVs and effective dynamics.
result Reconstructs coarse-grained kinetics and reproduces transition times across barriers.

A new method for ILO with transition model disparity using an intermediary policy.

problem Learning tasks from expert observations with different transition dynamics.
method Training an intermediary policy to match the state transitions of the expert dataset.
result Our method outperforms existing ILO approaches with transition model mismatch.

Study shows Merton model limits to Poisson process with log-normal intensity, improving default portfolio prediction.

problem Improving prediction of default portfolios using complex models.
method Applying Merton model with log-normal intensity function to Poisson process, discussing temporal correlation effects.
result Power decay model provides better generalization for long-term default portfolio data.

We present a global optimization algorithm for clustering data given the ratio of likelihoods that each pair of data points is in the same cluster or in different clusters. To define a clustering solution in terms of pairwise relationships, a necessary and sufficient condition is that belonging to the same cluster sati…

2015-06-09abs ↗pdf ↗

Efficient RL algorithm for multinomial logistic MDPs with provable guarantees.

problem Model-based RL for episodic MDPs with unknown transition probabilities.
method Upper confidence bound-based algorithm for exploration-exploitation balance.
result Achieves ildeO(dH3T) ilde{O}(d \sqrt{H^3 T}) regret bound for multinomial logistic models.

Study on neural networks' storage capacity and solution space structure.

problem Understanding the storage capacity and solution space structure of neural networks.
method Replica method from statistical physics.
result Storage capacity per parameter remains finite even with infinite width and weights exhibit negative correlations.

Develops CLTs for Markov chain transition probabilities and policies.

problem Estimating transition probabilities and policies in controlled Markov chains.
method Non-parametric estimator for transition matrices; CLTs for value, Q-, and advantage functions; goodness-of-fit tests.
result Asymptotic normality of estimators under specific logging policies.

Wide neural networks become linear, but adding bottlenecks makes them bilinear or multilinear.

problem Understanding the transition of neural networks from linearity to higher-order functions.
method Analyzing the behavior of randomly initialized wide neural networks with and without bottleneck layers.
result Bottleneck layers transform the network's function from linear to bilinear or multilinear.

A new RL paradigm reduces state-action-value function approximation inefficiency.

problem Challenges in state-action-value function approximation for RL.
method State Action Separable Reinforcement Learning (sasRL) decouples action space from value function learning.
result sasRL achieves up to 75% better performance than state-of-the-art MDP-based RL algorithms.

Generative Stochastic Networks (GSNs) have been recently introduced as an alternative to traditional probabilistic modeling: instead of parametrizing the data distribution directly, one parametrizes a transition operator for a Markov chain whose stationary distribution is an estimator of the data generating distributio…

2013-12-19abs ↗pdf ↗

Paper tackles transfer RL under unobserved context, developing methods to reduce bias.

problem Transfer RL with unobserved contextual information leading to biased models.
method Develops causal bounds on transition and reward functions using demonstrator's data.
result Proposes Q learning and UCB-Q learning algorithms that converge to true value function without bias.

New algorithm reduces suboptimality in imitation learning to nearly optimal levels.

problem Statistical limits of imitation learning in MDPs with known transitions.
method Mimic-MD algorithm and reduction to value estimation problem.
result Upper bound of O(SH3/2/N)O(|\mathcal{S}|H^{3/2}/N) for suboptimality, with efficient computation.

Value functions struggle to represent transition dynamics, impacting statistical efficiency.

problem Limited representational power of value functions in capturing transition dynamics.
method Case studies of various reinforcement learning problems to explore the limitations of value-based methods.
result Value-based methods can be as efficient as model-based ones in some cases but severely underperform in others due to information loss.

PCA improves detection of phase transitions in muon spectroscopy data from various materials.

problem Subtle changes in asymmetry function indicate phase transitions, but existing methods require material-specific knowledge.
method Applied unsupervised PCA to muon spectroscopy asymmetry data from multiple materials.
result PCA can recover phase transition indicators and improve detection of material-specific variations.