Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

25.0%50.0%75.0%100.0% · Jun 199319922001200920172026
48 results for Obata-Vétois argument

This is an extended write-up of a talk given in April, 1993 in honor of Raoul Bott's 70th birthday. We first illustrate how some traditional topological and geometric invariants obey ``gluing laws'' inspired by those in classical and quantum field theory. Here we discuss characteristic numbers, particularly the Euler n…

1994-06-28abs ↗pdf ↗

Since their invention, generative adversarial networks (GANs) have become a popular approach for learning to model a distribution of real (unlabeled) data. Convergence problems during training are overcome by Wasserstein GANs which minimize the distance between the model and the empirical distribution in terms of a dif…

2017-09-26abs ↗pdf ↗

MDP Playground tests RL agents across various dimensions for better understanding and debugging.

problem Understanding and debugging reinforcement learning agents across diverse environments and dimensions.
method Controlled testbed with adjustable dimensions for different RL challenges.
result Insights into agent performance and interaction with various dimensions.

Simplifies deep learning scaling analysis without sacrificing accuracy.

problem Interpreting feature learning mechanisms and determining network implicit bias in high-dimensional settings.
method Developed a heuristic approach for predicting data and width scales of feature learning patterns.
result Predictions align with known results and extend to complex architectures.

Non-symmetric rectangular correlation matrices occur in many problems in economics. We test the method of extracting statistically meaningful correlations between input and output variables of large dimensionality and build a toy model for artificially included correlations in large random time series.The results are t…

2010-04-26abs ↗pdf ↗

Study compares ZBDT model to BDT for financial derivatives valuation.

problem Valuation of financial derivatives under catastrophic events.
method Introduced Zero Black-Derman-Toy (ZBDT) model with jumps to zero interest rate.
result ZBDT model better matches financial slowdown risk.

We study the relation between the trading behavior of agents and volatility in toy markets of adaptive inductively rational agents. We show that excess volatility, in such simplified markets, arises as a consequence of {\em i)} the neglect of market impact implicit in price taking behavior and of {\em ii)} excessive re…

2000-04-21abs ↗pdf ↗

Toy model study shows resampling/reweighting can improve feature learning in imbalanced classification.

problem Improving feature learning in imbalanced classification problems.
method High-dimensional toy model with replica method, class-wise resampling/reweighting, and simplified model.
result No resampling/reweighting can sometimes give best feature learning performance.

This work examines how adversarial vulnerability changes with the dimensionality of the subspace of perturbations.

problem Understanding adversarial vulnerability in constrained input spaces.
method Investigates adversarial vulnerability in subspace VV of the input space XX with varying dimensions, using PGD attacks and analyzing the dependence on εε and dim(V)/dim(X)dim(V)/dim(X).
result Adversarial success of PGD attacks is a monotonically increasing function of $ε( rac{dim(V)}{dim(X)})^{ rac{1}{q}}$.

In this work we address the problem of argument search. The purpose of argument search is the distillation of pro and contra arguments for requested topics from large text corpora. In previous works, the usual approach is to use a standard search engine to extract text parts which are relevant to the given topic and su…

2019-05-26abs ↗pdf ↗

A toy model shows how locality can emerge in the universe's Hamiltonian and initial state.

problem Understanding the emergence of locality in the universe's Hamiltonian and initial state.
method A loss functional is minimized by gradient descent to find a tensor product structure.
result Local structure emerges in the universe's Hamiltonian and initial state through spontaneous symmetry breaking.

We propose a modification of the classical Black-Derman-Toy (BDT) interest rate tree model, which includes the possibility of a jump with small probability at each step to a practically zero interest rate. The corresponding BDT algorithms are consequently modified to calibrate the tree containing the zero interest rate…

2019-08-12abs ↗pdf ↗

The study reveals a transition in neural network performance from infinite-width to variance-limited behavior as dataset size increases.

problem Understanding the transition from infinite-width to variance-limited behavior in neural networks.
method Empirical study of the transition from infinite-width to variance-limited behavior as a function of sample size and network width.
result The critical sample size \( P^* \) is approximately \( \sqrt{N} \) for polynomial regression with ReLU networks.

This paper studies the concept of instantaneous arbitrage in continuous time and its relation to the instantaneous CAPM. Absence of instantaneous arbitrage is equivalent to the existence of a trading strategy which satisfies the CAPM beta pricing relation in place of the market. Thus the difference between the arbitrag…

2019-01-16abs ↗pdf ↗

We consider the class of short rate interest rate models for which the short rate is proportional to the exponential of a Gaussian Markov process x(t) in the terminal measure r(t) = a(t) exp(x(t)). These models include the Black, Derman, Toy and Black, Karasinski models in the terminal measure. We show that such intere…

2012-04-04abs ↗pdf ↗

Analyzed Guyon's volatility model for existence and uniqueness.

problem Existence and uniqueness of a strong solution for Guyon's volatility model.
method Proved existence and uniqueness of a strong solution, characterised boundary behavior, derived asymptotic option prices, and small-time estimates.
result Existence and uniqueness of a strong solution for Guyon's volatility model.

A clustering algorithm based on the Hausdorff distance is introduced and compared to the single and complete linkage. The three clustering procedures are applied to a toy example and to the time series of financial data. The dendrograms are scrutinized and their features confronted. The Hausdorff linkage relies of firm…

2008-01-07abs ↗pdf ↗

APD method decomposes neural network parameters into simple, faithful components.

problem Understanding the internal mechanisms learned by neural networks.
method Attribution-based Parameter Decomposition (APD) method.
result Demonstrated effectiveness in recovering features, separating computations, and identifying representations.

A theory of feature geometry using spectral analysis of weight matrices.

problem Current methods decompose neural network activations into sparse linear features, losing geometric structure.
method Develops a theory by analyzing the spectra of weight-derived matrices, introducing the frame operator.
result Features collapse onto single eigenspaces, organizing into tight frames, and admit discrete classification.

We construct an elementary, combinatorial kind of topological quantum field theory, based on curves, surfaces, and orientations. The construction derives from contact invariants in sutured Floer homology and is essentially an elaboration of a TQFT defined by Honda--Kazez--Matic. This topological field theory stores inf…

2012-01-22abs ↗pdf ↗

Study evaluates unsupervised disentanglement methods on a toy dataset.

problem Lack of clear disentanglement metrics capturing independent features.
method Empirical evaluation of six unsupervised disentanglement methods on MPI3D dataset.
result Beta-TCVAE outperforms other methods in metrics, but not in disentanglement quality.

Random matrix theory predicts neural representations generalize well.

problem Understanding why neural representations generalize well in practice.
method Applied random matrix theory to kernel regression and neural networks.
result GCV estimator accurately predicts generalization risk in overparameterized settings.

Graph Networks are used to make decisions in potentially complex scenarios but it is usually not obvious how or why they made them. In this work, we study the explainability of Graph Network decisions using two main classes of techniques, gradient-based and decomposition-based, on a toy dataset and a chemistry task. Ou…

2019-05-31abs ↗pdf ↗

This technical report proves components consistency for the Doubly Stochastic Dirichlet Process with exponential convergence of posterior probability. We also present the fundamental properties for DSDP as well as inference algorithms. Simulation toy experiment and real-world experiment results for single and multi-clu…

2016-05-24abs ↗pdf ↗

Maps sets to probability distributions to minimize information loss.

problem Learning to map sets to probability distributions to preserve information.
method Relates set operations to probability distribution interpolations and demonstrates a preliminary solution.
result Experimental results show the effectiveness of the set embedding approach.

Study extends Elkalla's work on subnormal subgroups to PD3PD_3-groups, but L2L^2-Betti numbers need verification.

problem Verifying L2L^2-Betti numbers for PD3PD_3-groups and group pairs.
method Algebraic arguments extending Elkalla's work, but reliant on unproven L2L^2-Betti number hypothesis.
result Need further research on L2L^2-Betti numbers for general PD3PD_3-groups.

SLT explains neural network success by closing theory-practice gap.

problem Failure of classical inference and learning theory in modern neural networks.
method Physics-inspired Singular Learning Theory (SLT) applied to neural networks.
result SLT recovers known and novel scaling laws for neural network phase transitions.

We use standard perturbation techniques originally formulated in quantum (statistical) mechanics in the analysis of a toy model of a stock market which is given in terms of bosonic operators. In particular we discuss the probability of transition from a given value of the {\em portfolio} of a certain trader to a differ…

2009-07-15abs ↗pdf ↗

EGFs use ergodicity to simplify generative flows for easier training and imitation learning.

problem Challenges in training generative flows, especially in continuous settings and for imitation learning.
method EGFs leverage ergodicity to build simple flows with universality guarantees and tractable FM loss. They introduce a KL-weakFM loss for IL training without a separate reward model.
result EGFs simplify generative flow training and enable effective imitation learning.

Study improves curvature estimate for stable marginally outer trapped hypersurfaces with a free boundary.

problem Curvature estimate for stable marginally outer trapped hypersurfaces with a free boundary.
method Iteration argument based on uniform area bound.
result Improved curvature estimate for stable marginally outer trapped hypersurfaces.

New method for hedging path-dependent options with price impact using probabilistic arguments.

problem Hedging of path-dependent options with price impact.
method Dual formulation using probabilistic arguments, proving existence of perfect hedging portfolios.
result Existence of a perfect hedging portfolio for path-dependent options with price impact.

New insights into how depth and width affect in-context learning in deep models.

problem Understanding how various resources impact in-context learning in deep models.
method Analyzed linear regression in a deep linear self-attention model, varying resources like depth, width, context length, and training steps.
result Increasing depth improves in-context learning even at infinite context length, contrary to previous findings.