Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

4998147196 · Jun 202019922001200920172026
48 results for universal least favorable submodel

ULFS-KDPE estimates parameters efficiently without influence functions.

problem Estimating pathwise differentiable parameters in nonparametric models.
method Kernel debiased plug-in estimator based on universal least favorable submodel.
result Semiparametric efficiency achieved without influence function derivation.

The paper quantifies and attributes uncertainty in complex system simulations.

problem Uncertainty in complex system simulations due to unknown or approximated subprocesses.
method Developed a framework for quantifying and attributing submodel uncertainty using bootstrapping, Bayesian model averaging, and tree-based methods.
result Individual submodels contribute to overall uncertainty, and their importance can be quantified.

This paper simplifies finding least favorable priors by reducing dimensionality.

problem Finding least favorable priors is challenging due to infinite-dimensional optimization.
method Develops a dimensionality reduction method using Bregman divergences.
result Allows use of gradient ascent algorithms for finding least favorable priors.

Federated learning was proposed with an intriguing vision of achieving collaborative machine learning among numerous clients without uploading their private data to a cloud server. However, the conventional framework requires each client to leverage the full model for learning, which can be prohibitively inefficient fo…

2019-11-06abs ↗pdf ↗

Data is one of the essential ingredients to power deep learning research. Small datasets, especially specific to medical institutes, bring challenges to deep learning training stage. This work aims to develop a practical deep multimodal that can classify patients into abnormal and normal categories accurately as well a…

2019-02-24abs ↗pdf ↗

We study the mixtures of factorizing probability distributions represented as visible marginal distributions in stochastic layered networks. We take the perspective of kernel transitions of distributions, which gives a unified picture of distributed representations arising from Deep Belief Networks (DBN) and other netw…

2012-11-05abs ↗pdf ↗

We prove the statistical consistency of kernel Partial Least Squares Regression applied to a bounded regression learning problem on a reproducing kernel Hilbert space. Partial Least Squares stands out of well-known classical approaches as e.g. Ridge Regression or Principal Components Regression, as it is not defined as…

2009-02-25abs ↗pdf ↗

Unified analysis of reweighted least-squares algorithms for linear models.

problem Recovering unknown signals from linear measurements using reweighted least squares.
method Unified asymptotic analysis of IRLS, lin-RFM, and alternating minimization algorithms.
result The algorithms can achieve favorable performance in a few iterations with appropriate reweighting.

Combining insights from machine learning and quantum Monte Carlo, the stochastic reconfiguration method with neural network Ansatz states is a promising new direction for high-precision ground state estimation of quantum many-body problems. Even though this method works well in practice, little is known about the learn…

2019-10-24abs ↗pdf ↗

There is a well known link between (maximal) polar representations and isotropy representations of symmetric spaces provided by Dadok. Moreover, the theory by Tits and Burns-Spatzier provides a link between irreducible symmetric spaces of non-compact type of rank at least three and irreducible topological spherical bui…

2012-05-28abs ↗pdf ↗

New analysis of Muon and SignSGD on matrix-valued least squares problems.

problem Understanding the behavior of Muon and SignSGD on matrix-valued least squares problems.
method Derive explicit deterministic dynamics to study learning behavior of Muon and SignSGD.
result Muon and SignSGD exhibit different optimal learning rates and convergence characteristics based on batch size and data covariance.

A method for identifying NPWARX models with arbitrary domains using probabilistic mixture models.

problem Identifying hybrid system models with discontinuous maps.
method Probabilistic mixture model with a neural network for nonlinear partitioning and Expectation Maximization for parameter estimation.
result Demonstrated on a nonlinear piece-wise problem with discontinuous maps.

Paper introduces \ell-DER for regression tasks using morphological operators and convex-concave procedure.

problem Developing a universal approximator for regression tasks.
method Introduces \ell-DER model, trains it using a convex-concave procedure (CCP) to minimize least-squares.
result Outperforms other hybrid morphological models and state-of-the-art approaches.

We applied machine learning to predict whether a gene is involved in axon regeneration. We extracted 31 features from different databases and trained five machine learning models. Our optimal model, a Random Forest Classifier with 50 submodels, yielded a test score of 85.71%, which is 4.1% higher than the baseline scor…

2017-10-30abs ↗pdf ↗

We show that deep narrow Boltzmann machines are universal approximators of probability distributions on the activities of their visible units, provided they have sufficiently many hidden layers, each containing the same number of units as the visible layer. We show that, within certain parameter domains, deep Boltzmann…

2014-11-14abs ↗pdf ↗

Functional PLS improves prediction and inference for scalar responses from functional predictors.

problem Estimating scalar responses from functional predictors in an ill-posed inverse problem.
method Functional partial least squares (PLS) estimator with adaptive early stopping and new tests.
result PLS attains nearly minimax-optimal convergence rates and detects local alternatives.

We study connected sum at infinity on smooth, open manifolds. This operation requires a choice of proper ray in each manifold summand. In favorable circumstances, the connected sum at infinity operation is independent of ray choices. For each m at least 3, we construct an infinite family of pairs of m-manifolds on whic…

2013-04-30abs ↗pdf ↗

We consider a problem of ranking and selection via simulation in the context of personalized decision making, where the best alternative is not universal but varies as a function of some observable covariates. The goal of ranking and selection with covariates (R&S-C) is to use simulation samples to obtain a selection p…

2017-10-07abs ↗pdf ↗

In this paper, we prove uniform lower bounds on the volume growth of balls in the universal covers of Riemannian surfaces and graphs. More precisely, there exists a constant δ>0δ>0 such that if (M,hyp)(M,hyp) is a closed hyperbolic surface and hh another metric on MM with $\area(M,h)\leq δ\area(M,hyp)$ then for every radiu…

2013-04-12abs ↗pdf ↗

CNNs identify stock market trend endpoints based on expert opinion.

problem Finding optimal entry and exit points for stock market trends.
method Three CNN submodels sequentially identify changepoints, locate them, and classify trends as upward, downward, or flat.
result CNNs can identify long-term trends based on expert opinion, offering a new approach to stock market analysis.

In a financial market with a continuous price process and proportional transaction costs we investigate the problem of utility maximization of terminal wealth. We give sufficient conditions for the existence of a shadow price process, i.e.~a least favorable frictionless market leading to the same optimal strategy and u…

2014-08-26abs ↗pdf ↗

If (M^n, g) is a complete Riemannian manifold with filling radius at least R, then we prove that it contains a ball of radius R and volume at least c(n)R^n. If (M^n, hyp) is a closed hyperbolic manifold and if g is another metric on M with volume at most c(n)Volume(M,hyp), then we prove that the universal cover of (M,g…

2006-10-06abs ↗pdf ↗

The study shows how nonnegative Ricci curvature and metric cones imply the existence of abelian subgroups in the fundamental group of open manifolds.

problem Understanding the structure of fundamental groups of open manifolds with specific curvature properties.
method Analyzing the properties of the Riemannian universal cover and its asymptotic cones.
result The fundamental group of an open manifold with nonnegative Ricci curvature and certain geometric properties contains an abelian subgroup of finite index.

This paper introduces a dual problem to study a continuous-time consumption and investment problem with incomplete markets and stochastic differential utility. For Epstein-Zin utility, duality between the primal and dual problems is established. Consequently the optimal strategy of the consumption and investment proble…

2016-01-14abs ↗pdf ↗

Study a simplified model of multiverse structure with synchronized timelines.

problem Understanding the complex structure of the local multiverse with multiple universes.
method Time-amalgamated globally hyperbolic model of multiverse as a collection of parallel universes.
result Elementary particles are transcosmic strings with multiple endpoints on parallel universes.

Dual-sPLS improves feature selection and prediction in high-dimensional data.

problem Relating variables to a response in high-dimensional chemometric problems.
method Generalizes PLS1 algorithm with dual norm penalizations and a shrinking ratio parameter.
result Favorably compares to similar regression methods on simulated and real chemical data.

UVU simplifies value uncertainty quantification in RL.

problem Estimating epistemic uncertainty in value functions for reinforcement learning.
method UVU uses squared prediction errors between an online learner and a fixed, randomly initialized target network, incorporating policy-conditional value uncertainty.
result UVU achieves equal performance to large ensembles on challenging offline RL settings, with computational savings.

We show that for a large class of hyperbolic knots and links, we can determine bounds on the volume of the link complement from combinatorial information given by a link diagram. Specifically, there is a universal constant C such that if a knot or link admits a prime, twist reduced diagram with at least 2 twist regions…

2006-04-21abs ↗pdf ↗

We study a distributionally robust mean square error estimation problem over a nonconvex Wasserstein ambiguity set containing only normal distributions. We show that the optimal estimator and the least favorable distribution form a Nash equilibrium. Despite the non-convex nature of the ambiguity set, we prove that the …

2018-09-24abs ↗pdf ↗

ESNs trained with Tikhonov least squares approximate ergodic dynamical systems in L2(μ) norm.

problem Approximating ergodic dynamical systems using ESNs.
method Tikhonov least squares regression on ESNs trained on observations from an ergodic dynamical system.
result ESNs trained with Tikhonov least squares approximate the target function in the L2(μ) norm.

We describe sufficient conditions which guarantee that a finite set of mapping classes generate a right-angled Artin group quasi-isometrically embedded in the mapping class group. Moreover, under these conditions, the orbit map to Teichmuller space is a quasi-isometric embedding for both of the standard metrics. As a …

2010-07-07abs ↗pdf ↗

We generalize the higher rank rigidity theorem to a class of Finsler spaces, i.e. Berwald spaces. More precisely, we prove that a complete connected Berwald space of finite volume and bounded nonpositive flag curvature with rank at least 22 whose universal cover is irreducible, is a locally symmetric space or a locall…

2015-10-15abs ↗pdf ↗

The study examines the universality of Gaussian data in high-dimensional generalized linear estimation.

problem Understanding when Gaussian data suffices for high-dimensional generalized linear estimation.
method Sharp asymptotic expressions for test and training errors in high-dimensional Gaussian mixture data with labels from a single-index model.
result The universality of Gaussian data in error estimation depends on the alignment between target weights and mixture cluster means and covariances.

Complex-valued neural networks can approximate any continuous function.

problem Generalizing the universal approximation theorem to complex-valued networks.
method Characterizing activation functions for complex networks to approximate any continuous function.
result Different activation functions are required for deep vs shallow complex networks to achieve universal approximation.

Paper proposes neural networks for learning functions from sets to graphs.

problem Challenges in learning Set2Graph functions, including computational and memory complexity.
method Develops a family of neural network models that are practical and of maximal expressive power, approximating arbitrary continuous Set2Graph functions.
result Models can approximate arbitrary continuous Set2Graph functions over compact sets.

Canonical Correlation Analysis (CCA) is a widely used statistical tool with both well established theory and favorable performance for a wide range of machine learning problems. However, computing CCA for huge datasets can be very slow since it involves implementing QR decomposition or singular value decomposition of h…

2014-07-16abs ↗pdf ↗