Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

227454681908 · Jun 202019922001200920172026
48 results for inner problem

A new algorithm tackles bilevel optimization with multiple inner minima.

problem Challenges in bilevel optimization with multiple inner minima.
method Reformulated as constrained optimization, solved via primal-dual bilevel optimization (PDBO) algorithm.
result First non-asymptotic convergence guarantee for bilevel optimization with multiple inner minima.

The paper challenges the belief that more inner iterations at test time improve performance in implicit deep learning.

problem The performance improvement of implicit deep learning models with increased inner iterations at test time.
method Theoretical analysis of a simple setting, validation on implicit deep learning problems.
result Overparametrization plays a key role; increasing the number of iterations at test time does not improve performance for overparametrized networks.

The paper solves the Andreadakis problem for specific groups using inner automorphisms.

problem Solving the Andreadakis problem for specific groups.
method Generalizing tools from [Dar19b] to study subgroups of IAn, focusing on the behavior of the Andreadakis problem with inner automorphisms.
result The Andreadakis equality holds for the pure braid group and the mapping class group of the n-punctured sphere.

A new stochastic method tackles bi-level optimization problems in deep learning.

problem Bi-level optimization problems in deep learning, including hyperparameter optimization and meta learning.
method Turning a BLO problem into a stochastic optimization, using SGLD MCMC and a recurrent algorithm to compute MC-estimated hypergradient.
result Our method is more robust to suboptimal inner optimization and non-unique inner minima, leading to more reliable solutions.

We present and discuss some open problems formulated by participants of the International Workshop "Knots, Braids, and Auto\-mor\-phism Groups" held in Novosibirsk, 2014. Problems are related to palindromic and commutator widths of groups; properties of Brunnian braids and two-colored braids, corresponding to an amalga…

2015-01-22abs ↗pdf ↗

New approach to bilevel optimization for machine learning using functional methods.

problem Solving bilevel optimization problems in machine learning, especially with over-parameterized neural networks.
method Functional point of view, scalable and efficient algorithms for functional bilevel optimization.
result Demonstrates benefits of functional approach on instrumental regression and reinforcement learning tasks.

This work speeds up hyperparameter selection for non-smooth convex models using implicit differentiation.

problem Optimizing hyperparameters of non-smooth convex models.
method Implicit differentiation of proximal gradient and coordinate descent methods.
result Implicit differentiation can speed up hyperparameter optimization, especially for non-smooth problems.

In this paper, we develop a loop group description of harmonic maps F:MG/K\mathcal{F}: M \rightarrow G/K ``of finite uniton type", from a Riemann surface MM into inner symmetric spaces of compact or non-compact type. This develops work of Uhlenbeck, Segal, and Burstall-Guest to non-compact inner symmetric spaces. To be mo…

2013-05-11abs ↗pdf ↗

A new pruning method improves neural network efficiency and accuracy.

problem Optimizing neural network efficiency and accuracy through pruning.
method Formulated as a Knapsack Problem, the method optimizes trade-off between neuron importance and computational cost. Channels are pruned while maintaining network structure, and fine-tuned using inner knowledge distillation from parent network.
result State-of-the-art pruning results on ImageNet, CIFAR-10, and CIFAR-100.

Study on combustion theory solutions, proving nondegeneracy and stability in limit.

problem One-phase singular perturbation problem in combustion theory.
method Introduce density condition to preserve nondegeneracy, classify stable solutions.
result Global stable solutions have flat level sets in dimensions ≤ 4.

The variance reduction class of algorithms including the representative ones, SVRG and SARAH, have well documented merits for empirical risk minimization problems. However, they require grid search to tune parameters (step size and the number of iterations per inner loop) for optimal performance. This work introduces `…

2019-08-25abs ↗pdf ↗

This paper explores the non-convex composition optimization in the form including inner and outer finite-sum functions with a large number of component functions. This problem arises in some important applications such as nonlinear embedding and reinforcement learning. Although existing approaches such as stochastic gr…

2017-11-13abs ↗pdf ↗

In this paper, we consider the convex and non-convex composition problem with the structure 1ni=1nFi(G(x))\frac{1}{n}\sum\nolimits_{i = 1}^n {{F_i}( {G( x )} )}, where G(x)=1nj=1nGj(x)G( x )=\frac{1}{n}\sum\nolimits_{j = 1}^n {{G_j}( x )} is the inner function, and Fi()F_i(\cdot) is the outer function. We explore the variance reduction based met…

2018-09-06abs ↗pdf ↗

Paper improves adversarial training using a learned optimizer.

problem Improving robustness of deep learning models against adversarial attacks.
method Empirically identified PGD attack's limitations and used a learning-to-learn framework to train an adaptive inner optimizer.
result The proposed framework consistently improves model robustness over traditional adversarial training methods.

New spectral functionals for Dirac operators with inner fluctuations computed.

problem Spectral functionals and Dirac operators with inner fluctuations.
method Extension of spectral functionals for Dirac operators with inner fluctuations.
result Computed spectral Einstein functional for Dirac operator with inner fluctuations on even-dimensional spin manifolds.

A graph (digraph) G=(V,E)G=(V,E) with a set TVT\subseteq V of terminals is called inner Eulerian if each nonterminal node vv has even degree (resp. the numbers of edges entering and leaving vv are equal). Cherkassky and Lovász showed that the maximum number of pairwise edge-disjoint TT-paths in an inner Eulerian graph $G…

2005-10-21abs ↗pdf ↗

This paper constructs quandles with abelian inner automorphism groups from graphs, proving their homogeneity.

problem Finding quandles with specific automorphism properties.
method Starting from simple graphs, the paper constructs quandles with abelian inner automorphism groups and proves their homogeneity.
result Homogeneous quandles with abelian inner automorphism groups are constructed from vertex-transitive graphs.

We consider the problem of designing locality sensitive hashes (LSH) for inner product similarity, and of the power of asymmetric hashes in this context. Shrivastava and Li argue that there is no symmetric LSH for the problem and propose an asymmetric LSH based on different mappings for query and database points. Howev…

2014-10-21abs ↗pdf ↗

Equivalent tests for SGD batch size selection found.

problem Finding equivalent tests for adaptive batch size selection in SGD.
method Norm and inner product/orthogonality tests equivalence demonstration.
result Norm and inner product/orthogonality tests are equivalent under specific conditions.

New algorithm for linear bandits tackles Optimal Transport problems.

problem Optimal Transport problems not covered by traditional linear bandits.
method Embed actions into a Hilbertian subspace, penalize optimism, use least-squares estimation.
result Achieves same regret bounds as OFUL but interpolates between ildeO(T) ilde{\mathcal O}(\sqrt{T}) and O(T){\mathcal O}(T).

Fewer data weight updates lead to faster convergence in machine learning models.

problem Improving robustness of machine learning models through data mixing.
method Analyzing convergence behavior of data mixing with a finite number of inner steps.
result The optimal number of inner steps scales with the budget and type of gradients used.

Groups with specific properties have vanishing 2\ell^2-Betti numbers.

problem Understanding 2\ell^2-Betti numbers for certain groups.
method Introduced cheap 1-rebuilding property and used structure theorem of Tucker-Drob.
result First 2\ell^2-Betti numbers vanish for specified groups.

The author reviews his results on locally compact homogeneous spaces with inner metric, in particular, homogeneous manifolds with inner metric. The latter are isometric to homogeneous (sub-)Finslerian manifolds; under some additional conditions they are isometric to homogeneous (sub)-Riemannian manifolds. The class ΩΩ

2014-12-26abs ↗pdf ↗

Researchers prove inner product recovery is impossible in latent space models.

problem Recovering inner products in latent space models with random geometric graphs.
method Rate-distortion theory applied to Gaussian or spherical latent locations.
result Impossible to recover inner products if dimensionality exceeds nh(p)n h(p), matching positive results' conditions.

Minwise hashing (Minhash) is a widely popular indexing scheme in practice. Minhash is designed for estimating set resemblance and is known to be suboptimal in many applications where the desired measure is set overlap (i.e., inner product between binary vectors) or set containment. Minhash has inherent bias towards sma…

2014-11-14abs ↗pdf ↗

In this paper, we determine the automorphism group of the pp-cones (p2p\neq 2) in dimension greater than two. In particular, we show that the automorphism group of those pp-cones are the positive scalar multiples of the generalized permutation matrices that fix the main axis of the cone. Next, we take a look at a pro…

2018-08-05abs ↗pdf ↗

This paper addresses the nearest neighbor search problem under inner product similarity and introduces a compact code-based approach. The idea is to approximate a vector using the composition of several elements selected from a source dictionary and to represent this vector by a short code composed of the indices of th…

2014-06-19abs ↗pdf ↗

Paper proposes a new method to optimize feature coordinates for better image classification.

problem Improving feature extraction for better machine learning classification.
method Mutual-energy inner product optimization method.
result The method enhances low-frequency features and suppresses high-frequency noise, leading to better classification results.

We classify homotopes of classical symmetric spaces (studied in Part I of this work). Our classification uses the fibered structure of homotopes: they are fibered as symmetric spaces, with flat fibers, over a non-degenerate base; the base spaces correspond to inner ideals in Jordan pairs. Using that inner ideals in cla…

2010-11-13abs ↗pdf ↗

ES-Single uses ES to estimate gradients in unrolled graphs, reducing variance and improving performance.

problem Estimating gradients in unrolled computation graphs with low variance and stability.
method Evolution strategies (ES) applied to unrolled graphs, with a single perturbation per particle.
result ES-Single reduces variance compared to PES, leading to better performance in various tasks.

Study of Gaussian distributions using entropic Gromov-Wasserstein and inner product Gromov-Wasserstein.

problem Optimal transportation between Gaussian distributions with different dimensions.
method Entropic Gromov-Wasserstein and inner product Gromov-Wasserstein, with closed-form expressions and von Neumann's trace inequality.
result Closed-form expressions for the entropic IGW and its unbalanced variant between Gaussian distributions.

We study hamiltonian actions of compact groups in the presence of compatible involutions. We show that the lagrangian fixed point set on the symplectically reduced space is isomorphic to the disjoint union of the involutively reduced spaces corresponding to involutions on the group strongly inner to the given one. Our …

2003-03-26abs ↗pdf ↗

A core capability of intelligent systems is the ability to quickly learn new tasks by drawing on prior experience. Gradient (or optimization) based meta-learning has recently emerged as an effective approach for few-shot learning. In this formulation, meta-parameters are learned in the outer loop, while task-specific m…

2019-09-10abs ↗pdf ↗

Study bounds the index of minimal submanifolds using energy measures and Yang-Mills-Higgs equations.

problem Bounding the index of codimension 2 minimal submanifolds.
method Second inner variation of energy, convergence of energy measures, and stress-energy tensors.
result Bound the Morse index of the submanifold by the index of critical points.

We propose a quantization based approach for fast approximate Maximum Inner Product Search (MIPS). Each database vector is quantized in multiple subspaces via a set of codebooks, learned directly by minimizing the inner product quantization error. Then, the inner product of a query to a database vector is approximated …

2015-09-04abs ↗pdf ↗

New method solves complex constrained optimization problems.

problem Constrained nonconvex-nonconcave minimax optimization problems.
method Inexact proximal gradient method using sequential convex programming.
result Established complexity guarantees for approximate stationary points.