Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,786 papers · 148 categories

Trend · papers per month

12.5%25.0%37.5%50.0% · Sep 199319922001200920172026
48 results for parallel applications

Study on charged parallel spinors and mass-charge inequalities.

problem Equality case of the spin positive mass theorem with charge.
method Investigation of charged parallel spinors and application to extremal charged manifolds.
result Characterization of the equality case of the mass-charge inequality.

Optimized parallel RNN training reaches up to 845x speedup.

problem Expensive RNN training through back-propagation through time (BPTT).
method Optimized parallel algorithm \opt based on ELM, leveraging GPU shared memory and QR factorization.
result Up to 845x speedup over sequential training and 20x less time to train.

We design and analyse variations of the classical Thompson sampling (TS) procedure for Bayesian optimisation (BO) in settings where function evaluations are expensive, but can be performed in parallel. Our theoretical analysis shows that a direct application of the sequential Thompson sampling algorithm in either synch…

2017-05-25abs ↗pdf ↗

Non-negative Matrix Factorization (NMF) is a key kernel for unsupervised dimension reduction used in a wide range of applications, including topic modeling, recommender systems and bioinformatics. Due to the compute-intensive nature of applications that must perform repeated NMF, several parallel implementations have b…

2019-04-16abs ↗pdf ↗

Curvature criteria for A-simple singularities and their parallel curves identified.

problem Determining singularity types of A-simple singularities and their parallel curves.
method Defined curvature parameters and criteria for A-simple singularities.
result Criteria to determine singularity types of A-simple singularities and their parallel curves.

Stochastic gradient descent~(SGD) and its variants have become more and more popular in machine learning due to their efficiency and effectiveness. To handle large-scale problems, researchers have recently proposed several parallel SGD methods for multicore systems. However, existing parallel SGD methods cannot achieve…

2015-08-24abs ↗pdf ↗

End-to-end DA method for domain-invariant CNNs using parallel audio recordings.

problem Distribution mismatches between training and application data in machine listening.
method Enforcing equal hidden layer representations for domain-parallel samples.
result Learn domain-invariant classifiers without requiring classification labels.

We prove that the length difference between a closed periodic curve and its parallel curve at a sufficiently small distance is proportional to the rotation index. As an application, the rotation index of a curve could be estimated by means of Cauchy-Crofton formula.

2007-11-11abs ↗pdf ↗

Motivated by the study of Weyl structures on conformal manifolds admitting parallel weightless forms, we define the notion of conformal product of conformal structures and study its basic properties. We obtain a classification of Weyl manifolds carrying parallel forms, and we use it to investigate the holonomy of the a…

2009-01-23abs ↗pdf ↗

Decomposes submanifolds with special tensors into simpler parts.

problem Understanding the structure of submanifolds with special tensors.
method Established a decomposition theorem for submanifolds with nonnegative sectional curvature and a Codazzi tensor with parallel mean curvature.
result Submanifolds with these tensors are locally isometric to a direct product of irreducible factors.

Method predicts how probability distributions evolve over time.

problem Predicting how systems described by probability distributions evolve under different conditions.
method Wasserstein Parallel Transport
result Wasserstein Parallel Transport provides counterfactual comparisons of distributional dynamics.

This paper optimizes object tracking on edge devices with small matrices.

problem Efficiently tracking objects in video sequences on edge devices with small matrices.
method Parallelized a Simple Online and Real-time Tracking (SORT) application on shared-memory multicores.
result Throughput-based parallelization technique outperforms multi-threading for small matrices.

This paper uses supervised learning to predict optimal chunk-size for parallel linear algebra operations.

problem Finding the optimal chunk-size for parallel linear algebra operations.
method The paper uses supervised learning models (logistic regression, neural networks, decision trees) to predict the optimal chunk-size for multiple linear algebra operations.
result The custom decision tree model outperforms classical decision trees and other models in predicting optimal chunk-size for linear algebra operations.

Absolute parallelism geometry is frequently used for physical applications. It has two main defects, from the point of view of applications. The first is the identical vanishing of its curvature tensor. The second is that its autoparallel paths do not represent physical trajectories. The present work shows how these de…

2002-09-17abs ↗pdf ↗

In many applications of black-box optimization, one can evaluate multiple points simultaneously, e.g. when evaluating the performances of several different neural network architectures in a parallel computing environment. In this paper, we develop a novel batch Bayesian optimization algorithm --- the parallel knowledge…

2016-06-14abs ↗pdf ↗

We give criteria for which a principal curvature becomes a bounded CC^\infty-function at non-degenerate singular points of wave fronts by using geometric invariants. As applications, we study singularities of parallel surfaces and extended distance squared functions of wave fronts. Moreover, we relate these singularit…

2016-12-02abs ↗pdf ↗

Study torsion parallel spinors on Lorentzian 4-manifolds and their evolution flows.

problem Investigate torsion parallel spinors on Lorentzian four-manifolds.
method Geometric study via spinorial polyforms and supersymmetric NS-NS system.
result Globally hyperbolic evolution flow determined by supersymmetric solutions.

The paper explores Lorentzian connections with parallel skew torsion.

problem Understanding metric connections with parallel skew-symmetric torsion in Lorentzian signature.
method Analyzing holonomy algebras, torsion, and curvature; constructing examples; classifying homogeneous spaces.
result Complete classification of Lorentzian naturally reductive homogeneous spaces in low dimensions.

We develop a parallel variational inference (VI) procedure for use in data-distributed settings, where each machine only has access to a subset of data and runs VI independently, without communicating with other machines. This type of "embarrassingly parallel" procedure has recently been developed for MCMC inference al…

2015-10-14abs ↗pdf ↗

In real world industrial applications of topic modeling, the ability to capture gigantic conceptual space by learning an ultra-high dimensional topical representation, i.e., the so-called "big model", is becoming the next desideratum after enthusiasms on "big data", especially for fine-grained downstream tasks such as …

2014-11-10abs ↗pdf ↗

We introduce a novel approach for parallelizing MCMC inference in models with spatially determined conditional independence relationships, for which existing techniques exploiting graphical model structure are not applicable. Our approach is motivated by a model of seismic events and signals, where events detected in d…

2016-12-02abs ↗pdf ↗

New method solves blind inverse problems by optimizing both operator and image parameters.

problem Solving blind inverse problems with known forward operator.
method Parallel reverse diffusion guided by gradients from intermediate stages.
result State-of-the-art performance on blind deblurring and imaging through turbulence.

I prove the bistability of linear evolution equations x=A(t)xx' = A(t)x in a Banach space EE, where the operator-valued function AA is of the form A(t)=f(t)G(t,f(t))A(t) = f'(t)G(t,f(t)) for a binary operator-valued function GG and a scalar function ff. The constant that bounds the solutions of the equation is computed explicitly; it i…

2015-02-12abs ↗pdf ↗

OptEx accelerates first-order optimization with parallelized iterations.

problem Inefficiencies in first-order optimization algorithms for complex tasks.
method Approximately parallelized iterations using kernelized gradient estimation.
result OptEx achieves substantial efficiency improvements with an effective acceleration rate of Ω(N)Ω(\sqrt{N}).

We study special almost Kaehler manifolds whose curvature tensor satisfies the second curvature condition of Gray. It is shown that for such manifolds, the torsion of the first canonical Hermitian is parallel. This enables us to show that every AK_2-manifold has parallel torsion. Some applications of this result, conce…

2003-01-21abs ↗pdf ↗

The numerical solution of large-scale PDEs, such as those occurring in data-driven applications, unavoidably require powerful parallel computers and tailored parallel algorithms to make the best possible use of them. In fact, considerations about the parallelization and scalability of realistic problems are often criti…

2017-05-10abs ↗pdf ↗

Clustering samples according to an effective metric and/or vector space representation is a challenging unsupervised learning task with a wide spectrum of applications. Among several clustering algorithms, k-means and its kernelized version have still a wide audience because of their conceptual simplicity and efficacy.…

2017-10-09abs ↗pdf ↗

Batch-splitting (data-parallelism) is the dominant distributed Deep Neural Network (DNN) training strategy, due to its universal applicability and its amenability to Single-Program-Multiple-Data (SPMD) programming. However, batch-splitting suffers from problems including the inability to train very large models (due to…

2018-11-05abs ↗pdf ↗