Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

158315473630 · Jun 202019922001200920182026
48 results for Parallel States

We prove a conjecture formulated by Pablo M. Chacon and Guillermo A. Lobos in [Pseudo-parallel Lagrangian submanifolds in complex space forms, Differential Geom. Appl.] stating that every Lagrangian pseudo-parallel submanifold of a complex space form of dimension at least 3 is semi-parallel.

2008-11-21abs ↗pdf ↗

Predictability enables efficient parallelization of nonlinear models.

problem Understanding which nonlinear state space models can be efficiently parallelized.
method Established a relationship between system dynamics and optimization problem conditioning, quantified by the largest Lyapunov exponent.
result Predictable systems can be evaluated in O((logT)2)O((\log T)^2) time, improving over conventional sequential approaches.

Parallelizes MCTS for continuous domains using leaf and root parallelization.

problem Solving challenging tasks in continuous domains using MCTS.
method Extends existing parallelization strategies to continuous domains, focusing on leaf and root parallelization.
result Proposes two final selection strategies for continuous states in root parallelization.

Stochastic gradient descent~(SGD) and its variants have become more and more popular in machine learning due to their efficiency and effectiveness. To handle large-scale problems, researchers have recently proposed several parallel SGD methods for multicore systems. However, existing parallel SGD methods cannot achieve…

2015-08-24abs ↗pdf ↗

A submanifold of a Riemannian symmetric space is called parallel if its second fundamental form is a parallel section of the appropriate tensor bundle. We classify parallel submanifolds of the Grassmannian $\rmG^+_2(\R^{n+2})$ which parameterizes the oriented 2-planes of the Euclidean space Rn+2\R^{n+2}\,. Our main resul…

2011-07-28abs ↗pdf ↗

Cyclic Data Parallelism reduces memory usage and balances gradient communications.

problem Training large deep learning models requires efficient parallelism to scale.
method Cyclic Data Parallelism shifts micro-batches from simultaneous to sequential execution, balancing memory and gradient communications.
result Cyclic Data Parallelism reduces total memory usage and balances gradient communications.

Study on deep and wide echo state networks for forecasting complex time series.

problem Performance analysis of deep reservoir computing models.
method Investigates the impact of partitioning neurons and parallel pathways on forecasting accuracy.
result Wide and deep networks outperform shallow models in forecasting multiscale spatiotemporal data.

Monte Carlo (MC) methods are widely used for Bayesian inference and optimization in statistics, signal processing and machine learning. A well-known class of MC methods are Markov Chain Monte Carlo (MCMC) algorithms. In order to foster better exploration of the state space, specially in high-dimensional applications, s…

2015-07-30abs ↗pdf ↗

Enhances parallelism in decentralized learning for larger networks.

problem Scalability limitations in decentralized learning with increasing number of machines.
method Proposes Decentralized Anytime SGD, a novel algorithm that extends parallelism threshold.
result Establishes a theoretical upper bound on parallelism surpassing current state-of-the-art.

New deep ESN architectures improve memory capacity and prediction accuracy.

problem Improving memory capacity and prediction accuracy of ESNs.
method Two new deep ESN architectures: parallel and series. Analysis of memory capacity and prediction accuracy.
result Parallel deep ESNs have equivalent memory capacity to shallow ESNs, while series deep ESNs have smaller memory capacity.

New method solves blind inverse problems by optimizing both operator and image parameters.

problem Solving blind inverse problems with known forward operator.
method Parallel reverse diffusion guided by gradients from intermediate stages.
result State-of-the-art performance on blind deblurring and imaging through turbulence.

Parallel-in-time solver reduces ODE simulation time from linear to logarithmic.

problem Efficiently solving ordinary differential equations (ODEs) with reduced computational cost.
method Formulated a parallel-in-time probabilistic numerical ODE solver using time-parallel formulation of iterated extended Kalman smoothers.
result Reduces span cost from linear to logarithmic in the number of time steps.

AgEBO-Tabular combines NAS and hyperparameter tuning for fast, high-performing tabular models.

problem Developing high-performing predictive models for large tabular data sets is challenging.
method Combines aging evolution NAS and asynchronous Bayesian optimization for hyperparameter tuning in data-parallel training.
result Automatically discovered neural network models outperform state-of-the-art AutoML ensembles in inference speed by two orders of magnitude.

Study of bound states in quantum layers with confining potentials.

problem Investigating bound states in quantum layers with confining potentials.
method Developed a general approach using parallel coordinates based on the surface but outside its cut locus.
result Discrete eigenvalues exist for certain quantum layers with positive total Gauss curvature.

Study timelike surfaces with parallel mean curvature in Minkowski 4-space.

problem Existence and uniqueness of timelike surfaces with parallel mean curvature.
method Introduce canonical parameters and prove existence and uniqueness theorem.
result Each timelike surface with parallel mean curvature is determined by three geometric functions.

New BO methods exploit parallel experiments, reducing search time and improving solution quality.

problem Limitation of Bayesian optimization in exploiting parallel experiments.
method Propose new parallel BO paradigms that exploit the structure of the system to partition the design space.
result Significantly reduce search time and increase probability of finding global solutions.

Mesh-TensorFlow enables efficient deep learning on large clusters.

problem Memory constraints and inefficiency in batch-splitting for large models.
method Introduces Mesh-TensorFlow for specifying general tensor computations across a multi-dimensional mesh of processors.
result Trains Transformer models with up to 5 billion parameters on TPU meshes of up to 512 cores.

The study characterizes compact homogeneous manifolds with Bismut parallel torsion.

problem Characterizing compact homogeneous manifolds with specific geometric properties.
method Investigating Hermitian manifolds with Bismut parallel torsion, focusing on locally homogeneous manifolds.
result Characterization of compact Chern flat BTP manifolds and properties of BTP compact Hermitian locally homogeneous manifolds.

Efficient event generation for collider phenomenology using parallel Langevin sampling and learned Stein diagnostics.

problem Event generation for precision collider phenomenology.
method Parallel Langevin sampling with learned Stein diagnostics.
result Relaxation time is estimated using a data-driven approach.

Paper tackles robust knowledge transfer in parallel RL tasks.

problem Transfer knowledge from low-tier to high-tier tasks in parallel RL without shared dynamics or reward functions.
method Identifies Optimal Value Dominance condition and proposes online learning algorithms for both tasks.
result Achieves constant regret on partial states and near-optimal regret when tasks are dissimilar.

Paper proposes DCT for efficient hybrid parallel training of large recommendation models.

problem Training large recommendation models at scale with efficient communication.
method Dynamic Communication Thresholding (DCT) for both Data Parallelism and Model Parallelism.
result Reduces communication by 100x and 20x during DP and MP, respectively, improving training time by 37%.

KalMamba improves RL efficiency with probabilistic SSMs.

problem Efficiency in learning and inference for probabilistic SSMs in RL.
method Combines Mamba's scalability with Kalman filtering for efficient probabilistic SSMs.
result KalMamba outperforms state-of-the-art SSMs in RL, especially on longer sequences.

A new GNN model SPIN achieves state-of-the-art performance on diverse real-world datasets.

problem Graph classification efficiency and accuracy.
method Parallel neighborhood aggregations (PA-GNNs) and SPIN model.
result SPIN model achieves state-of-the-art performance on diverse real-world datasets.

QEM uses parallel importance weighting for fast approximate Bayesian inference.

problem Bayesian inference challenges in large models with many observations and latent variables.
method Expectation Maximization (EM) with massively parallel importance weighting.
result QEM is faster and more scalable than RWS and VI.