Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

219437656874 · Jun 202019922001200920172026
48 results for MM algorithms

MM (majorization--minimization) algorithms are an increasingly popular tool for solving optimization problems in machine learning and statistical estimation. This article introduces the MM algorithm framework in general and via three popular example applications: Gaussian mixture regressions, multinomial logistic regre…

2016-11-12abs ↗pdf ↗

Deep-learning improves 6x6-mm OCTA angiograms by reducing noise and artifacts.

problem Reduced scan quality in 6x6-mm OCTA angiograms due to undersampling.
method Deep-learning-based high-resolution angiogram reconstruction network (HARNet) trained on 3x3-mm and 6x6-mm angiogram data.
result Reconstructed 6x6-mm angiograms have lower noise and better vascular connectivity.

Non-convex optimization is ubiquitous in machine learning. Majorization-Minimization (MM) is a powerful iterative procedure for optimizing non-convex functions that works by optimizing a sequence of bounds on the function. In MM, the bound at each iteration is required to \emph{touch} the objective function at the opti…

2015-06-25abs ↗pdf ↗

Unified approach for federated learning using MM optimization.

problem Scaling stochastic optimization to federated learning.
method Unified Majorize-Minimize (MM) framework for stochastic optimization, extended to federated learning.
result Unified algorithm \QSMM\ for federated learning that aggregates surrogate majorizing functions.

Penalized estimation can conduct variable selection and parameter estimation simultaneously. The general framework is to minimize a loss function subject to a penalty designed to generate sparse variable selection. The majorization-minimization (MM) algorithm is a computational scheme for stability and simplicity, and …

2019-12-23abs ↗pdf ↗

The purpose of this paper is the study of the roots in the mapping class groups. Let ΣΣ be a compact oriented surface, possibly with boundary, let $\PP$ be a finite set of punctures in the interior of ΣΣ, and let $\MM (Σ, \PP)$ denote the mapping class group of $(Σ, \PP)$. We prove that, if ΣΣ is of genus 0, then ea…

2006-07-12abs ↗pdf ↗

New algorithm improves on EM for streaming data, outperforming existing methods.

problem Processing high-volume, streaming data efficiently.
method Incremental stochastic Majorization-Minimization (MM) algorithm.
result The algorithm converges to a stationary point with vanishing gradient.

The paper classifies Poincaré complexes as topological manifolds.

problem Classifying Poincaré complexes as topological manifolds.
method Using spherical fibrations and CW-complexes, the paper proves stability and homotopy equivalence.
result A sufficient condition for Poincaré complexes to be homotopy types of topological manifolds.

This paper proposes MM-DAGs for analyzing traffic congestion, learning multiple DAGs jointly.

problem Analyzing multi-modal traffic data with overlapping and distinct variables.
method Developed MM-DAGs for multi-task, multi-modal DAG learning, using multi-modal regression and CD measure.
result Proved the effectiveness of MM-DAGs in traffic congestion analysis.

The Bradley-Terry model is a popular approach to describe probabilities of the possible outcomes when elements of a set are repeatedly compared with one another in pairs. It has found many applications including animal behaviour, chess ranking and multiclass classification. Numerous extensions of the basic model have a…

2010-11-08abs ↗pdf ↗

We present a selective sampling method designed to accelerate the training of deep neural networks. To this end, we introduce a novel measurement, the minimal margin score (MMS), which measures the minimal amount of displacement an input should take until its predicted classification is switched. For multi-class linear…

2019-11-16abs ↗pdf ↗

A new method trains physics-constrained neural networks more efficiently.

problem Training machine learning tools with limited data and physical constraints.
method Dual-Dimer method for searching saddle points in nonconvex-nonconcave functions.
result The Dual-Dimer method improves training efficiency and convergence speed.

Proposes MM-DUST for efficient generalized lasso solution paths.

problem Efficiently solve generalized lasso problems in large-scale and non-linear models.
method Majorization-minimization dual stagewise algorithm incorporating quadratic majorizers and stagewise learning.
result Established the uniform convergence of approximated solution paths.

Proposes MM-KTD for efficient RL learning with reduced sample size.

problem High sensitivity to parameter selection and overfitting in DNN-based RL methods.
method Adapts Kalman filter parameters using observed states and rewards, enhances sampling efficiency through active learning.
result Significantly reduced number of samples needed to learn optimal policy.

We consider Markov models of stochastic processes where the next-step conditional distribution is defined by a kernel density estimator (KDE), similar to Markov forecast densities and certain time-series bootstrap schemes. The KDE Markov models (KDE-MMs) we discuss are nonlinear, nonparametric, fully probabilistic repr…

2018-07-30abs ↗pdf ↗

We consider practical data characteristics underlying federated learning, where unbalanced and non-i.i.d. data from clients have a block-cyclic structure: each cycle contains several blocks, and each client's training data follow block-specific and non-i.i.d. distributions. Such a data structure would introduce client …

2020-02-18abs ↗pdf ↗

MM-DREX adapts LLM experts for financial trading via dynamic routing.

problem Challenges of non-stationary financial markets and static expert designs.
method MM-DREX uses a VLM-powered dynamic router to allocate expert weights and designs heterogeneous trading experts.
result Significantly outperforms 15 baselines across key metrics.

Let EfE_f be the energy of some knot ττ for any ff from certain class of functions. The problem is to find knots with extremal values of energy. We discuss the notion of the locally perturbed knot. The knot circle minimizes some energies EfE_f and maximizes some others. So, is there any energy such that the circle ne…

2004-11-03abs ↗pdf ↗

We model the behavior of three agent classes acting dynamically in a limit order book of a financial asset. Namely, we consider market makers (MM), high-frequency trading (HFT) firms, and institutional brokers (IB). Given a prior dynamic of the order book, similar to the one considered in the Queue-Reactive models [14,…

2018-02-16abs ↗pdf ↗

New lower bounds improve logistic log-likelihood optimization and inference.

problem Designing computationally tractable lower bounds for logistic log-likelihoods.
method Developed a piece-wise quadratic lower bound that uniformly improves tangent quadratic minorizers.
result Improves the speed of convergence and accuracy of variational Bayes approximations.

New algorithm solves fair PCA, robust PCA, and sparse PCA problems efficiently.

problem Fair Principal Component Analysis (FPCA) to ensure fairness in PCA solutions.
method Iterative MM algorithm with SDP reformulation to quadratic program.
result Algorithm monotonically improves fairness objectives at each iteration.

A new method estimates expectations from subtractive mixture models without sampling.

problem Estimating expectations from multimodal distributions using SMMs.
method Difference representation of SMMs to create unbiased IS estimator (ΔextExΔ ext{Ex}).
result Demonstrates that ΔextExΔ ext{Ex} can achieve comparable estimation quality to auto-regressive sampling but is faster.

New methods for parameter estimation in mechanistic models using data-consistent inversion.

problem Parameter estimation bias in Bayesian analysis for mechanistic models.
method Data-consistent inversion methods based on rejection sampling, MCMC, GANs, and constrained optimization.
result Improved parameter estimation without bias from uninformative priors.

This paper uses deep RL to optimize market quotes from LOB data.

problem Optimizing quotes for market making from complex LOB data.
method Attn-LOB neural network with convolutional filters and attention mechanism for feature extraction; hybrid reward function for continuous action space.
result The RL agent outperforms traditional methods in market making tasks.

IMM uses imitation learning and predictive representation learning to improve market making strategies.

problem Challenges in training RL agents for multi-price level market making strategies.
method IMM combines RL and imitation learning, introducing effective state and action representations and a representation learning unit.
result IMM outperforms existing RL-based market making strategies in financial criteria.

In this paper we introduce some new copulas emerging from shock models. It was shown earlier that reflected maxmin copulas (RMM for short) are not just some specific singular copulas; they contain many important absolutely continuous copulas including the negative quadrant dependent part of the Eyraud-Farlie-Gumbel-Mor…

2018-08-23abs ↗pdf ↗

Proposes a new method for estimating sparse precision matrices in GMRF-MM models.

problem Difficulty in learning GMMs with large parameters and limited data.
method Restricts GMM to GMRF-MM, proposes efficient optimization for sparse precision matrices, and debiases the estimates.
result Debiasing approach outperforms GLASSO in single-GMRF and GMRF-MM cases.

Unified NMF models for various noise distributions, improving feature extraction.

problem Inadequate assumptions for NMF under complex data distributions.
method Unified framework using MM-algorithms for traditional and convex NMF under Tweedie and Negative Binomial models.
result Unified multiplicative update rules for all models, including novel updates for convex NMF.