Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,932 papers · 148 categories

Trend · papers per month

4.6%9.3%13.9%18.5% · Mar 202019922001200920172026
48 results for MSA processing

Seq-SetNet processes sequence sets directly, improving protein structure prediction.

problem Processing sequence sets (MSAs) for structural inference without considering sequence order.
method Developed a symmetric function module to integrate features from MSAs.
result Seq-SetNet outperforms state-of-the-art approaches by 3.6% in precision.

This paper explores integration and contagion among US metropolitan housing markets. The analysis applies Federal Housing Finance Agency (FHFA) house price repeat sales indexes from 384 metropolitan areas to estimate a multi-factor model of U.S. housing market integration. It then identifies statistical jumps in metrop…

2011-10-18abs ↗pdf ↗

Paper tackles Arabic question similarity, outperforming state-of-the-art.

problem Detecting semantically similar questions in Arabic is challenging.
method Utilizes contextualized word representations (ELMo embeddings) trained on MSA and dialectic sentences, combined with a pairwise similarity layer.
result Achieves 93% F1-score on Modern Standard Arabic benchmark and 82% on dialectical benchmark.

MSA compares neural representations' intrinsic geometry for better understanding.

problem Existing similarity measures fail to capture subtle distinctions between neural network solutions.
method Metric similarity analysis (MSA) using Riemannian geometry.
result MSA can disentangle features of neural computations and compare nonlinear dynamics.

Paper introduces MSA for weakly supervised covariance alignment in MEG signals.

problem Limited labeled signals in target datasets for MEG applications.
method Mixing model Stiefel Adaptation (MSA) leveraging unlabeled data.
result MSA outperforms recent methods in brain-age regression with MEG signals.

Improves speaker verification for variable-duration utterances using a feature pyramid module.

problem Improving robustness for variable-duration utterances in speaker verification.
method Integrates a feature pyramid module into multi-scale aggregation to enhance speaker-discriminative information from multiple layers.
result Improves performance for both short and long utterances compared to state-of-the-art approaches.

This study analyzes information flow networks in Chinese stock sectors using transfer entropy.

problem Understanding information transmission and market dynamics in Chinese stock sectors.
method Daily closing price data of 28 sectors from 2000 to 2017, transfer entropy, maximum spanning arborescence (MSA).
result The composite sector is an information source, and the non-bank financial sector is an information sink.

This paper investigates the risk-return relationship in determination of housing asset pricing. In so doing, the paper evaluates behavioral hypotheses advanced by Case and Shiller (1988, 2002, 2009) in studies of boom and post-boom housing markets. The paper specifies and tests a multi-factor housing asset pricing mode…

2011-03-30abs ↗pdf ↗

Unified deep learning framework improves SV in noisy, reverberant, and long non-speech segments.

problem Robust speaker verification in adverse environments, especially short speech segments.
method Feature Pyramid Module (FPM)-based Multi-scale Aggregation (MSA), Self-adaptive Soft VAD (SAS-VAD), Masking-based Speech Enhancement (SE).
result The proposed method outperforms baseline systems in challenging conditions.

This chapter is an attempt to present a mathematical theory of compound fractional Poisson processes. The chapter begins with the characterization of a well-known Lévy process: The compound Poisson process. The semi-Markov extension of the compound Poisson process naturally leads to the compound fractional Poisson proc…

2011-03-03abs ↗pdf ↗

A deep Neyman-Scott process uses Poisson processes for efficient inference in complex point processes.

problem Efficient inference in complex hierarchical point processes.
method Developed an efficient posterior sampling via Markov chain Monte Carlo for likelihood-based inference.
result More hidden Poisson processes improve likelihood fitting and event prediction.

The study examines Hawkes processes and their long-term behavior.

problem Understanding the long-term behavior of Hawkes processes.
method Proving functional limit theorems under various conditions on the dispersion of child events.
result Functional limit theorems hold for Hawkes processes with different levels of child event dispersion.

Elliptical processes generalize Gaussian and Student-t models with fat tails and computational efficiency.

problem Need for models with fat tails and computational tractability.
method Represent elliptical distributions as continuous mixtures of Gaussian distributions, derive closed-form expressions for marginal and conditional distributions.
result Elliptical processes offer advantages in robust regression compared to Gaussian processes.

We investigate the Student-t process as an alternative to the Gaussian process as a nonparametric prior over functions. We derive closed form expressions for the marginal likelihood and predictive distribution of a Student-t process, by integrating away an inverse Wishart process prior over the covariance kernel of a G…

2014-02-18abs ↗pdf ↗

Efficient methods for Lévy models using SINH-regular processes.

problem Efficient numerical methods for evaluating Lévy models.
method Defining SL-processes and sSL-processes, deriving properties of characteristic exponent, and showing all popular Lévy processes can be subordinated to Brownian motion.
result All crucial properties of characteristic exponent are consequences of a specific representation, and all popular Lévy processes are SL- or sSL-subordinated Brownian motion.

The aim of process discovery, originating from the area of process mining, is to discover a process model based on business process execution data. A majority of process discovery techniques relies on an event log as an input. An event log is a static source of historical data capturing the execution of a business proc…

2017-04-25abs ↗pdf ↗

Researchers study the geometric properties of a specific type of stable processes.

problem Understanding the information geometry of tempered stable processes.
method Derivation of α-divergence, Fisher information matrices, and α-connections.
result Obtained Fisher information matrices and α-connections for statistical manifolds.

This study bridges discrete and continuous state spaces using the Ehrenfest process and diffusion models.

problem Understanding the relationship between discrete and continuous state spaces in stochastic processes.
method Investigates time-continuous Markov jump processes on discrete state spaces and their correspondence to state-continuous diffusion processes.
result The time-reversal of the Ehrenfest process converges to the time-reversed Ornstein-Uhlenbeck process, bridging discrete and continuous state spaces.

The fractional Poisson process (FPP) is a counting process with independent and identically distributed inter-event times following the Mittag-Leffler distribution. This process is very useful in several fields of applied and theoretical physics including models for anomalous diffusion. Contrary to the well-known Poiss…

2011-04-21abs ↗pdf ↗

Gaussian process priors are commonly used in aerospace design for performing Bayesian optimization. Nonetheless, Gaussian processes suffer two significant drawbacks: outliers are a priori assumed unlikely, and the posterior variance conditioned on observed data depends only on the locations of those data, not the assoc…

2018-01-18abs ↗pdf ↗

We characterize the combinatorial structure of conditionally-i.i.d. sequences of negative binomial processes with a common beta process base measure. In Bayesian nonparametric applications, such processes have served as models for latent multisets of features underlying data. Analogously, random subsets arise from cond…

2013-12-31abs ↗pdf ↗

The paper analyzes multivariate Hawkes processes and their induced population processes.

problem Analyzing the time-dependent joint probability distribution of multivariate Hawkes processes.
method Exact and asymptotic analysis of general multivariate Hawkes processes and their induced population processes.
result Full characterization of the time-dependent joint transform of the multivariate population process and its intensity process.

Study on error probability for classification of heavy-tailed renewal processes.

problem Error probability in classification of heavy-tailed renewal processes.
method Asymptotic expressions for Bhattacharyya bound on misclassification error probabilities.
result Obtained asymptotic expressions for misclassification error probabilities.

Study shows convergence rates for BSDEs approximated by compound Poisson processes.

problem Analyzing convergence rates of BSDEs driven by Lévy processes.
method Approximating Lévy processes by compound Poisson processes and studying BSDEs.
result Optimal convergence rates derived for BSDEs in L2\mathbb L^2-norm and Wasserstein distance.

Non-Markovian point process shows power-law scaling, similar to nonlinear Markovian process.

problem Understanding the scaling behavior of non-Markovian point processes.
method Analyzed a confined fractional Brownian motion-driven point process and compared it to a nonlinear Markovian process.
result A nonlinear Markovian process can reproduce the power-law scaling behavior of a non-Markovian point process.

State spaces of multifactor approximations of nonnegative Volterra processes are linear transformations of the nonnegative orthant.

problem Characterizing state spaces of multifactor approximations of nonnegative Volterra processes.
method Explicit linear transformation of the nonnegative orthant.
result State spaces of multifactor approximations of nonnegative Volterra processes are given by explicit linear transformation of the nonnegative orthant.

Automated process discovery is a class of process mining methods that allow analysts to extract business process models from event logs. Traditional process discovery methods extract process models from a snapshot of an event log stored in its entirety. In some scenarios, however, events keep coming with a high arrival…

2018-04-08abs ↗pdf ↗

We introduce Dirac processes, using Dirac delta functions, for short-rate-type pricing of financial derivatives. Dirac processes add spikes to the existing building blocks of diffusions and jumps. Dirac processes are Generalized Processes, which have not been used directly before because the dollar value of non-Real nu…

2015-04-17abs ↗pdf ↗

Improved Gaussian process experts model for complex data.

problem Limitations of standard Gaussian processes: scalability and predictive performance.
method Proposes a new mixture model of Gaussian process experts based on kernel stick-breaking processes.
result Improved predictive performance compared to existing models.

New self-exciting random evolutions (SEREs) for modeling traffic and transport processes.

problem Modeling self-exciting and clustering effects in traffic and transport processes.
method Introducing a new process based on a superposition of a Markov chain and a Hawkes process, and constructing self-exciting random evolutions (SEREs).
result Developed new models and limit theorems for SEREs, including averaging and diffusion approximation.

Deep learning outperforms traditional methods in estimating OU process parameters.

problem Parameter estimation of the Ornstein-Uhlenbeck process is challenging.
method Used a multi-layer perceptron to estimate OU process parameters compared to traditional methods like Kalman filter and maximum likelihood estimation.
result Deep learning method outperforms traditional methods in parameter estimation of the OU process.