Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

137273410546 · Jun 202019922001200920172026
48 results for bimodal distribution

Proposes new loss functions for better handling bimodal predictive uncertainty.

problem Bimodal predictive uncertainty in machine learning models.
method Family of distribution-aware loss functions integrating normalized RMSE with Wasserstein and Cramér distances.
result Proposed loss functions reduce predictive uncertainty estimation error by 45% on complex bimodal datasets.

We find stationary distributions in a financial model with trends and mean-reversion.

problem Financial markets with competing trends and mean-reversion.
method Analytical derivation of stationary distributions in various noise and feedback regimes.
result The distributions are unimodal Gaussians in small noise, small feedback limits, but can be bimodal for stronger trends.

A new method generates synthetic data with realistic marginal distributions.

problem Generating synthetic data with bimodal and skewed marginal distributions.
method Pre-transformation variational autoencoders (PTVAEs) with separate parameter optimization for each variable.
result PTVAEs outperform other methods in generating synthetic data with bimodal and skewed distributions.

This work improved clustering methods by analyzing various datasets and dendrograms.

problem Avoiding false positives in clustering, especially for unimodal and bimodal data.
method Applied agglomerative clustering methods (single, average, median, complete, centroid, Ward's) to various datasets.
result Many methods detected two clusters in unimodal data, with single-linkage being more resilient.

This study examines how earnings announcements affect option volatility and pricing.

problem The impact of earnings announcements on option volatility and pricing.
method Analysis of extremely short-term options data to study bimodality and concavity in IV curves.
result Investors pay a premium to hedge against extreme volatility during earnings announcements in the presence of concave IV smiles.

A simple spin system is constructed to simulate dynamics of asset prices and studied numerically. The outcome for the distribution of prices is shown to depend both on the dimension of the system and the introduction of price into the link measure. For dimensions below 2, the associated risk is high and the price distr…

2014-08-01abs ↗pdf ↗

Deep learning speeds up pressure prediction in carbon storage reservoirs.

problem Accurately forecasting reservoir pressure in geologic carbon storage projects with sparse well data.
method Combining InSAR surface displacement data with deep learning and data assimilation techniques.
result Workflow can predict reservoir pressure with high efficiency and uncertainty quantification.

Improved financial market calibration reveals large excess volatility.

problem Large excess volatility in financial markets.
method Extended Chiarella model to handle long-term value drifts, calibrated on multiple asset classes.
result Large excess volatility (factor ≈ 4 for stock indices) and bimodal mispricing distribution.

The two phase behavior in financial markets actually means the bifurcation phenomenon, which represents the change of the conditional probability from an unimodal to a bimodal distribution. In this paper, the bifurcation phenomenon in Hang-Seng index is carefully investigated. It is observed that the bifurcation phenom…

2007-12-30abs ↗pdf ↗

Speech emotion recognition is a challenging task and an important step towards more natural human-machine interaction. We show that pre-trained language models can be fine-tuned for text emotion recognition, achieving an accuracy of 69.5% on Task 4A of SemEval 2017, improving upon the previous state of the art by over …

2019-11-29abs ↗pdf ↗

Deep networks achieve linear separability through progressive folding of data in higher dimensions.

problem How feed-forward networks achieve linear separability for classification tasks.
method Progressive folding of the data manifold in unoccupied higher dimensions.
result The folding operation allows efficient solutions by providing access to arbitrary regions in the distribution.

Filtering data with a pre-trained model improves multimodal contrastive learning performance.

problem Improving the quality of internet-scale multimodal datasets.
method Characterized the performance of filtered contrastive learning under a bimodal data generation model.
result Data filtering using a pre-trained model reduces contrastive learning error by a factor of η\sqrt{η} in the large ηη regime.

New methods improve prediction regions for high-dimensional data.

problem Creating effective prediction regions for high-dimensional data.
method CD-split and HPD-split methods that combine split method and data-driven partition.
result CD-split and HPD-split converge to oracle highest predictive density set and satisfy local and asymptotic conditional validity.

AMF-VI uses adaptive mixtures of flows for robust VI across diverse distributions.

problem Inconsistent behavior of single-flow models across different distributions.
method Sequential expert training of individual flows and adaptive global weight estimation via likelihood-driven updates.
result AMF-VI achieves lower negative log-likelihood and stable gains in transport metrics across various posterior families.

Modified lognormal distribution with flexible tails for skewed data.

problem Skewed and fat-tailed data in natural and engineering datasets.
method Developed a family of three-parameter non-Gaussian probability density functions based on generalized kappa-exponential and kappa-logarithm functions.
result Closed-form analytic expressions for statistical functions and maximum-likelihood estimation.

One aim of data mining is the identification of interesting structures in data. For better analytical results, the basic properties of an empirical distribution, such as skewness and eventual clipping, i.e. hard limits in value ranges, need to be assessed. Of particular interest is the question of whether the data orig…

2019-08-15abs ↗pdf ↗

New model corrects bias in crowdsourced ratings for diverse items.

problem Bias and noise in crowdsourced ratings for training data.
method Bayesian rating model with item-level effects for difficulty, discriminativeness, and guessability.
result New model avoids bias in training data, improving model goodness of fit.

Much research has been conducted arguing that tipping points at which complex systems experience phase transitions are difficult to identify. To test the existence of tipping points in financial markets, based on the alternating offer strategic model we propose a network of bargaining agents who mutually either coopera…

2015-09-16abs ↗pdf ↗

A new model that combines economic growth rate fluctuations at the microscopic and macroscopic level is presented. At the microscopic level, firms are growing at different rates while also being exposed to idiosyncratic shocks at the firm and sector level. We describe such fluctuations as independent Lévy-stable fluctu…

2017-08-26abs ↗pdf ↗

Improved simulation of phase transitions using hierarchical autoregressive networks.

problem Simulating phase transitions in complex systems.
method Hierarchical Autoregressive Neural (HAN) network sampling algorithm.
result Significant improvement in statistical uncertainty compared to the Wolff cluster algorithm.

Cross-sectional signatures of market panic were recently discussed on daily time scales in [1], extended here to a study of cross-sectional properties of stocks on intra-day time scales. We confirm specific intra-day patterns of dispersion and kurtosis, and find that the correlation across stocks increases in times of …

2010-10-23abs ↗pdf ↗

Most speech recognition tasks pertain to mapping words across two modalities: acoustic and orthographic. In this work, we suggest learning encoders that map variable-length, acoustic or phonetic, sequences that represent words into fixed-dimensional vectors in a shared latent space; such that the distance between two w…

2019-08-01abs ↗pdf ↗

Populations of species in ecosystems are often constrained by availability of resources within their environment. In effect this means that a growth of one population, needs to be balanced by comparable reduction in populations of others. In neutral models of biodiversity all populations are assumed to change increment…

2015-03-02abs ↗pdf ↗

Bayesian model enhances phenotype discovery in asthma EHRs.

problem Lack of interpretability in unsupervised learning phenotyping of EHR data.
method Operationalized a Bayesian latent class framework with clinical knowledge priors.
result Identified an asthma sub-phenotype with elevated eosinophil levels and allergy markers.

Model shows how capital accumulation can lead to poverty traps and well-being states.

problem Capital accumulation and its effects on poverty and well-being.
method Stochastic Solow growth model with sigmoidal saving fraction and bimodal steady state distribution.
result Existence of poverty trap with fluctuation-driven transitions between poverty and well-being states.