Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

51101152202 · May 202619922001200920172026
48 results for bimodal target

Proposes new loss functions for better handling bimodal predictive uncertainty.

problem Bimodal predictive uncertainty in machine learning models.
method Family of distribution-aware loss functions integrating normalized RMSE with Wasserstein and Cramér distances.
result Proposed loss functions reduce predictive uncertainty estimation error by 45% on complex bimodal datasets.

New methods improve prediction regions for high-dimensional data.

problem Creating effective prediction regions for high-dimensional data.
method CD-split and HPD-split methods that combine split method and data-driven partition.
result CD-split and HPD-split converge to oracle highest predictive density set and satisfy local and asymptotic conditional validity.

We find stationary distributions in a financial model with trends and mean-reversion.

problem Financial markets with competing trends and mean-reversion.
method Analytical derivation of stationary distributions in various noise and feedback regimes.
result The distributions are unimodal Gaussians in small noise, small feedback limits, but can be bimodal for stronger trends.

A new method generates synthetic data with realistic marginal distributions.

problem Generating synthetic data with bimodal and skewed marginal distributions.
method Pre-transformation variational autoencoders (PTVAEs) with separate parameter optimization for each variable.
result PTVAEs outperform other methods in generating synthetic data with bimodal and skewed distributions.

This study examines how earnings announcements affect option volatility and pricing.

problem The impact of earnings announcements on option volatility and pricing.
method Analysis of extremely short-term options data to study bimodality and concavity in IV curves.
result Investors pay a premium to hedge against extreme volatility during earnings announcements in the presence of concave IV smiles.

Speech emotion recognition is a challenging task and an important step towards more natural human-machine interaction. We show that pre-trained language models can be fine-tuned for text emotion recognition, achieving an accuracy of 69.5% on Task 4A of SemEval 2017, improving upon the previous state of the art by over …

2019-11-29abs ↗pdf ↗

This work improved clustering methods by analyzing various datasets and dendrograms.

problem Avoiding false positives in clustering, especially for unimodal and bimodal data.
method Applied agglomerative clustering methods (single, average, median, complete, centroid, Ward's) to various datasets.
result Many methods detected two clusters in unimodal data, with single-linkage being more resilient.

Deep networks achieve linear separability through progressive folding of data in higher dimensions.

problem How feed-forward networks achieve linear separability for classification tasks.
method Progressive folding of the data manifold in unoccupied higher dimensions.
result The folding operation allows efficient solutions by providing access to arbitrary regions in the distribution.

Filtering data with a pre-trained model improves multimodal contrastive learning performance.

problem Improving the quality of internet-scale multimodal datasets.
method Characterized the performance of filtered contrastive learning under a bimodal data generation model.
result Data filtering using a pre-trained model reduces contrastive learning error by a factor of η\sqrt{η} in the large ηη regime.

Deep learning speeds up pressure prediction in carbon storage reservoirs.

problem Accurately forecasting reservoir pressure in geologic carbon storage projects with sparse well data.
method Combining InSAR surface displacement data with deep learning and data assimilation techniques.
result Workflow can predict reservoir pressure with high efficiency and uncertainty quantification.

Improved financial market calibration reveals large excess volatility.

problem Large excess volatility in financial markets.
method Extended Chiarella model to handle long-term value drifts, calibrated on multiple asset classes.
result Large excess volatility (factor ≈ 4 for stock indices) and bimodal mispricing distribution.

A simple spin system is constructed to simulate dynamics of asset prices and studied numerically. The outcome for the distribution of prices is shown to depend both on the dimension of the system and the introduction of price into the link measure. For dimensions below 2, the associated risk is high and the price distr…

2014-08-01abs ↗pdf ↗

Most speech recognition tasks pertain to mapping words across two modalities: acoustic and orthographic. In this work, we suggest learning encoders that map variable-length, acoustic or phonetic, sequences that represent words into fixed-dimensional vectors in a shared latent space; such that the distance between two w…

2019-08-01abs ↗pdf ↗

A new method normalizes flow mixtures for better inference across different data types.

problem Inference failure across diverse posterior geometries in normalizing flows.
method Introduces a two-stage framework with a stable global weighting mechanism based on sEMA.
result Achieves consistent NLL improvements and stable weight trajectories over baselines.

The two phase behavior in financial markets actually means the bifurcation phenomenon, which represents the change of the conditional probability from an unimodal to a bimodal distribution. In this paper, the bifurcation phenomenon in Hang-Seng index is carefully investigated. It is observed that the bifurcation phenom…

2007-12-30abs ↗pdf ↗

AMF-VI uses adaptive mixtures of flows for robust VI across diverse distributions.

problem Inconsistent behavior of single-flow models across different distributions.
method Sequential expert training of individual flows and adaptive global weight estimation via likelihood-driven updates.
result AMF-VI achieves lower negative log-likelihood and stable gains in transport metrics across various posterior families.

New model corrects bias in crowdsourced ratings for diverse items.

problem Bias and noise in crowdsourced ratings for training data.
method Bayesian rating model with item-level effects for difficulty, discriminativeness, and guessability.
result New model avoids bias in training data, improving model goodness of fit.

Much research has been conducted arguing that tipping points at which complex systems experience phase transitions are difficult to identify. To test the existence of tipping points in financial markets, based on the alternating offer strategic model we propose a network of bargaining agents who mutually either coopera…

2015-09-16abs ↗pdf ↗

A new model that combines economic growth rate fluctuations at the microscopic and macroscopic level is presented. At the microscopic level, firms are growing at different rates while also being exposed to idiosyncratic shocks at the firm and sector level. We describe such fluctuations as independent Lévy-stable fluctu…

2017-08-26abs ↗pdf ↗

Improved simulation of phase transitions using hierarchical autoregressive networks.

problem Simulating phase transitions in complex systems.
method Hierarchical Autoregressive Neural (HAN) network sampling algorithm.
result Significant improvement in statistical uncertainty compared to the Wolff cluster algorithm.

Human language is a rich multimodal signal consisting of spoken words, facial expressions, body gestures, and vocal intonations. Learning representations for these spoken utterances is a complex research problem due to the presence of multiple heterogeneous sources of information. Recent advances in multimodal learning…

2019-05-14abs ↗pdf ↗

Cross-sectional signatures of market panic were recently discussed on daily time scales in [1], extended here to a study of cross-sectional properties of stocks on intra-day time scales. We confirm specific intra-day patterns of dispersion and kurtosis, and find that the correlation across stocks increases in times of …

2010-10-23abs ↗pdf ↗

Detects corruption in agentic models during execution.

problem Inconsistent context, retrieval errors, or adversarial inputs corrupt intermediate steps of reasoning chains.
method Analyzes token graphs induced by attention and computes spectral statistics to emit accept/reject signals.
result A single threshold on the high frequency energy ratio optimally detects context inconsistency in agentic models.

Proposes a new normalization method for deep neural networks in financial forecasting.

problem Deep neural networks are sensitive to input variable range and prone to numerical issues, especially with financial time-series.
method Bilinear input normalization method that handles high-frequency financial time-series without expert knowledge.
result Significant improvements in forecasting future stock price dynamics over other normalization techniques.

Proposes a flexible framework for implied volatility surfaces with random parameters.

problem Inconsistent calibration of parametric implied volatility models when market volatility deviates from the model's regime.
method Introduces random coefficients for parametric implied volatility formulas, preserving analytic flexibility and efficiency.
result Demonstrates improved modeling of implied volatility curves, especially for short-term options and earnings announcements.

Suppose one buys two very similar stocks and is curious about how much, after some time T, one of them will contribute to the overall asset, expecting, of course, that it should be around 1/2 of the sum. Here we examine this question within the classical Black and Scholes (BS) model, focusing on the evolution of the pr…

2010-05-11abs ↗pdf ↗

Optimizes contrastive learning with individualized temperatures for better performance on imbalanced datasets.

problem The common practice of using a global temperature parameter ignores the varying semantic similarity across different anchor data.
method Proposes a new robust contrastive loss inspired by distributionally robust optimization (DRO) and an efficient stochastic algorithm for automatic temperature individualization.
result Our method automatically learns a suitable temperature for each sample, improving performance on imbalanced datasets.