Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,291 papers · 148 categories

Trend · papers per month

306191121 · May 202619922001200920182026
48 results for threshold identification

New algorithm identifies good arms with fewer samples when thresholds are close.

problem Good arm identification in bandit problems with small threshold gaps.
method Proposes lil'HDoC algorithm to improve GAI under small threshold gaps.
result Sample complexity of first λ output arm is nearly identical to HDoC algorithm when thresholds are close.

Study lenient regret and good-action identification in Gaussian process bandits.

problem Optimizing function values above a certain threshold in Gaussian process bandits.
method Study lenient regret notions and introduce algorithms for finding good actions.
result Upper and lower bounds on lenient regret for GP-UCB and elimination algorithms.

Paper improves gas species identification in complex mixtures using neural networks.

problem Identifying gas species in multi-gas mixtures with high accuracy.
method Multi-label neural networks with optimal thresholding for IR spectroscopy.
result Optimal thresholding improves classification performance over conventional methods.

The paper analyzes methods for sparse Bayesian regression in nonlinear system identification.

problem Learning sparse models in Bayesian regression with nonlinear applications.
method Two classes of methods: regularization and thresholding based, built on automatic relevance determination (ARD).
result Analytical demonstration of favorable performance with sparse solutions in linear problems.

Paper proposes a method to identify optimal threshold for stock market networks.

problem Challenges in identifying the optimal threshold for reliable stock network construction.
method Dynamic consistence between threshold network and stock market, optimal threshold maximized by consistence function.
result Optimal threshold value of 0.28 for stocks in S&P 500 Index.

Donors who defer their donations volunteer less in the future.

problem Volunteer labor can be less beneficial to charities than its costs.
method Regression discontinuity design with a procedure to handle manipulation.
result Donor manipulation invalidates standard regression discontinuity design, but a new method provides partial identification bounds.

This work optimizes identifying good arms in nonparametric multi-armed bandits.

problem Efficiently identifying arms with high means in nonparametric settings.
method Combining reward-maximizing sampling with a nonparametric sequential test for anytime-valid labeling.
result Achieves minimax optimal stopping times for identifying arms above a threshold.

Optimal top-2 method improves best arm identification with reduced error.

problem Identifying the arm with the highest mean in a set of arms.
method A novel top-2 algorithm that pulls the empirical best arm with probability β and the challenger arm otherwise.
result The proposed algorithm matches the information theoretic lower bound on sample complexity as δ approaches 0.

New algorithm achieves near optimal sample complexity for 1-identification problem.

problem Determining if an arm's mean reward is at least a known threshold with high probability.
method Design of Sequential-Exploration-Exploitation (SEE) algorithm with non-asymptotic analysis.
result Achieves near optimality in sample complexity, matching upper and lower bounds up to a polynomial logarithmic factor.

The study examines when to trust confidence thresholding in pseudo-labelling regression.

problem Calibrated probabilities from classifiers used for pseudo-labelling need careful handling to avoid bias in downstream regression.
method Developed a diagnostic apparatus to predict and bound the bias induced by confidence thresholding, derived a closed-form expression for the attenuation bias.
result The bias can be predicted from the residual score variance VV^{*}, motivating a structural separation between classifier features and downstream controls.

VA-LUCB identifies best arm with variance constraint, achieving optimal sample complexity.

problem Identifying the best arm with variance constraint under fixed confidence.
method Parameter-free algorithm VA-LUCB, analyzing sample complexity and proving lower bounds.
result Optimal sample complexity up to a logarithmic factor in HVAH_{VA}, demonstrated by experiments.

The labeled stochastic block model is a random graph model representing networks with community structure and interactions of multiple types. In its simplest form, it consists of two communities of approximately equal size, and the edges are drawn and labeled at random with probability depending on whether their two en…

2015-02-11abs ↗pdf ↗

Study improves early warning models for currency and stock market crises.

problem Predicting currency and stock market crises.
method Synthetic review and comparison of early warning models, focusing on crisis identifications and predictive models.
result SWARCH model with elastic thresholding methodology most accurately classifies crisis observations.

This work improves SINDy-type algorithms for system identification using score-guided dictionary selection.

problem Improving accuracy and interpretability in dynamical system identification.
method Score-guided library selection to refine dictionary terms in sparse regression.
result Score-guided methods enhance SINDy's robustness in discovering governing equations.

Sharp large deviations and Gibbs conditioning for portfolio credit risk models.

problem Analyzing the risk of default in financial portfolios with dependent factors.
method Sharp large deviation estimates and conditional Bahadur-Rao estimates for threshold models with diverging latent factors.
result Conditioned on a large exceedance event, default indicators become asymptotically i.i.d., and loss-given-default is exponentially tilted.

New techniques improve the accuracy of identifying nonlinear systems from noisy data.

problem Identifying nonlinear dynamical systems from noisy state measurements.
method Comparative study of local and global smoothing techniques to denoise state measurements and improve sparse regression methods.
result Global smoothing methods outperform local methods in improving the accuracy of governing equation recovery.

Extends extreme value mixture models to identify changepoints in financial extreme regimes.

problem Inference over financial extreme regimes is affected by threshold choice.
method Extends extreme value mixture models to account for distributional extreme changepoints using MCMC algorithms.
result Inclusion of different extreme regimes improves financial applications compared to static and dynamic approaches.

Unified algorithm for efficient pure exploration using dual variables.

problem Efficiently achieving a specific goal through adaptive experimentation.
method Introducing dual variables to derive optimal allocation conditions, leading to Information-Directed Selection.
result Top-two Thompson sampling attains asymptotic optimality for Gaussian best-arm identification.

New method identifies network structure without regularization for sparse teacher couplings.

problem Identifying network structure in inverse Ising problems with model mismatch.
method Ridge linear regression with two-stage estimator.
result Perfect identification of network structure possible without regularization for sparse teacher couplings.

Study on the limits of learning HMM parameters under various conditions.

problem Understanding the conditions under which hidden Markov model parameters can be learned.
method Nonasymptotic minimax upper and lower bounds, thresholds analysis.
result Nonasymptotic minimax bounds match up to constants, showing learnable thresholds.

WSINDy algorithm proves robust to noise in identifying differential equations.

problem Identifying differential equations from noisy data.
method Weak-form sparse identification of nonlinear dynamics (WSINDy) algorithm.
result WSINDy is asymptotically consistent for a wide class of models, including Navier-Stokes and Kuramoto-Sivashinsky equations.

A two-phase algorithm identifies the best arm in sparse linear bandits with fixed budget.

problem Best arm identification in sparse linear bandits with limited budget.
method Lasso and Optimal-Design (Lasso-OD) based linear best-arm identification.
result Lasso-OD achieves significant performance improvement for sparse and high-dimensional linear bandits.

Researchers identify critical protein residues using advanced graph theory.

problem Identifying essential residues in proteins for function.
method Learning Random Geometric Graphs (RGG) with Cramer's V correlation and organic thresholding.
result Advanced RGG methods accurately identify critical residues compared to existing techniques.

DeepSupp detects financial support levels using attention mechanisms.

problem Traditional SR identification methods fail to adapt to modern markets.
method Multi-head attention mechanisms, dynamic correlation matrices, DBSCAN clustering.
result DeepSupp outperforms six baseline methods across six financial metrics.

Improved reliability of machine learning predictions using variational auto-encoders.

problem Individual unreliability of machine learning models.
method Modified variational auto-encoders to identify a low-dimensional space for reliable classification.
result Improved reliability of predictions and robust identification of adversarial samples.

Study on inferring dynamic communities and links in networks with memory.

problem Inferring communities and links in dynamic networks with memory.
method Maximum likelihood inference from single snapshot observations, analytical and numerical analysis.
result Link persistence makes community detection harder, while community persistence makes it easier.

M-learner estimates treatment effects in mediation models with subgroup identification.

problem Estimating heterogeneous treatment effects in mediation models.
method Four-step procedure: compute conditional effects, construct distance matrix, apply tSNE and K-means clustering, refine clusters.
result Validates robustness and effectiveness in real-world dataset.

Paper tackles adaptive sampling for identifying largest gaps between distributions.

problem Adaptive sampling from K distributions to identify the largest gap between any two adjacent means.
method Proposes elimination and UCB-style algorithms, showing minimax optimality.
result UCB-style algorithms require 6-8x fewer samples than non-adaptive sampling.

This paper presents the first theoretical results showing that stable identification of overcomplete μμ-coherent dictionaries ΦRd×KΦ\in \mathbb{R}^{d\times K} is locally possible from training signals with sparsity levels SS up to the order O(μ2)O(μ^{-2}) and signal to noise ratios up to O(d)O(\sqrt{d}). In particular the di…

2014-01-24abs ↗pdf ↗

Study identifies contagion in aggregated defaults despite environmental changes.

problem Identify contagion in aggregated default counts with fluctuating probabilities.
method Compare three contagion mechanisms (Davis-Lo, Torri, Vasicek) under i.i.d. and hierarchical specifications.
result Threshold contagion is largely absorbed into environmental heterogeneity, while cumulative contagion leaves a persistent signature.