Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,657 papers · 148 categories

Trend · papers per month

61121182242 · May 202619922001200920172026
48 results for scope localization

Convolutional neural network improves assertion detection in multi-label clinical text.

problem Detecting assertions in multi-label clinical text with rich descriptions.
method Developed a CNN architecture for multi-label scope detection.
result At least 12% improvement over state-of-the-art on multi-label clinical text.

Many machine learning models, such as logistic regression~(LR) and support vector machine~(SVM), can be formulated as composite optimization problems. Recently, many distributed stochastic optimization~(DSO) methods have been proposed to solve the large-scale composite optimization problems, which have shown better per…

2016-01-30abs ↗pdf ↗

Proposes dynamic model type recommendation for OLP technique.

problem Limited local competence of base-classifiers in uneven data distributions.
method Builds a multi-label meta-classifier to recommend model types based on local data complexity.
result Statistically similar performance to original OLP with fixed base-classifier model.

SCOPE iteratively optimizes sparsity-constrained problems without tuning hyperparameters.

problem Optimizing sparsity-constrained problems in signal processing, statistics, and machine learning.
method SCOPE (Sparsity-Constrained Optimization via sPlicing itEration) replaces gradient steps with a splicing operation guided by the objective value.
result SCOPE achieves linear convergence and superior support recovery performance.

Scoping review finds EEG key in MCI research, identifying ERP/EEG, QEEG, and machine learning.

problem Identifying MCI early and accurately.
method Scoping review with co-occurrence analysis and PAGER framework.
result Main research themes identified: ERP/EEG, QEEG, and EEG-based machine learning.

SCOPE-FE improves feature engineering efficiency for high-dimensional datasets.

problem Expanding and reducing feature space in tabular learning becomes computationally expensive with increased dimensionality.
method SCOPE-FE controls the search space by regulating operator and feature-pair spaces, using OperatorProbing and FeatureClustering.
result SCOPE-FE reduces feature engineering time while maintaining competitive predictive performance.

We study the convergence of the Expectation-Maximization (EM) algorithm for mixtures of linear regressions with an arbitrary number kk of components. We show that as long as signal-to-noise ratio (SNR) is Ω~(k)\tildeΩ(k), well-initialized EM converges to the true regression parameters. Previous results for k3k \geq 3 hav…

2019-05-28abs ↗pdf ↗

Scoping review of EO-ML methods for causal inference in poverty geography.

problem Lack of thorough documentation and best practices for EO-ML methods in causal analysis.
method Comprehensive scoping review cataloging five principal approaches.
result Detailed protocol for integrating EO data into causal analysis.

This text discusses several popular explanatory methods that go beyond the error measurements and plots traditionally used to assess machine learning models. Some of the explanatory methods are accepted tools of the trade while others are rigorously derived and backed by long-standing theory. The methods, decision tree…

2018-10-05abs ↗pdf ↗

Scoping review and benchmarking of synthetic EHR data generation methods.

problem Creating realistic synthetic electronic health records for research and training.
method Conducted a scoping review and benchmarked seven methods on open-source EHR datasets.
result GAN-based methods excel in fidelity and utility, while rule-based methods excel in privacy protection.

Study finds carbon emissions affect stock value, but not bought emissions.

problem Determining if carbon emissions impact stock value and whether this is due to direct or indirect emissions.
method Fixed-effects analysis with propensity score weighting to control for selection bias.
result Firms with higher Scope 1 emissions have a statistically significant positive carbon premium, but Scope 2 emissions do not.

Two new estimators reduce costs and improve accuracy for EHR outcome prediction.

problem Sparse estimate distributions, high computational cost, and high sampling variance in EHR outcome prediction.
method Proposed SCOPE and REACH estimators that leverage next-token probability distributions.
result SCOPE and REACH match Monte Carlo accuracy with token reductions of 2.5-3.4 times and variance guarantees.

Localized Multidirectional Correction improves non-refusal target-response behavior in foundation models.

problem Controlled post-training refusal suppression in routed MoE and hybrid-MoE foundation models.
method Introduce Localized Multidirectional Correction (LoMC), a support-gated intervention framework.
result Substantially improves non-refusal target-response behavior while maintaining general capability under a compact intervention footprint.

Study reduces emissions in portfolios with error-prone emissions data.

problem Portfolio optimization with firm-level emissions intensities measured inaccurately.
method Introduced a scope-specific penalty operator to rescale asset payoffs based on revenue-normalized emissions intensity.
result Reduces average Scope~1 emissions intensity by roughly 92% while maintaining similar Sharpe ratios.

Machine learning predicts liquid water properties from cluster data.

problem Accuracy of bulk properties from machine-learned potentials is limited by training data.
method Local, atom-centred descriptors enable prediction of bulk properties from cluster data.
result Excellent agreement with experimental and theoretical counterparts of liquid water properties.

This paper examines how different loss functions affect neural network features and performance.

problem Investigating which loss function is best for deep neural networks.
method Examining last-layer features of deep networks and drawing inspiration from the Neural Collapse phenomenon.
result All relevant loss functions (CE, LS, FL, MSE) produce equivalent features and similar performance.

Let XX be a compact Kähler manifold. Given a big cohomology class {θ}\{θ\}, there is a natural equivalence relation on the space of θθ-psh functions giving rise to S(X,θ)\mathcal S(X,θ), the space of singularity types of potentials. We introduce a natural pseudometric dSd_{\mathcal {S}} on S(X,θ)\mathcal S(X,θ) that is non-de…

2019-09-02abs ↗pdf ↗

The small-ball method was introduced as a way of obtaining a high probability, isomorphic lower bound on the quadratic empirical process, under weak assumptions on the indexing class. The key assumption was that class members satisfy a uniform small-ball estimate: that Pr(fκfL2)δPr(|f| \geq κ\|f\|_{L_2}) \geq δ for given const…

2017-09-04abs ↗pdf ↗

Machine learning pipelines often rely on optimization procedures to make discrete decisions (e.g., sorting, picking closest neighbors, or shortest paths). Although these discrete decisions are easily computed, they break the back-propagation of computational graphs. In order to expand the scope of learning problems tha…

2020-02-20abs ↗pdf ↗

Defines Learning Analytics' foundational structure and scope.

problem Lack of theoretical foundation in Learning Analytics.
method Proposes an axiomatic theory based on psychological learning and LA methodology.
result Clarifies the epistemological stance of Learning Analytics and its limitations.

SCOPE estimator improves covariance and precision matrix estimation.

problem Estimating covariance and precision matrices accurately.
method Distributionally robust optimization with convex spectral divergence.
result SCOPE estimator reduces spectral bias and improves condition number.

Introduces a new triple coproduct for knots on surfaces, preserving local crossing patterns.

problem Tackles the lack of fine-grained detection in classical cobrackets for local crossing patterns.
method Defines an integer-valued invariant using a coproduct and intersection theory, extending Turaev's cobracket theory.
result Reveals an intrinsic simplicity in the algebraic framework, uniquely determining relations in the word space.

Randomized methods of neural network learning suffer from a problem with the generation of random parameters as they are difficult to set optimally to obtain a good projection space. The standard method draws the parameters from a fixed interval which is independent of the data scope and activation function type. This …

2019-08-11abs ↗pdf ↗

In this paper we characterise the propensity of big capital investments to systematically deliver poor outcomes as "fragility," a notion suggested by Nassim Taleb. A thing or system that is easily harmed by randomness is fragile. We argue that, contrary to their appearance, big capital investments break easily - i.e. d…

2016-03-04abs ↗pdf ↗

Unified theoretical guarantees for distribution-free changepoint detection and testing.

problem Distribution-free changepoint inference with finite-sample validity and consistency.
method Distribution-free changepoint localization using conformal p-values with theoretical guarantees.
result Unified distribution-free guarantees for changepoint detection, localization, and testing.

BAxUS optimizes high-dimensional functions adaptively, avoiding performance degradation and failure.

problem State-of-the-art HDBO methods degrade or fail with increasing dimensions.
method BAxUS uses nested random subspaces to adaptively optimize high-dimensional functions.
result BAxUS outperforms state-of-the-art methods across various applications.

We introduce a measure to quantify ambiguity in deep learning models, improving their reliability.

problem Deep learning models make mistakes on seemingly trivial cases and fail in recognizing what they don't know.
method We define ambiguity based on decision boundaries and convex hulls in feature space, developing a theoretical framework to identify unknowns.
result A single ambiguity measure can detect a significant portion of model mistakes, including adversarial and out-of-distribution inputs.

This summarizes the study of the financial and economic crisis in Europe. The starting questions were: 1) Why do we have a crisis? Unde venis? 2) What will be the outcome? Quo vadis? Here is the reasoning which touches many areas, ranging from financial to politics and from psychology and economy.

2013-05-23abs ↗pdf ↗