Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,051 papers · 148 categories

Trend · papers per month

2545097631,017 · Jun 202019922001200920172026
48 results for performance curve

Learn2Evaluate uses learning curves to estimate high-dimensional prediction performance.

problem Estimating test performance in high-dimensional data settings is challenging.
method Learn2Evaluate uses learning curves to estimate test performance at the total sample size.
result Learn2Evaluate provides a lower confidence bound for performance estimation.

Learning curves show more data doesn't always improve performance.

problem The surprising finding that more data doesn't always lead to better generalization performance.
method A survey of learning curves, focusing on those that indicate more data doesn't necessarily improve performance.
result Learning curves can show that more data doesn't always lead to better generalization performance.

LC-PFN predicts learning curve performance more accurately and faster than MCMC.

problem Bayesian extrapolation of learning curves is computationally expensive and overly restrictive.
method Prior-Data Fitted Neural Networks (PFNs) for approximate Bayesian inference.
result LC-PFN outperforms MCMC in accuracy and is significantly faster.

Probabilistic models predict neural network performance across varying hyperparameters.

problem Predicting neural network performance with different hyperparameters.
method Probabilistic models based on random forests and Bayesian recurrent neural networks.
result Models outperform state-of-the-art hyperparameter optimization methods.

New model predicts neural network performance from early training epochs, incorporating architecture impact.

problem Predicting neural network performance from early training epochs, neglecting architecture impact.
method Architecture-aware graph ordinary differential equation model.
result Model outperforms state-of-the-art methods for MLP and CNN learning curves.

Yield curve forecasting is an important problem in finance. In this work we explore the use of Gaussian Processes in conjunction with a dynamic modeling strategy, much like the Kalman Filter, to model the yield curve. Gaussian Processes have been successfully applied to model functional data in a variety of application…

2017-03-04abs ↗pdf ↗

Automatically explores geometric loci of curves using software networking.

problem Exploring hyperbolisms and geometric loci of plane curves.
method Parametric equations, Groebner bases, and elimination for deriving polynomial equations.
result Derives new constructions of lemniscates and other geometric loci.

The paper proposes using theoretical ROC curves to categorize classifier responses.

problem The lack of explicit probability distributions for classifier responses in machine learning.
method Fit beta distributions to classifier responses and use them to categorize responses into different classes.
result Established a categorization of classifier responses into classes with different ROC curve extremal behaviors.

Paper estimates optimal ROC curve arc length and AUC, improving classification performance.

problem Estimating optimal ROC curve arc length and AUC in imbalanced binary classification.
method Expresses arc length and AUC as variational objectives, estimating using positive and negative samples.
result Proposed classification procedure maximizes an approximate lower bound of maximal AUC.

No Shimura-Teichmüller curves found in genus 5.

problem Classifying Shimura-Teichmüller curves in genus 5.
method Utilized the equivalence of Shimura-Teichmüller curves to having completely degenerate Kontsevich-Zorich spectrum, and implemented a computer search to exclude remaining cases.
result No Shimura-Teichmüller curves exist in genus 5.

The paper studies multi-curve interest rate models and their consistency and finite-dimensional realizations.

problem Consistency and existence of finite-dimensional realizations for multi-curve interest rate models.
method Geometric approach, characterizing consistency and existence of finite-dimensional realizations for multi-curve models.
result Characterization of consistency and existence of finite-dimensional realizations for multi-curve models.

A meta-learning approach for efficient algorithm selection in budget-limited scenarios.

problem Efficiently selecting the best-performing machine learning algorithm with limited computational resources.
method A Markov Decision Process framework where an agent decides whether to train, wake up, or start new algorithms based on partial learning curves.
result Meta-learning from learning curves improves algorithm selection, especially when learning curves do not intersect frequently.

New method estimates individual dose-response curves for any number of treatments.

problem Estimating individual dose-response curves for varied exposures.
method Neural network approach for learning counterfactual representations.
result Set a new state-of-the-art in estimating individual dose-response curves.

GLMM trees identify subgroups with different growth patterns in longitudinal data.

problem Identifying subgroups with distinct growth trajectories in longitudinal studies.
method Extended GLMM trees for longitudinal data.
result Extended GLMM trees outperform other methods in accuracy and speed.

This work studies learning curves for revenue maximization algorithms.

problem Understanding the performance of revenue-maximizing algorithms as they learn from more data.
method Initiates the study of learning curves for revenue maximization, providing a near-complete characterization of their rate of decay.
result Learning curves for revenue maximization can decay arbitrarily slowly or almost exponentially fast, depending on the distribution and optimal revenue.

DAC provides interpretable feature importance curves for tree ensembles.

problem Challenges in understanding feature importance and prediction in tree ensembles.
method Disentangled Attribution Curves (DAC) method to visualize feature importance.
result DAC improves interpretability of tree ensembles while maintaining accuracy.

Comparison of decision curve analysis and cost curves for model evaluation.

problem Evaluating classification performance across different operating contexts.
method Comparison of Decision Curve Analysis (DCA) and Cost Curves.
result DCA and Cost Curves are closely related, with Brier curves being more generally applicable.

Confidence bands for tuning curves improve hyperparameter comparison in NLP.

problem Ambiguity in comparing hyperparameter tuning methods.
method Constructs exact, simultaneous, and distribution-free confidence bands for tuning curves.
result Confidence bands provide a robust basis for comparing methods rigorously.

The study improves the assessment of fairness in face recognition using ROC curves and statistical guarantees.

problem Improving the assessment of fairness in face recognition systems.
method Proves asymptotic guarantees for empirical ROC curves and fairness metrics, and introduces a recentering technique to avoid bootstrap pitfalls.
result Demonstrates the practical relevance of the methods for assessing fairness in face recognition systems.

HAMLET optimizes algorithm selection for machine learning tasks.

problem Limited time budgets and computational resources make traditional bandit approaches ineffective for automated algorithm selection.
method HAMLET incorporates learning curve extrapolation and time-awareness to select machine learning algorithms.
result HAMLET variants outperform other bandit-based strategies in experiments with recorded hyperparameter tuning traces.

ABROCA assesses algorithmic bias, revealing skewed distributions that inflate results.

problem Detecting nuanced performance differences in classifier fairness.
method Study of ABROCA metric's statistical properties under various conditions.
result ABROCA distributions are skewed, inflating results by chance in imbalanced classes.

It is given the diffeomorphism classification on generic singularities of tangent varieties to curves with arbitrary codimension in a projective space. The generic classifications are performed in terms of certain geometric structures and differential systems on flag manifolds, via several techniques in differentiable …

2012-01-13abs ↗pdf ↗

In this paper we study the shape space of curves with values in a homogeneous space M=G/KM = G/K, where GG is a Lie group and KK is a compact Lie subgroup. We generalize the square root velocity framework to obtain a reparametrization invariant metric on the space of curves in MM. By identifying curves in MM with thei…

2017-06-09abs ↗pdf ↗

Framework for transferring discount curve estimates across fixed-income product classes.

problem Challenges in estimating discount curves from sparse or noisy data.
method Proposes a vector-valued kernel ridge regression (KR) framework with economic regularization.
result Transfer learning tightens confidence intervals and improves extrapolation performance.

Entrocraft addresses RL performance saturation in LLMs by customizing entropy curves.

problem Performance saturation in RL algorithms for LLMs.
method Entrocraft uses rejection sampling to bias advantage distributions for customized entropy schedules.
result Entrocraft significantly improves generalization, output diversity, and long-term training in 4B models.

DynAEsti models ability as a time-varying curve, improving IRT for dynamic settings.

problem Traditional IRT models assume static ability; new approach models dynamic ability.
method DynAEsti augments traditional IRT Expectation Maximization with CurvFiFE for curve fitting.
result DynAEsti successfully recovers human performance dynamics in golf.

Efficiently models learning curves using Gaussian processes with latent Kronecker structure.

problem Joint modeling of machine learning model performance across hyper-parameters and training progress.
method Imposes latent Kronecker structure to leverage efficient product kernels and handle missing values.
result Matches the performance of a Transformer on a learning curve prediction task.