AR model forecasts partially observed dynamical time series by estimating evolution function and imputing missing variables.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
We consider the classical optimal dividends problem under the Cramér-Lundberg model with exponential claim sizes subject to a constraint on the time of ruin. We introduce the dual problem and show that the complementary slackness conditions are satisfied, thus there is no duality gap. Therefore the optimal value functi…
The paper proves geometric and spectral alignment for deep neural networks.
An augmented Lagrangian (AL) can convert a constrained optimization problem into a sequence of simpler (e.g., unconstrained) problems, which are then usually solved with local solvers. Recently, surrogate-based Bayesian optimization (BO) sub-solvers have been successfully deployed in the AL framework for a more global …
Finite-time queue peaks in stochastic networks have logarithmic scaling after geometric thresholds.
EDSVM uses elite observations to guide SVM classification.
We introduce a longevity feature to the classical optimal dividend problem by adding a constraint on the time of ruin of the firm. We extend the results in \cite{HJ15}, now in context of one-sided Lévy risk models. We consider de Finetti's problem in both scenarios with and without fix transaction costs, e.g. taxes. We…
We introduce a rich model for multi-objective clustering with lexicographic ordering over objectives and a slack. The slack denotes the allowed multiplicative deviation from the optimal objective value of the higher priority objective to facilitate improvement in lower-priority objectives. We then propose an algorithm …
Support Vector Machine (SVM) is an efficient classification approach, which finds a hyperplane to separate data from different classes. This hyperplane is determined by support vectors. In existing SVM formulations, the objective function uses L2 norm or L1 norm on slack variables. The number of support vectors is a me…
A number of machine learning (ML) methods have been proposed recently to maximize model predictive accuracy while enforcing notions of group parity or fairness across sub-populations. We propose a desirable property for these procedures, slack-consistency: For any individual, the predictions of the model should be mono…
Optimization with inequality constraints using embedded gradient vector field method
Storytelling algorithms aim to 'connect the dots' between disparate documents by linking starting and ending documents through a series of intermediate documents. Existing storytelling algorithms are based on notions of coherence and connectivity, and thus the primary way by which users can steer the story construction…
DiffSlack learns neural networks with nonlinear constraints via learnable slack variables.
New analysis reveals gaps in selective classifiers, guiding improvements.
Improved robustness of machine learning models with controlled Lipschitz constants.
Study bandwidth-limited training and inference of language models.
We aim to predict and explain service failures in supply-chain networks, more precisely among last-mile pickup and delivery services to customers. We analyze a dataset of 500,000 services using (1) supervised classification with Random Forests, and (2) Association Rules. Our classifier reaches an average sensitivity of…
Tail-Safe hedging uses reinforcement learning with a safety layer to manage financial risks.
The exact nonnegative matrix factorization (exact NMF) problem is the following: given an -by- nonnegative matrix and a factorization rank , find, if possible, an -by- nonnegative matrix and an -by- nonnegative matrix such that . In this paper, we propose two heuristics for exac…
Proposes a method for evaluating multiple dimensions of organizational effectiveness using DEA.
Graph neural networks optimize radio resource management policies for wireless networks.
Efficient dispatching rule in manufacturing industry is key to ensure product on-time delivery and minimum past-due and inventory cost. Manufacturing, especially in the developed world, is moving towards on-demand manufacturing meaning a high mix, low volume product mix. This requires efficient dispatching that can wor…
Algorithm stabilizes queues in asymmetric systems with unknown service rates.
We present algorithms for efficiently learning regularizers that improve generalization. Our approach is based on the insight that regularizers can be viewed as upper bounds on the generalization gap, and that reducing the slack in the bound can improve performance on test data. For a broad class of regularizers, the h…
Empirical risk minimization frequently employs convex surrogates to underlying discrete loss functions in order to achieve computational tractability during optimization. However, classical convex surrogates can only tightly bound modular loss functions, sub-modular functions or supermodular functions separately while …
The explosion of time series data in recent years has brought a flourish of new time series analysis methods, for forecasting, clustering, classification and other tasks. The evaluation of these new methods requires either collecting or simulating a diverse set of time series benchmarking data to enable reliable compar…
Previous studies indicate that nonlinear properties of Gaussian time series with long-range correlations, , can be detected and quantified by studying the correlations in the magnitude series , i.e., the ``volatility''. However, the origin for this empirical observation still remains unclear, and the exact …
Research into time series classification has tended to focus on the case of series of uniform length. However, it is common for real-world time series data to have unequal lengths. Differing time series lengths may arise from a number of fundamentally different mechanisms. In this work, we identify and evaluate two cla…
Modeling regime shifts in co-evolving time series with interactions and time-dependency.
We provide the proof that the space of time series data is a Kolmogorov space with -separation axiom using the loop space of time series data. In our approach we define a cyclic coordinate of intrinsic time scale of time series data after empirical mode decomposition. A spinor field of time series data comes fro…
Capturing the dynamical properties of time series concisely as interpretable feature vectors can enable efficient clustering and classification for time-series applications across science and industry. Selecting an appropriate feature-based representation of time series for a given application can be achieved through s…
Overview of high-dimensional time series regression methods.
Improved prediction of hierarchical time series using structured regularization.
Time series motifs play an important role in the time series analysis. The motif-based time series clustering is used for the discovery of higher-order patterns or structures in time series data. Inspired by the convolutional neural network (CNN) classifier based on the image representations of time series, motif diffe…
Introduces a new benchmark for time series extrinsic regression.
In this paper, we present a new approach to time series forecasting. Time series data are prevalent in many scientific and engineering disciplines. Time series forecasting is a crucial task in modeling time series data, and is an important area of machine learning. In this work we developed a novel method that employs …
Few-shot learning improves time-series forecasting with limited data.
Meta-learning for Koopman spectral analysis with short time-series data.
Transformers improve time series modeling by capturing long-range dependencies.
Archive of 20 time series datasets for forecasting evaluation.
Multidimensional time series are sequences of real valued vectors. They occur in different areas, for example handwritten characters, GPS tracking, and gestures of modern virtual reality motion controllers. Within these areas, a common task is to search for similar time series. Dynamic Time Warping (DTW) is a common di…
theft package simplifies feature extraction for time series analysis in R.
Method summarizes and predicts time series data for COVID-19 cases and deaths.
Research into the classification of time series has made enormous progress in the last decade. The UCR time series archive has played a significant role in challenging and guiding the development of new learners for time series classification. The largest dataset in the UCR archive holds 10 thousand time series only; w…
Feature-based time series representations have attracted substantial attention in a wide range of time series analysis methods. Recently, the use of time series features for forecast model averaging has been an emerging research focus in the forecasting community. Nonetheless, most of the existing approaches depend on …
New deep probabilistic model handles missing data in time series forecasting.
Paper introduces novel distances for clustering ordinal time series.
A new framework for generating predictive features in noisy multivariate time series.