This study examines the interaction between CDS and stock indices, revealing significant short and long-term impacts.
problem Understanding the interaction between Credit Default Swaps (CDS) and national stock indices.
method ARDL technique applied to analyze short and long-run interactions between BIST-100 index and CDS prices over a specific period.
result The study finds that changes in CDS and BIST-100 index prices have significant impacts on each other, with long-term effects being more pronounced.
Study enhances financial forecasting with machine learning and fuzzy MCDM.
problem Increasing financial uncertainty and market complexity.
method Integrates machine learning (XGBoost, LSTM, GNN) and intuitionistic fuzzy MCDM.
result High forecasting accuracy with low MAPE and narrow confidence intervals.
Study uses XAI and transformers for stock price prediction of top 100 BIST banks.
problem Enhancing interpretability and accuracy of stock price predictions.
method Combines transformer-based time series models with XAI techniques.
result Transformer models show strong predictive capabilities and provide feature transparency.
QLSTM outperforms LSTM in predicting KSE 100 index movements.
problem Predicting stock market movement in uncertain economic conditions.
method Used LSTM and QLSTM models on monthly data of economic indicators.
result QLSTM provided more accurate predictions of KSE 100 index values.
Compact method proves Brown-York mass positivity and connects to major conjectures.
problem Proving positivity of Brown-York's mass and its connections to conjectures.
method Compact approach to proving mass positivity and exploring connections.
result Proved the positivity of Brown-York's mass and its relation to conjectures.
The κ-generalised distribution fits daily stock returns well.
problem Stock returns are often heavy-tailed, not normally distributed.
method Used the κ-generalised distribution with a Monte-Carlo goodness of fit test. result The κ-generalised distribution fits historic daily stock returns well for a significant proportion of analyzed stocks. Persistence is studied in a financial context by mapping the time evolution of the values of the shares quoted on the London Financial Times Stock Exchange 100 index (FTSE 100) onto Ising spins. By following the time dependence of the spins, we find evidence for power law decay of the proportion of shares that remain e…
Study interprets neural network generalization and memorization on corrupted data.
problem Understanding when a neural network has memorized corrupted data versus learned the underlying rule.
method Analyzes multi-layer perceptrons and Transformers on modular arithmetic tasks with corrupted labels.
result Regularization methods can force networks to ignore corrupted data, improving accuracy on uncorrupted data.
Clarifies the confidence interval approach for bioequivalence testing.
problem Ensuring the reliability of bioequivalence testing methods.
method Clarifies the conditions under which a 100(1-2α)% confidence interval yields a size-α test.
result A 100(1-2α)% confidence interval approach for bioequivalence testing yields a size-α test only when the two one-sided tests are 'equal-tailed'.
We propose a novel investment decision strategy (IDS) based on deep learning. The performance of many IDSs is affected by stock similarity. Most existing stock similarity measurements have the problems: (a) The linear nature of many measurements cannot capture nonlinear stock dynamics; (b) The estimation of many simila…
A new index rebalancing strategy reduces large constituent weights without undesirable effects.
problem Undesirable effects of current Nasdaq-100 index rebalancing.
method A simple rebalancing strategy that avoids undesirable effects.
result Preserves the order of index weights and prevents maximum weight increase.
The paper calibrates uncertainty in dropout variational inference models.
problem Miscalibration of model uncertainty in dropout variational inference.
method Logit scaling methods are extended to recalibrate model uncertainty.
result Logit scaling reduces miscalibration, improving reliability of predictions.
Convolutional DKMs improve kernel methods on MNIST, CIFAR-10, and CIFAR-100.
problem Improving kernel methods for image classification.
method Developed a novel inter-domain inducing point approximation and introduced various techniques to extend DKMs to convolutional networks.
result Achieved state-of-the-art performance on image classification benchmarks.
This paper examines how to calibrate ensemble members for better prediction accuracy.
problem Improper calibration of deep neural networks leads to unreliable probability estimates.
method Theoretical analysis and empirical evaluation on CIFAR-100 dataset.
result Well-calibrated ensemble members do not guarantee a well-calibrated ensemble prediction, but a well-calibrated ensemble prediction cannot exceed the average performance of its members.
CutMix training technique improves spatial locality in Vision Transformers.
problem Improving spatial locality in Vision Transformers trained from scratch.
method Comparison of Baseline and Modern training protocols on CIFAR-10, CIFAR-100, and Tiny-ImageNet.
result CutMix training component significantly reduces Mean Attention Distance (MAD) in early layers of Vision Transformers.
A restricted Boltzmann machine (RBM) is a generative neural-network model with many novel applications such as collaborative filtering and acoustic modeling. An RBM lacks the capacity to retain memory, making it inappropriate for dynamic data modeling as in time-series analysis. In this paper we address this issue by p…
A framework for multi-label sentiment analysis in 100 languages with dynamic weighting.
problem Cross-lingual sentiment analysis in multi-label settings with label imbalance.
method Dynamic weighting method, focal loss adaptation, optimal class-specific thresholds.
result State-of-the-art performance in 7 out of 9 metrics across 3 languages.
The methodology presented provides a quantitative way to characterize investor behavior and price dynamics within a particular asset class and time period. The methodology is applied to a data set consisting of over 250,000 data points of the S&P 100 stocks during 2004-2018. Using a two-way fixed-effects model, we unco…
New approach reduces malware detection memory requirements and speeds up training.
problem Efficiently classifying long sequences of malware detection data.
method Developed a new temporal max pooling method and global channel gating design.
result 116x more memory efficient and 25.8x faster training on original dataset.
Simplified non-contrastive learning avoids representation collapse.
problem Training failure modes in self-supervised learning.
method Hyperdimensional computing and inductive bias.
result The approach avoids representation collapses.
Unified view of contrastive and supervised learning improves model performance.
problem Improving model performance in classification, robustness, and detection.
method Hybrid discriminative-generative training of energy-based models.
result Improved performance on classification, robustness, out-of-distribution detection, and calibration.
Colored noise improves neural network robustness against adversarial attacks.
problem Vulnerability of neural networks to adversarial perturbations.
method Injection of colored noise into network weights and activations during adversarial training.
result Our approach outperforms previous methods in terms of adversarial accuracy on CIFAR-10 and CIFAR-100 datasets.
Ensembles, where multiple neural networks are trained individually and their predictions are averaged, have been shown to be widely successful for improving both the accuracy and predictive uncertainty of single neural networks. However, an ensemble's cost for both training and testing increases linearly with the numbe…
Graph Ricci flow reveals hidden hierarchies in stock market correlations.
problem Detecting hidden structures in the complex stock market graph.
method Using graph Ricci curvature and flow techniques to analyze the NASDAQ 100 index.
result Algorithm detects hidden hierarchies, community behavior, and clustering in financial markets.
Locally Optimal Block Preconditioned Conjugate Gradient (LOBPCG) is demonstrated to efficiently solve eigenvalue problems for graph Laplacians that appear in spectral clustering. For static graph partitioning, 10-20 iterations of LOBPCG without preconditioning result in ~10x error reduction, enough to achieve 100% corr…
Improved sample efficiency with normalized RBF kernels in neural networks.
problem Learning more with less data in deep learning models.
method Two-phase method to train neural networks with normalized RBF kernels as output layer.
result Normalized RBF kernel networks achieve higher sample efficiency, compactness, and separability.
Study improves risk management for volatile markets using expectiles.
problem Limitations of traditional risk measures during market stress.
method Develops expectile-based framework for FTSE 100 index.
result Expectile-based Value-at-Risk (EVaR) outperforms traditional VaR measures.
CMTF improves financial market forecasting by fusing multiple data types.
problem Lack of effective integration of diverse financial data sources.
method Transformer-based deep learning framework with tensor interpretation and auto-training.
result CMTF outperforms classical and deep learning models in price direction classification.
MALT improves adversarial attacks by targeting classes more efficiently.
problem Naive targeting of adversarial attacks based on classifier confidence.
method MALT - Mesoscopic Almost Linearity Targeting, based on medium-scale almost linearity assumptions.
result MALT wins over AutoAttack on CIFAR-100 and ImageNet datasets, five times faster.
SPAT improves adversarial robustness by preserving semantics in adversarial training.
problem Adversarial examples often have different semantics than original data, introducing unintended biases.
method Semantics-preserving adversarial training (SPAT) that encourages pixel perturbation shared among all classes.
result SPAT improves adversarial robustness and achieves state-of-the-art results in CIFAR-10 and CIFAR-100.
New method enhances adversarial robustness of deep learning models.
problem Improving the robustness of deep learning models against adversarial attacks.
method Optimal transport regularized divergences applied to distributionally robust optimization.
result Improved adversarial robustness on CIFAR-10 and CIFAR-100 datasets.
A new mutual information optimization method using self-supervised binary contrastive learning.
problem Improving self-supervised contrastive learning for better model performance.
method Proposes a novel loss function for contrastive learning that optimizes mutual information in positive and negative pairs.
result The proposed method outperforms state-of-the-art self-supervised contrastive frameworks on various benchmark datasets.
A knot is called minimal if its knot group admits epimorphisms onto the knot groups of only the trivial knot and itself. In this paper, we determine which two-bridge knot b(p,q) is minimal where q≤6 or p≤100.
We give a "soft" proof of Alberti's Luzin-type theorem in [1] (G. Alberti, A Lusintype theorem for gradients, J. Funct. Anal. 100 (1991)), using elementary geometric measure theory and topology. Applications to the C2-rectifiability problem are also discussed.
The paper addresses optimal control in modern tontines with bequest preferences, showing a linear investment strategy.
problem Optimal controls and decreasing allocation in modern tontines with bequest preferences.
method Dual approach to solve optimal control problems with power utilities, modeling bequest preferences.
result Investment strategy almost linearly adjusts from 0% to 100% over time.
We derive an explicit solution for deterministic market impact parameters in the Graewe and Horst (2017) portfolio liquidation model. The model allows to combine various forms of market impact, namely instantaneous, permanent and temporary. We show that the solutions to the two benchmark models of Almgren and Chris (20…
New bound on neural network generalization error using geometric complexity.
problem Understanding the generalization capabilities of deep neural networks.
method Derive a new upper bound on generalization error using margin-normalized geometric complexity.
result Empirical validation of the bound for ResNet-18 on CIFAR-10 and CIFAR-100 datasets.
In this report we examine the effectiveness of WISER in identification of a chemical culprit during a chemical based Mass Casualty Incident (MCI). We also evaluate and compare Binary Decision Tree (BDT) and Artificial Neural Networks (ANN) using the same experimental conditions as WISER. The reverse engineered set of S…
This paper explores adversarial training limits and improves model robustness against norm-bounded perturbations.
problem Understanding and improving adversarial robustness of deep neural networks.
method Systematic study of adversarial training with various factors, including model size, activation functions, and unlabeled data.
result Training robust models that go beyond state-of-the-art results by combining larger models, Swish/SiLU activations, and model weight averaging.
We propose sequenced-replacement sampling (SRS) for training deep neural networks. The basic idea is to assign a fixed sequence index to each sample in the dataset. Once a mini-batch is randomly drawn in each training iteration, we refill the original dataset by successively adding samples according to their sequence i…
Reduces test set maintenance effort by 80-100%.
problem Lack of proper and up-to-date test sets in real-world scenarios.
method Simple technique to reduce labeling effort.
result Significant reduction in test set maintenance effort (80-100%).
High-dimensional partial differential equations (PDE) appear in a number of models from the financial industry, such as in derivative pricing models, credit valuation adjustment (CVA) models, or portfolio optimization models. The PDEs in such applications are high-dimensional as the dimension corresponds to the number …
Bitcoin's integration with major financial indices intensifies, suggesting a shift from alternative to integrated asset.
problem Understanding Bitcoin's evolving role in financial markets and its correlation dynamics.
method Rolling-window correlation, static correlation coefficients, and event-study framework on daily data from 2018 to 2025.
result Correlation levels between Bitcoin and major indices reached 0.87 in 2024, indicating a more integrated role.
One of the ways to train deep neural networks effectively is to use residual connections. Residual connections can be classified as being either identity connections or bridge-connections with a reshaping convolution. Empirical observations on CIFAR-10 and CIFAR-100 datasets using a baseline Resnet model, with bridge-c…
Many recent works have shown that adversarial examples that fool classifiers can be found by minimally perturbing a normal input. Recent theoretical results, starting with Gilmer et al. (2018b), show that if the inputs are drawn from a concentrated metric probability space, then adversarial examples with small perturba…
The Sornette-Ide differential equation of herding and rational trader behaviour together with very small random noise is shown to lead to crashes or bubbles where the price change goes to infinity after an unpredictable time. About 100 time steps before this singularity, a few predictable roughly log-periodic oscillati…
Geometry of the tracks left by a bicycle is closely related with the so-called Prytz planimeter and with linear fractional transformations of the complex plane. We describe these relations, along with the history of the problem, and give a proof of a conjecture made by Menzin in 1906.
Study improves top-k set prediction with low cardinality.
problem Improving top-k set prediction accuracy with low cardinality.
method Introduces new target loss function and surrogate losses.
result Demonstrates effectiveness of cardinality-aware algorithms.