Study shows skewed data labels significantly impact decentralized ML accuracy.
problem Skewed data labels across devices/locations cause significant accuracy loss in decentralized ML.
method Detailed experimental study on skewed data labels, presenting SkewScout system-level approach.
result Skewed data labels are a fundamental challenge for decentralized learning, affecting many applications and models.
New method corrects skewed confidence for PbN classification.
problem Weakly supervised binary classification with biased negative data.
method Corrects skewed confidence in negative data to improve classifier.
result Reduces distortion in posterior probability for PbN classification.
Layer normalization improves federated learning with skewed labels.
problem Label skewness in federated learning datasets.
method Identified feature normalization as key mechanism; applied to latent features before classifier.
result Normalization accelerates global training and improves convergence under extreme label shift.
This study assesses the impact of non-IID data in federated learning, revealing significant performance drops.
problem The impact of non-IID data on federated learning model performance.
method Empirical analysis using Hellinger Distance to measure distribution differences, benchmarking four strategies for handling non-IID data.
result Significant performance drops occur at specific HD thresholds, especially for extreme non-IIDness.
Proposes a method to identify elements in a skewness matrix for multivariate skew-elliptical distributions.
problem Label switching issue in Bayesian estimation of skewness matrix.
method Imposes a positive lower-triangular constraint and uses Bayesian sparse estimation with horseshoe prior.
result Successfully estimates the true structure of skewness dependency.
A new active learning method for skewed data sets.
problem Severe class imbalance and small initial training data in sequential active learning.
method HAL: a hybrid active learning algorithm balancing labeled and unlabeled data.
result HAL makes better choices for labeling points than strong baselines.
Machine learning approaches to multi-label document classification have to date largely relied on discriminative modeling techniques such as support vector machines. A drawback of these approaches is that performance rapidly drops off as the total number of labels and the number of labels per document increase. This pr…
We improve kernel ridge regression for skewed responses using oversampling and adaptive partitioning.
problem Kernel ridge regression struggles with skewed response variables, leading to poor estimates.
method Combines adaptive partitioning with oversampling to address skewed responses in kernel ridge regression.
result The proposed method yields estimates with smaller risk compared to classical methods under mild conditions.
Paper analyzes D-SGD convergence with heterogeneous data and proposes topology learning.
problem Efficiently dealing with data heterogeneity in decentralized learning.
method Revisits D-SGD analysis, introduces neighborhood heterogeneity, and proposes topology learning.
result Formulates topology learning as a tractable optimization problem and demonstrates its effectiveness.
The paper analyzes skewness and kurtosis measures for skew-elliptical distributions.
problem Examining skewness and kurtosis measures for skew-elliptical distributions.
method Deriving exact expressions for skewness and kurtosis measures for skew-elliptical distributions, constructing test statistics, and comparing measures through simulations and real data analysis.
result Exact expressions and test statistics for skewness and kurtosis measures for various skew-elliptical distributions.
Let M be a smooth closed orientable surface. Let F be the space of Morse functions on M having fixed number of critical points of each index, moreover at least χ(M)+1 critical points are labeled by different labels (enumerated). A notion of a skew cylindric-polyhedral complex, which generalizes the notion of a …
New RESK distributions improve robust clustering of skewed data.
problem Robustly clustering non-symmetric, heavy-tailed data clusters.
method Proposes RESK distributions and an EM algorithm with robust skew-Huber M-estimator.
result Numerical experiments confirm the effectiveness of the proposed methods.
A new clustering method for functional data using skewed distributions.
problem Clustering functional data with skewed distributions.
method Mixtures of functional linear regression models and three skewed multivariate distributions (variance-gamma, skew-t, normal-inverse Gaussian).
result The proposed method funWeightClustSkew performs well on simulated and real data.
A new method predicts true classes from positive and unlabeled data with additional labeled observations.
problem Predicting true classes from positive and unlabeled data with selection bias.
method Introduces augmented PU prediction, allowing feature-dependent labeling, and compares various empirical Bayes rules.
result The variational autoencoder-based method performs similarly or better than other methods and improves accuracy for unlabeled samples.
WES improves neural network regression by stretching distribution error.
problem Improving prediction performance in neural-network-based regression.
method Proposed weighted empirical stretching (WES) loss function.
result WES outperforms existing loss functions, especially in extreme domains.
Improves binary classification from positive data with skewed confidence.
problem Skewed confidence in positive data affects the performance of Pconf classifiers.
method Parameterized model of skewed confidence and hyperparameter selection.
result Proposed method effectively cancels out the negative impact of skewed confidence.
A mixture of common skew-t factor analyzers model is introduced for model-based clustering of high-dimensional data. By assuming common component factor loadings, this model allows clustering to be performed in the presence of a large number of mixture components or when the number of dimensions is too large to be well…
The paper improves asset allocation using a skew-normal distribution in the Black-Litterman model.
problem Improving asset allocation under skewed return distributions.
method Using the Black-Litterman model with hidden truncation skew-normal distribution and Simaan's three-moment risk model.
result Optimal portfolios have less risk and higher skewness compared to classical BL model.
This paper classifies 4D spin manifolds with skew Killing spinors.
problem Classifying 4D Riemannian spin manifolds with skew Killing spinors.
method Analyzing skew Killing spinors with skew-symmetric endomorphisms A, considering both degenerate and non-degenerate cases.
result In the degenerate case, the manifold is locally isometric to R x N with N having a skew Killing spinor.
Study shows different types of volatility and skewness changes affect stock prices.
problem Different types of volatility and skewness changes affect stock prices.
method Used intraday data for individual stocks to analyze cross-section of asset returns.
result Idiosyncratic transitory and persistent shocks to volatility and skewness are priced differently in stock returns.
In recent years, data have become increasingly higher dimensional and, therefore, an increased need has arisen for dimension reduction techniques for clustering. Although such techniques are firmly established in the literature for multivariate data, there is a relative paucity in the area of matrix variate, or three-w…
This paper analyzes how differential privacy and data skewness affect membership inference attacks.
problem Membership inference attacks on privately trained models.
method Developed MPLens system to evaluate membership inference vulnerability.
result Membership inference risk is higher with skewed training data and differential privacy has trade-offs.
Researchers develop a new spatial process model for non-Gaussian data.
problem Non-Gaussian spatial data with asymmetry and heavy-tailedness.
method Re-parameterized Unified Skew-Normal (SUN) distribution, GSUN process, neural Bayes inference with GATs.
result GSUN process captures non-Gaussian spatial data properties and outperforms conventional models.
Protein-protein interaction (PPI) prediction is an important problem in machine learning and computational biology. However, there is no data set for training or evaluation purposes, where all the instances are accurately labeled. Instead, what is available are instances of positive class (with possibly noisy labels) a…
The paper defines MTCov for skewed elliptical distributions.
problem No specific problem stated, but dealing with skewed elliptical distributions.
method Defined MTCov for generalized skew-elliptical distributions and compared with skewed and non-skewed normal distributions.
result Special formula for MTCov of generalized skew-elliptical distributions.
New algorithm optimizes privacy and utility in multi-task learning with skewed data.
problem Privacy constraints in multi-task learning with uneven data distribution.
method Adaptive reweighting of privacy budget allocation among tasks.
result Significant improvement in utility with state-of-the-art performance on benchmarks.
Study models extreme skew surges along French Atlantic coast.
problem Appropriate modelling of extreme skew surges for coastal risk management.
method Peak-over-threshold framework, multivariate generalized Pareto distribution, extreme regression framework.
result Reconstructed historical skew surge time series at stations with limited data.
Framework tackles class imbalance and noisy labels in active learning.
problem Class imbalance and noisy labels in real-world datasets.
method Uses foundation model priors to select informative samples for active learning.
result Substantial annotation savings (over 50%) with preserved performance and robustness.
Dynamic skewness models improve financial time series analysis.
problem Modeling financial time series with skewness and heavy tails.
method Dynamic skewness stochastic volatility models with penalized priors and HMC estimation.
result Penalized priors outperform classical choices in model performance.
SkewPNN uses probabilistic neural networks with skew-normal kernels to improve classification of imbalanced data.
problem Imbalanced data distribution leading to biased predictions for minority classes.
method Probabilistic neural networks with skew-normal kernel function and Bat optimization algorithm for hyperparameter tuning.
result SkewPNN and BA-SkewPNN outperform other methods in both balanced and imbalanced datasets.
The paper calculates moments and conditional risks for skewed elliptical distributions.
problem Estimating moments and tail conditional risks for skewed elliptical distributions.
method Derives explicit expressions for multivariate doubly truncated moments and conditional risks for generalized skew-elliptical distributions.
result Explicit formulas for multivariate doubly truncated moments and conditional risks are derived for various skewed elliptical distributions.
Paper introduces a new performance metric for class imbalance datasets.
problem Challenges in selecting and comparing models for imbalanced datasets.
method Proposes a new performance measure based on the harmonic mean of Recall and Selectivity normalized in class labels.
result The proposed measure is less sensitive to changes in the majority class and more sensitive to changes in the minority class.
Two different formulas for macro F1 lead to significant differences in classification evaluation.
problem Evaluation discrepancies in binary, multi-class, and multi-label classification problems.
method Comparison of two formulas for macro F1 metric.
result The two formulas can result in up to a 0.5 difference and different classifier rankings.
Bayesian VI copula models capture asymmetric intraday equity dependence.
problem Modeling asymmetric and extreme tail dependence in financial data.
method Bayesian variational inference for skew-t copula models in high dimensions.
result The copula captures substantial heterogeneity in asymmetric dependence over equity pairs and time.
We review and illustrate how the volatility smile translates into a probability distribution, the market-implied probability distribution representing believes priced in. The effects of changes in the smile are examined. Special attention is given to the effects of slope, which might appear at first counter-intuitive. …
Proposes vMF distribution for skewed elliptical distributions.
problem Skewed distributions not adequately modeled by symmetric distributions.
method Introduces von-Mises-Fisher (vMF) distribution to represent skewed elliptical distributions.
result vMF distribution provides an explicit and simple probability representation of skewed elliptical distributions.
Enhances knot counting invariant using skew braces.
problem Counting invariant for virtual knots and links.
method Introduces new invariants using skew brace structures.
result New invariants not determined by the counting invariant.
Develops a robust model for skewed and heavy-tailed data in periodontal studies.
problem Skewed and heavy-tailed data in periodontal pocket depth measurements.
method Flexible two-piece scale Student-t error distribution and deep neural network with monotonicity constraints.
result Robust mode-based estimation resistant to outliers with clinical interpretability.
Study the geometry and dynamics of skew evolutes and involutes, related to bicycle kinematics.
problem Understanding the geometry and dynamics of skew evolutes and involutes.
method Investigate the skew evolute and involute maps, comparing them to bicycle kinematics.
result The skew evolute and involute maps have properties analogous to bicycle kinematics.
Study on simplicity of Lie skew braces, proving new results for compact cases.
problem Simplicity of Lie skew braces, focusing on compact connected cases.
method Reviewing correspondence, investigating ideals and rigidity, proving main result for compact Lie skew braces.
result Compact connected simple Lie skew braces are either trivial or have simple underlying Lie groups.
The paper examines smoothness in graded skew Clifford algebras.
problem Smoothness of graded skew Clifford algebras.
method Investigation of differential smoothness.
result Results on the differential smoothness of graded skew Clifford algebras.
Proposes a new model for clustering with heavier tails.
problem Clustering with heavy-tailed data.
method Finite mixture of skewed sub-Gaussian stable distributions, maximum likelihood estimation, EM algorithm.
result The proposed model can robustly handle heavy-tailed data.
A skew loop is a closed curve without parallel tangent lines. We prove: The only complete surfaces in euclidean 3-space with a point of positive curvature and no skew loops are the quadrics. In particular, ellipsoids are the only closed surfaces without skew loops. We also prove results about skew loops on cylinders an…
Given a constant magnetic field on Euclidean space Rp determined by a skew-symmetric (p×p) matrix Θ, and a Zp-invariant probability measure μ on the disorder set Σ which is by hypothesis a Cantor set, where the action is assumed to be minimal, the corresponding Integrated Density…
Simple method solves Quanto Skew problem.
problem Quanto Skew problem in Equities and FX.
method Analytical method that accommodates Equity and FX volatility skew.
result Highly efficient and fast performance.
Skewness dispersion predicts future stock market returns, especially in months with monetary policy announcements.
problem Predicting future stock market returns using skewness dispersion.
method Cross-sectional analysis of firm-level realized skewness and stock market returns.
result Skewness dispersion is a significant predictor of future stock market returns, robust to various estimation methods.
New topological biquandles created using skew braces.
problem Creating nontrivial topological biquandles.
method Using the concept of skew braces.
result Constructs nontrivial examples of topological biquandles.
Examines differential smoothness in a specific skew PBW extension family.
problem Differential smoothness in skew PBW extensions.
method Investigates a specific family of skew PBW extensions.
result Results on differential smoothness of the family.