New proof shows how to identify DAGs with weakly increasing errors.
problem Identifying the true DAG in models with weakly increasing error variances.
method Minimum-trace DAG method and hill climbing algorithm with R2R neighborhood.
result Hill climbing algorithm without strict local optima under weakly increasing error variances.
We describe an approach to Grammatical Error Correction (GEC) that is effective at making use of models trained on large amounts of weakly supervised bitext. We train the Transformer sequence-to-sequence model on 4B tokens of Wikipedia revisions and employ an iterative decoding strategy that is tailored to the loosely-…
Gradient descent performs well on weakly convex losses, offering generalization guarantees.
problem Learning with weakly convex losses using gradient descent.
method Analyzing the stability of gradient descent through the smallest eigenvalue of the Hessian.
result Generalization error bounds hold under a wider range of step sizes.
Randomized gradient-based ensemble improves prediction accuracy.
problem Improving prediction accuracy in machine learning.
method Randomization and gradient-based aggregation of weakly-correlated estimators.
result The method outperforms existing techniques in terms of increased accuracy.
Improves label propagation for weakly supervised learning.
problem Reducing the need for labeled data in machine learning.
method Label Propagation with Weak Supervision (LPA) analysis.
result Demonstrated improvements over existing methods on weakly supervised classification tasks.
Dark blockchain venues increase miners' profits but raise users' execution risk.
problem Exploitable information leakage in blockchain transactions.
method Economic incentive analysis and empirical study of dark venues.
result Dark venues increase miners' profits but raise users' execution risk.
The complexity of a learning task is increased by transformations in the input space that preserve class identity. Visual object recognition for example is affected by changes in viewpoint, scale, illumination or planar transformations. While drastically altering the visual appearance, these changes are orthogonal to r…
Weakly supervised data are widespread and have attracted much attention. However, since label quality is often difficult to guarantee, sometimes the use of weakly supervised data will lead to unsatisfactory performance, i.e., performance degradation or poor performance gains. Moreover, it is usually not feasible to man…
Paper tackles drowsy driving by learning from weakly labeled car acceleration data.
problem Lack of labeled data for estimating driver drowsiness.
method Weakly supervised learning, scalable stochastic optimization.
result Algorithm learns from weakly labeled data, outperforming baseline methods.
Solves weakly supervised regression using low-rank approximations and manifold regularization.
problem Weakly supervised regression with known, unknown, and uncertain labels.
method Combines manifold regularization and low-rank matrix decomposition for optimization.
result Improves solution quality and stability for large datasets.
Neural network approximates weakly efficient frontier of convex vector optimization problems.
problem Approximating the weakly efficient frontier of convex vector optimization problems.
method Designing a neural network architecture to approximate the weakly efficient frontier of convex vector optimization problems (CVOP) satisfying Slater's condition.
result The proposed algorithm effectively approximates the true weakly efficient frontier of CVOPs, even for large problems.
New AI error correctors improve classifier performance with provable guarantees.
problem Improving AI classifier performance with scarce training data.
method Weakly supervised AI error correctors with performance guarantees.
result Provable performance guarantees for AI error correction.
New framework for weakly supervised learning from label proportions.
problem Lack of consistent learning procedure and theoretical training criterion for LLP.
method Pose LLP as mutual contamination models (MCMs) and establish unbiased losses and generalization error bounds.
result Established novel technical results for MCMs and proposed a new experimental setting.
Proposes a constrained labeling method for weakly supervised learning.
problem Combining weak supervision signals while navigating misleading correlations.
method Randomized constrained labeling within a defined space.
result Randomized constrained labeling converges after few iterations and outperforms other methods.
This research provides theoretical guarantees for hyperparameter estimation in complex network dynamical systems.
problem Theoretical guarantees for hyperparameter estimation in large, inhomogeneous complex network dynamical systems.
method Formulating the system's evolution in a measure transport perspective, proposing a theoretical framework for estimating hyperparameters with mean-type observations.
result A nonasymptotic bound for the deviation of hyperparameter estimates in inhomogeneous complex network dynamical systems with respect to network population size.
Paper tackles weakly supervised learning from similarity-confidence data.
problem Learning binary classifier from unlabeled data pairs with confidence of similarity.
method Proposes an unbiased estimator of classification risk from Sconf data and risk correction scheme.
result Demonstrates effectiveness of proposed methods through experiments.
Example shows learnable distributions not privately learnable.
problem Learnable distributions under non-private conditions not transferable to differential privacy.
method Example of a distribution class learnable up to constant error in total variation distance but not under differential privacy.
result Contradicts conjecture of Ashtiani on learnability under differential privacy.
We consider the task of training classifiers without labels. We propose a weakly supervised method---adversarial label learning---that trains classifiers to perform well against an adversary that chooses labels for training data. The weak supervision constrains what labels the adversary can choose. The method therefore…
This paper explains why double descent sometimes occurs weakly or not at all from an optimization perspective.
problem Understanding the role of optimization in the phenomenon of double descent.
method Investigates model-wise double descent from an optimization perspective, proposing a unified explanation for its occurrence.
result Model-wise double descent is observed if and only if the optimizer can find a sufficiently low-loss minimum.
Unified approach for multicalibration in weakly supervised learning.
problem Existing multicalibration methods require clean input-label pairs, which are unavailable in weakly supervised learning.
method Developed estimators and post-hoc correction methods for multicalibration under weak supervision.
result Unified framework for estimating and correcting multicalibration under weak supervision with finite-sample guarantees.
Binary PheNorm extends phenotype labeling for EHRs using binary silver labels.
problem Lack of gold-standard phenotype labels in EHR studies.
method Proposes Binary PheNorm, an extension that uses binary silver labels directly in phenotype scoring.
result Binary PheNorm achieved strong discrimination using binary labels alone and improved performance when combined with count labels.
The study connects lattices, Garside structures, and weakly modular graphs.
problem Exploring combinatorial non-positive curvature in various simplicial complexes.
method Analyzing lattices with Z-actions and their quotients. result Lattices and their quotients give rise to weakly modular graphs.
Unified framework for N-tuples learning improves weakly supervised tasks.
problem Reducing annotation burden in supervised learning.
method Empirical risk minimization framework integrating pointwise unlabeled data.
result Framework improves generalization across various N-tuples learning tasks.
Deep neural networks are gaining increasing popularity for the classic text classification task, due to their strong expressive power and less requirement for feature engineering. Despite such attractiveness, neural text classification models suffer from the lack of training data in many real-world applications. Althou…
Paper analyzes convergence of stochastic methods under heavy-tailed noise.
problem Analyzing convergence of stochastic methods under heavy-tailed noise.
method Investigates vanilla and clipped stochastic subgradient descent methods.
result Demonstrates convergence properties under sub-Weibull and p-BCM noise assumptions.
Study compares unsupervised and weakly-supervised methods for anomaly detection at the LHC.
problem Detecting new physics signals at the LHC with model-agnostic techniques.
method Compared unsupervised autoencoder (AE) and weakly-supervised Classification Without Labels (CWoLa) methods.
result Both methods complement each other, providing sensitivity to different types of signals.
This paper shows how to construct sequential tests with power one against weakly compact sets in Polish spaces.
problem Testing composite null hypotheses involving weakly compact sets in Polish spaces.
method Develops sequential tests for i.i.d. laws in Polish spaces, providing a sufficient condition for power one.
result Power-one sequential tests exist for weakly compact sets against their complements in i.i.d. laws in Polish spaces.
Paper tackles efficient learning of non-convex hypotheses in metric spaces.
problem Efficiently find consistent hypotheses for non-convex hypotheses composed of possibly several disconnected regions.
method Proposes a general domain-independent algorithm for finding consistent weakly convex hypotheses and proves sufficient conditions for its efficiency.
result Shows that consistent hypothesis finding problem can be solved in polynomial time for a broad class of weakly convex hypotheses over metric spaces.
Paper tackles robust deep learning from weakly dependent data with unbounded loss and input.
problem Tackles robust deep learning from weakly dependent data with unbounded loss and input.
method Establishes non-asymptotic bounds for expected excess risk under strong mixing and ψ-weak dependence assumptions. result Derives a relationship between bounds and r, and shows convergence rate close to i.i.d. results for r=∞. Weakly Einstein Kähler surfaces are characterized and classified.
problem Characterizing and classifying weakly Einstein Kähler surfaces.
method Several conditions and constructions to characterize and classify weakly Einstein Kähler surfaces.
result Classification of weakly Einstein Kähler surfaces with specific properties and construction of new examples.
Bayesian inference with deep, weakly nonlinear networks is solved rigorously.
problem Bayesian inference with neural networks of specific structure.
method Perturbative analysis of fully connected neural networks with a shaped nonlinearity.
result Neural network Bayesian inference can be equivalent to kernel methods under certain conditions.
We study a class of weakly identifiable location-scale mixture models for which the maximum likelihood estimates based on n i.i.d. samples are known to have lower accuracy than the classical n−21 error. We investigate whether the Expectation-Maximization (EM) algorithm also converges slowly for these m…
The study examines weakly Einstein Lie groups and proves non-existence for certain types.
problem Characterizing and proving the non-existence of weakly Einstein Lie groups.
method Analyzing left-invariant metrics on Lie groups and using algebraic properties.
result No weakly Einstein non-abelian 2-step nilpotent Lie groups exist.
Classifies weakly Einstein submanifolds in space forms satisfying specific equalities.
problem Characterizing submanifolds in space forms with certain geometric properties.
method Classification based on Chen's equality and semisymmetric conditions.
result Classification of weakly Einstein submanifolds in space forms.
The study explores weakly p-Kähler hyperbolic manifolds.
problem Generalization and application of weakly p-Kähler hyperbolic manifolds. method Investigation of generalizations and applications.
result Exploration of weakly p-Kähler hyperbolic manifolds. Variational autoencoders learn unsupervised data representations, but these models frequently converge to minima that fail to preserve meaningful semantic information. For example, variational autoencoders with autoregressive decoders often collapse into autodecoders, where they learn to ignore the encoder input. In th…
We consider the weakly supervised binary classification problem where the labels are randomly flipped with probability 1−α. Although there exist numerous algorithms for this problem, it remains theoretically unexplored how the statistical accuracies and computational efficiency of these algorithms depend on the degr…
Study weakly weighted Einstein-Finsler metrics, showing specific curvature properties and characterizing them.
problem Characterizing weakly weighted Einstein-Finsler metrics.
method Showed isotropic S-curvature under certain conditions. Characterized via navigation expressions and α and β. result Weakly weighted Einstein-Kropina metrics have isotropic S-curvature and can be completely characterized.
The increasing occurrence of ordinal data, mainly sociodemographic, led to a renewed research interest in ordinal regression, i.e. the prediction of ordered classes. Besides model accuracy, the interpretation of these models itself is of high relevance, and existing approaches therefore enforce e.g. model sparsity. For…
The paper provides examples of keen weakly reducible bridge spheres for links in b-bridge position.
problem Characterizing and finding examples of keen weakly reducible bridge spheres.
method Analyzing bridge spheres and their properties in terms of compressing disks and width complex.
result Infinitely many examples of keen weakly reducible bridge spheres for links in b-bridge position.
We obtain a Bernstein-type inequality for sums of Banach-valued random variables satisfying a weak dependence assumption of general type and under certain smoothness assumptions of the underlying Banach norm. We use this inequality in order to investigate in the asymptotical regime the error upper bounds for the broad …
Minimal displacement set in weakly systolic complexes is systolic and embeds isometrically.
problem Structure of minimal displacement set in weakly systolic complexes.
method Investigation of minimal displacement set properties and embeddings.
result Minimal displacement set is systolic and embeds isometrically into the complex.
Study on extended weakly symmetric spaces, classifying and providing an example.
problem Understanding geometric properties of extended weakly symmetric spaces.
method Classification and presentation of a non-trivial example.
result Existence of extended weakly symmetric spaces established.
We present a novel method for variable selection in regression models when covariates are measured with error. The iterative algorithm we propose, MEBoost, follows a path defined by estimating equations that correct for covariate measurement error. Via simulation, we evaluated our method and compare its performance to …
Study characterizes and mitigates imbalances in neurosymbolic learning.
problem Characterizing and mitigating class-specific risks in neural classifiers.
method Theoretical analysis and practical techniques including estimating marginal gold labels and mitigating imbalances at training and testing time.
result Learning imbalances can be greatly impacted by the symbolic component σ, unlike in supervised and weakly supervised learning.
There is a well developed theory of weakly symmetric Riemannian manifolds. Here it is shown that several results in the Riemannian case are also valid for weakly symmetric pseudo-Riemannian manifolds, but some require additional hypotheses. The topics discussed are homogeneity, geodesic completeness, the geodesic orbit…
The Bellman error is a poor proxy for value function accuracy, even with all state-action pairs.
problem The Bellman error is a poor proxy for the accuracy of the value function.
method Study of the Bellman equation as a surrogate objective for value prediction accuracy.
result The magnitude of the Bellman error is only weakly related to the distance to the true value function, even with all state-action pairs.
Neural point estimators improve parameter estimation from replicated data.
problem Making inference from replicated data in weakly-identified and highly-parameterised models.
method Permutation-invariant neural networks for likelihood-free parameter estimation.
result Neural point estimators can quickly and optimally estimate parameters.