A method for combining classifiers from multiple views using Bregman divergences.
problem Combining classifiers from multiple views with limited labeled data.
method Jointly learns view-specific and overall weighted majority vote classifiers using Bregman divergences.
result Empirical results show improved classifier performance with limited labeled data.
Majority voting neural networks improve binary compressed sensing for sparse signal recovery.
problem Sparse signal recovery in binary compressed sensing.
method Majority voting neural networks with a cross entropy-like term and L1 regularization.
result The majority voting neural network achieves excellent recovery performance, approaching optimal performance as the number of component nets grows.
Majority-of-Three is Optimal
problem Optimality of Voting Learners
method Majority vote of three classifiers
result Proves optimality for the simplest voting scheme
Paper proposes a voting method to improve acoustic scene classification.
problem Improving acoustic scene classification accuracy.
method Punishment voting algorithm based on super categories construction.
result Punishment voting significantly improves classification performance.
New margin bound improves generalization for voting classifiers.
problem Improving generalization bounds for voting classifiers.
method Established a new margin-based generalization bound.
result Derives an optimal weak-to-strong learner with matching theoretical lower bound.
The paper studies a stochastic majority vote approach to improve classifier accuracy.
problem Improving classifier accuracy over ensembles of classifiers.
method Minimizing a PAC-Bayes generalization bound with Dirichlet distributions.
result Achieves state-of-the-art accuracy and tight generalization bounds.
Novel analysis improves weighted majority vote in multiclass classification.
problem Improving the performance of weighted majority vote in multiclass classification.
method Analyzes expected risk of weighted majority vote, considering prediction correlations and provides a bound for efficient minimization.
result Minimization of the new bound typically does not degrade the test error of the ensemble.
The paper challenges the assumption that majority voting rights equate to 'effective control' in foreign ownership regulations.
problem The assumption that majority voting rights determine 'effective control' in foreign ownership regulations is flawed.
method The paper proposes and demonstrates a method for calculating 'effective control' based on voting thresholds and weights.
result The 'effective control' of a foreign minority stockholder can be higher than their shareholding size, challenging the assumption that majority voting rights equate to 'effective control'.
New algorithms minimize PAC-Bayesian C-Bound for majority voting, leading to scalable and accurate predictors.
problem Improving majority vote classifiers using PAC-Bayesian bounds.
method Directly optimizing PAC-Bayesian guarantees on the C-Bound with gradient descent.
result Self-bounding majority vote learning algorithms with scalable and accurate predictors.
In his seminal work, Schapire (1990) proved that weak classifiers could be improved to achieve arbitrarily high accuracy, but he never implied that a simple majority-vote mechanism could always do the trick. By comparing the asymptotic misclassification error of the majority-vote classifier with the average individual …
Approach to detect emotion from speech using majority voting and selected features.
problem Detecting human emotion from speech.
method Majority voting technique over machine learning models (NN, DT, SVM, KNN).
result Majority voting technique achieves better accuracy than individual models.
Ensemble CNNs improve mode classification in smartphone travel surveys.
problem Classifying transportation modes from smartphone travel survey data.
method Developed an ensemble of CNN models with different architectures and hyper-parameters, combined using average voting, majority voting, optimal weights, and a Random Forest meta-learner.
result The ensemble method with Random Forest as meta-learner achieved 91.8% accuracy, surpassing other methods.
Best-of-Majority improves inference performance in Pass@k settings.
problem Inference in difficult tasks often underperforms with single-shot selection methods.
method Combining majority voting and Best-of-N, Best-of-Majority restricts candidates to high-frequency responses.
result Best-of-Majority achieves minimax optimal regret and outperforms other methods.
RCAM-based ensemble combines binary classifiers using similarity and vote scheme.
problem Improving binary classification accuracy through ensemble methods.
method RCAM-based ensemble combining classifiers using similarity and recurrent consult-vote scheme.
result RCAM-based ensemble outperforms individual classifiers and majority voting.
Mathematical analysis shows Brexit affects EU voting power in unexpected ways.
problem Effects of Brexit on EU voting power and distribution of power.
method Mathematical analysis using Penrose--Banzhaf Index and normal approximation.
result Non-monotonic effects of Brexit on EU voting power, exacerbated by EU population vector.
Crowdsourcing has become an effective and popular tool for human-powered computation to label large datasets. Since the workers can be unreliable, it is common in crowdsourcing to assign multiple workers to one task, and to aggregate the labels in order to obtain results of high quality. In this paper, we provide finit…
New model captures intransitive preferences without concave likelihood.
problem Complex human choices not accounted for by traditional models.
method Inspired by Condorcet method, Majority Vote model using RUMs.
result Three-dimensional model can represent strong, long intransitive cycles.
Majority Vote is optimal for reliable data labeling under certain conditions.
problem Reliable data labeling requires aggregating multiple annotators' labels, but the optimality of Majority Vote is not well understood.
method Characterized conditions under which Majority Vote achieves the optimal label estimation error.
result Majority Vote optimally recovers labels for a given class distribution under tolerable annotation noise limits.
Simpler majority vote of three classifiers achieves optimal error bounds.
problem Developing an optimal PAC learning algorithm in the realizable setting.
method Returning the majority vote of three ERM classifiers.
result Achieves optimal in-expectation bound on error.
The C-bound risk analysis improves majority voting in binary classification.
problem Improving binary classification accuracy through majority voting.
method PAC-Bayesian analysis and MinCq learning algorithm.
result MinCq learning algorithm achieves state-of-the-art performance.
Machine learning ensemble improves accuracy by considering minority answers as more likely true.
problem Ensemble methods often rely on majority voting, which can fail when the majority is wrong.
method Proposes Bayesian Truth Serum for classification problems, detecting surprising majority answers.
result Better classification performance achieved by considering minority answers as more likely true.
The paper improves confidence regions for band-limited functions using tighter norm bounds and majority voting.
problem Constructing reliable confidence regions for band-limited functions from noisy data.
method Improved norm bounds using Hoeffding's inequality and empirical Bernstein bound, majority voting to aggregate intervals.
result Confidence intervals retain their simultaneous coverage guarantee even when aggregated from random subsamples.
New PAC-Bayesian bounds for multi-view learning using Rényi divergence.
problem Applying PAC-Bayesian theory to multi-view learning.
method Introducing novel PAC-Bayesian bounds based on Rényi divergence for multi-view learning.
result Efficient optimization algorithms that align with theoretical bounds.
Study generalization of voting classifiers using margin-based bounds.
problem Understanding the generalization of ensemble classifiers like voting.
method Proved margin-based generalization bounds using PAC-Bayes theory and Dirichlet posteriors.
result Provided state-of-the-art guarantees on classification tasks.
New bounds on majority voting's accuracy for multi-class classification problems.
problem Determining the accuracy of majority voting for multi-class classification.
method Analyzing the majority voting function under different voter conditions and distributions.
result The error rate of majority voting exponentially decays or grows with the number of voters under certain conditions.
Paper proposes FVC for functional data classification.
problem Challenges in high-dimensional temporal data.
method Ensemble learning of diverse functional representations.
result FVC enhances predictive accuracy over individual models.
New bound improves on weighted majority vote risk estimation.
problem Improving risk estimation for weighted majority vote.
method Novel Chebyshev-Cantelli inequality and PAC-Bayes-Bennett inequality.
result New bounds improve on existing methods.
We tackle the PAC-Bayesian Domain Adaptation (DA) problem. This arrives when one desires to learn, from a source distribution, a good weighted majority vote (over a set of classifiers) on a different target distribution. In this context, the disagreement between classifiers is known crucial to control. In non-DA superv…
Crowdsourcing is an effective tool for human-powered computation on many tasks challenging for computers. In this paper, we provide finite-sample exponential bounds on the error rate (in probability and in expectation) of hyperplane binary labeling rules under the Dawid-Skene crowdsourcing model. The bounds can be appl…
Study improves fair opinion aggregation by balancing voter attributes.
problem Aggregation of opinions can be biased by voter attributes.
method Combines majority voting and D&S model with fairness options.
result Effective combination of Soft D&S and fairness options for different data types.
Prefix consistency improves model reliability by weighting answers based on their reproducibility.
problem Improving the reliability of large language models' reasoning traces.
method Use prefix consistency to weight candidate answers based on their reproducibility during regeneration.
result Prefix consistency is the best correctness predictor, reaching Standard MV plateau accuracy with up to 21x fewer tokens.
In machine learning, Domain Adaptation (DA) arises when the distribution gen- erating the test (target) data differs from the one generating the learning (source) data. It is well known that DA is an hard task even under strong assumptions, among which the covariate-shift where the source and target distributions diver…
In machine learning, the domain adaptation problem arrives when the test (target) and the train (source) data are generated from different distributions. A key applied issue is thus the design of algorithms able to generalize on a new distribution, for which we have no label information. We focus on learning classifica…
Geometric framework determines optimal number of ensemble classifiers.
problem Determining the optimal number of component classifiers for ensemble prediction accuracy.
method Geometric framework for a priori determining ensemble size, applicable to batch and online environments.
result The framework proves the existence of an ideal number of components for WMV, equal to the number of class labels.
Optimal number of voters for a voting ensemble can be estimated from the distribution of classifier errors.
problem Finding the optimal number of voters for a voting ensemble to minimize error rate.
method Estimate the distribution of classifier errors and infer error rates for different numbers of voters.
result Lower-variance estimates of error rates can be obtained by inferring them for different numbers of voters.
For classifying time series, a nearest-neighbor approach is widely used in practice with performance often competitive with or better than more elaborate methods such as neural networks, decision trees, and support vector machines. We develop theoretical justification for the effectiveness of nearest-neighbor-like clas…
Boosting improves accuracy by combining weak learners into a voting classifier.
problem Boosting's theoretical performance is sub-optimal, especially for voting classifiers.
method Proposes a randomized boosting algorithm that outputs voting classifiers with a single logarithmic dependency on sample size.
result Randomized boosting achieves a generalization error with a single logarithmic dependency on the sample size.
New PAC-Bayesian approach for domain adaptation.
problem Learning a model from a source domain for a target domain with distributional shift.
method PAC-Bayesian analysis, deriving an upper-bound on target risk.
result Upper-bound on target risk with distribution divergence controlling trade-off.
We revisit the classical decision-theoretic problem of weighted expert voting from a statistical learning perspective. In particular, we examine the consistency (both asymptotic and finitary) of the optimal Nitzan-Paroush weighted majority and related rules. In the case of known expert competence levels, we give sharp …
Researchers combined linear classifiers using score functions and found simple and trimmed averages to be the best combination strategies.
problem Combining linear classifiers using their score functions.
method Two score functions tested; four combination strategies investigated; comparison with majority voting and model averaging.
result Simple and trimmed average combination strategies were the best.
The paper offers a simple proof of Condorcet's jury theorem.
problem The relationship between majority voting and Condorcet's jury theorem.
method A simple derivation of Condorcet's jury theorem.
result Condorcet's jury theorem is more likely to choose correctly when individual votes are often correct and independent.
New algorithm reduces communication costs in distributed deep learning.
problem High communication costs in distributed deep learning.
method Sparse-SignSGD with Majority Vote (S3GD-MV).
result Significantly reduces communication costs while maintaining accuracy.
QAlign improves language model alignment with less compute, outperforming existing methods.
problem Improving language model performance with limited test-time computation.
method QAlign: sampling from optimal aligned distribution using Markov chain Monte Carlo.
result Consistent improvements over existing methods on various benchmarks.
This paper analyzes voter coalitions in MakerDAO's decentralized governance.
problem Understanding the governance structure and influence of voter coalitions in DAOs.
method Applied clustering algorithm to voting history of MakerDAO to identify voter coalitions.
result The emergence of a dominant voter coalition signals governance centralization in DAOs.
The paper analyzes how multiple classifiers' disagreement and polarization affect overall accuracy.
problem Improving accuracy through ensembling multiple classifiers.
method The paper derives an upper bound for polarization, proposes a neural polarization law, and presents a tight upper bound for the error of majority vote classifiers.
result Disagreement and polarization among classifiers are linearly correlated with the target, and polarization is nearly constant for a dataset.
Removing or filtering outliers and mislabeled instances prior to training a learning algorithm has been shown to increase classification accuracy. A popular approach for handling outliers and mislabeled instances is to remove any instance that is misclassified by a learning algorithm. However, an examination of which l…
ARIMLE optimizes classifier fusion for brain-computer interface.
problem Improving ensemble classifier aggregation performance.
method ARIMLE uses agreement rate to estimate classifier accuracy, then refines a maximum likelihood estimator.
result ARIMLE outperforms majority voting and other methods in brain-computer interface applications.
This paper generalizes an important result from the PAC-Bayesian literature for binary classification to the case of ensemble methods for structured outputs. We prove a generic version of the \Cbound, an upper bound over the risk of models expressed as a weighted majority vote that is based on the first and second stat…