Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

1122 · Jun 202019922001200920172026
33 results for electoral polls

Complex dynamical systems driven by the unravelling of information can be modelled effectively by treating the underlying flow of information as the model input. Complicated dynamical behaviour of the system is then derived as an output. Such an information-based approach is in sharp contrast to the conventional mathem…

2019-04-21abs ↗pdf ↗

This paper examines how voter concentration affects election outcomes in district-based systems.

problem How does the spatial concentration of electors impact election results?
method The authors frame the spatial distribution of electors in a probabilistic setting and explore models to capture intra-district polarization. They use Likelihood-free Inference under the Approximate Bayesian Computation framework and supervised regression methods to estimate parameters.
result The models can capture statistical properties of real elections and show how voter distributions can change election results.

Donald Trump was lagging behind in nearly all opinion polls leading up to the 2016 US presidential election, but he surprisingly won the election. This raises the following important questions: 1) why most opinion polls were not accurate in 2016? and 2) how to improve the accuracies of opinion polls? In this paper, we …

2018-12-31abs ↗pdf ↗

Calculates winning probability for three candidates based on support rates and information timing.

problem Determining optimal strategy for three candidates in an election.
method Closed-form solution using support rates, political spectrum positioning, time left, and information revelation rate.
result Optimal strategy can be complex, especially for candidates in the center of a polarized electorate.

Study on continual learning with Twitter data, developing ConGraD algorithm.

problem Personalized online language learning on a massive scale.
method Developed POLL problem setting, collected Firehose datasets, and introduced ConGraD algorithm.
result ConGraD algorithm outperforms prior continual learning methods on Firehose datasets.

This paper presents a data set describing the evolution of results in the Portuguese Parliamentary Elections of October 6th^{th} 2019. The data spans a time interval of 4 hours and 25 minutes, in intervals of 5 minutes, concerning the results of the 27 parties involved in the electoral event. The data set is tailored f…

2019-12-05abs ↗pdf ↗

The problem of population recovery refers to estimating a distribution based on incomplete or corrupted samples. Consider a random poll of sample size nn conducted on a population of individuals, where each pollee is asked to answer dd binary questions. We consider one of the two polling impediments: (a) in lossy pop…

2017-02-18abs ↗pdf ↗

GPI uses GenAI models to infer causal and predictive effects from unstructured data.

problem Estimating causal and predictive effects from unstructured data like text and images.
method Leverages open-source GenAI models to generate and represent unstructured data, applying machine learning to these representations.
result GPI efficiently estimates causal and predictive effects with quantified uncertainty, without fine-tuning.

The paper uses Black-Scholes model to analyze political support and coalition agreements.

problem Determining the minimum support level for a minor party in a pre-electoral coalition.
method Modeling political support as a stochastic process with a deterministic growth rate and applying Black-Scholes option pricing theory.
result The minimum support level for a minor party to gain a representative in a pre-electoral coalition.

Modern applications of machine learning (ML) deal with increasingly heterogeneous datasets comprised of data collected from overlapping latent subpopulations. As a result, traditional models trained over large datasets may fail to recognize highly predictive localized effects in favour of weakly predictive global patte…

2019-10-15abs ↗pdf ↗

Enhances topic-metadata relationship modeling using Bayesian methods.

problem Estimating relationships between latent topics and metadata in topic modeling.
method Proposes modifications to the method of composition, using Beta regression and a fully Bayesian approach.
result Improves quantification of uncertainty in topic-metadata relationships.

MakerDAO's governance is centralized despite its decentralized claim.

problem Decentralization illusion in Decentralized Finance (DeFi) governance.
method Empirical analysis using financial, transaction, network, and sentiment indicators.
result Centralized governance impacts Maker protocol and voting power distribution.

The paper explores learning from label proportions, showing differences in efficiency between LLP and PAC learning.

problem Learning from label proportions (LLP) in unlabeled data with given label proportions.
method Formal definition and computational complexity analysis of LLP learning.
result LLP learning is more restrictive than PAC learning for finite VC classes, and some classes are uncharacterizable.

QuEst combines model predictions with observed data to estimate quantile-based measures.

problem Limited applicability of current hybrid-inference tools for quantile-based distributional measures.
method Principled framework merging observed and imputed data for a wide range of quantile-based measures.
result QuEst delivers point estimates and rigorous confidence intervals for quantile-based measures.

Wasserstein t-SNE embeds hierarchical datasets considering within-unit distributions.

problem Exploring hierarchical datasets where units are compared based on means of sample distributions.
method Uses Wasserstein distance metric for 2D embeddings of units, approximating Gaussian distributions for efficiency.
result Demonstrates effective embedding of hierarchical datasets, uncovering meaningful structure.

Majority Vote is optimal for reliable data labeling under certain conditions.

problem Reliable data labeling requires aggregating multiple annotators' labels, but the optimality of Majority Vote is not well understood.
method Characterized conditions under which Majority Vote achieves the optimal label estimation error.
result Majority Vote optimally recovers labels for a given class distribution under tolerable annotation noise limits.

Study improves pension scheme efficiency in Kenya through governance and risk management.

problem Limited research on efficiency of Kenyan pension schemes under governance structures.
method Quantitative panel regression analysis on 128 Kenyan pension schemes over 7 years.
result Employee board members have a significant positive effect on pension scheme efficiency.

We consider the estimation of binary election outcomes as martingales and propose an arbitrage pricing when one continuously updates estimates. We argue that the estimator needs to be priced as a binary option as the arbitrage valuation minimizes the conventionally used Brier score for tracking the accuracy of probabil…

2017-03-18abs ↗pdf ↗

Hybrid engine analyzes news sentiment for markets in real-time.

problem Real-time market analysis of news sentiment.
method Three-way ensemble learning combining financial lexicon, adaptive TF-IDF clustering, and auto-calibrated weighting.
result Adaptive statistical clustering learner improves adaptability to market changes.

This paper examines how institutional liquidity affects prediction markets.

problem How institutional liquidity impacts prediction markets and their quality.
method Defines a market-quality lens, separates channels, and uses synthetic microstructure lab.
result Institutional liquidity does not necessarily translate to equal gains for all traders.

Imitative and contrarian behaviors are the two typical opposite attitudes of investors in stock markets. We introduce a simple model to investigate their interplay in a stock market where agents can take only two states, bullish or bearish. Each bullish (bearish) agent polls m "friends'' and changes her opinion to bear…

2001-09-21abs ↗pdf ↗

Study optimizes data collection from biased, costly sources to minimize risk.

problem Estimating population means and group-conditional means from multiple sources with varying costs and biases.
method Develops a sampling plan that maximizes effective sample size, paired with a post-stratification estimator.
result Achieves budgeted minimax optimal risk for estimating population means and group-conditional means.