Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,181 papers · 148 categories

Trend · papers per month

21426283 · Jun 202019922001200920182026
48 results for crowd-sourced fact checking

Crowd-powered system flags misinformation for fact checking.

problem Reduce the spread of fake news and misinformation on social media.
method Flexible temporal point process representation and scalable online algorithm Curb for optimal fact checking selection.
result Our scalable algorithm Curb can effectively reduce the spread of fake news and misinformation.

Time-aware fact-checking improves veracity predictions for time-sensitive claims.

problem Fact-checking decisions should consider temporal information of claims and evidence.
method Investigated four temporal ranking methods to optimize evidence ranking for fact-checking models.
result Time-aware evidence ranking surpasses relevance assumptions and improves veracity predictions for time-sensitive claims.

Task focuses on fact checking in Q&A forums, improving over baseline systems.

problem Fact checking in community Q&A forums to distinguish factual from opinion.
method Two subtasks: distinguishing factual vs. opinion/advice/socializing, predicting answer truthfulness.
result Improved over baseline systems for both subtasks, but not for Subtask B.

We create a large dataset for fact checking claims and improve prediction accuracy.

problem Fact checking claims from multiple sources is challenging.
method We created a comprehensive dataset and developed a novel method for automatic veracity prediction.
result Our model achieves a Macro F1 of 49.2%, showing significant performance improvements.

A novel fact-checking method using debate dynamics on knowledge graphs.

problem Fact-checking on knowledge graphs with user comprehension and interactive reasoning.
method Reinforcement learning agents debate on paths in the graph to classify facts as true or false.
result Interactive reasoning and user understanding of AI decisions on knowledge graphs.

Paper develops a model for verifying facts in tables without pre-retrieved evidence.

problem Verification of factual claims in structured data, especially in open-domain settings.
method Joint reranking-and-verification model that fuses evidence documents.
result Model achieves comparable performance to closed-domain state-of-the-art on TabFact dataset.

Deep learning model detects and corrects outliers in crowd-sourced weather data.

problem Data quality issues in crowd-sourced weather data.
method Bayesian deep learning approach with Gaussian-uniform mixture density network.
result Automated outlier detection in spatio-temporal environmental modeling.

Formalizes interpreting natural language rules for answering questions, collecting 32k task instances.

problem Interpreting regulations and answering 'Can I...?' or 'Do I have to...?' questions.
method Formalization of task, crowd-sourcing strategy to collect 32k instances, analysis of challenges, evaluation of performance.
result Promising results when no background knowledge is needed, substantial room for improvement when background knowledge is needed.

Developed a neural topic model for classifying COVID-19 disinformation.

problem Tackles the challenge of disinformation during the COVID-19 pandemic.
method Classification-aware neural topic model (CANTM) for COVID-19 disinformation.
result Demonstrated the effectiveness of CANTM in classifying COVID-19 disinformation.

We describe the infinitesimal moduli space of pairs (Y,V)(Y, V) where YY is a manifold with G2G_2 holonomy, and VV is a vector bundle on YY with an instanton connection. These structures arise in connection to the moduli space of heterotic string compactifications on compact and non-compact seven dimensional spaces, e.…

2016-07-12abs ↗pdf ↗

By exploiting standard facts about N=1N=1 and N=2N=2 supersymmetric Yang-Mills theory, the Donaldson invariants of four-manifolds that admit a Kahler metric can be computed. The results are in agreement with available mathematical computations, and provide a powerful check on the standard claims about supersymmetric Yang…

1994-03-31abs ↗pdf ↗

Paper optimizes summarization of multiple document groups for better distinction.

problem Comparative document summarization to select representative documents from multiple groups.
method Formulated new objective functions based on binary classification and maximum mean discrepancy, using gradient-based optimization.
result Gradient-based optimization outperforms other methods in automatic and crowd-sourced evaluations.

The paper examines circle graphs of Gauss diagrams and finds counterexamples to previous descriptions.

problem Problems with previous descriptions of realizable Gauss diagrams.
method Experimental checking and formulation of new descriptions of realizable circle graphs.
result New descriptions of realizable circle graphs and an algorithm for checking realizability.

Bayesian model improves truth inference from highly redundant crowd annotations.

problem Inferring true annotations from highly redundant crowd annotations.
method Bayesian graphical model with conjugate priors and iterative expectation-maximisation inference.
result Our technique significantly outperforms majority vote heuristic at one-sided level 0.025.

The topological underpinnings are presented for a new algorithm which answers the question: `Is a given knot the unknot?' The algorithm uses the braid foliation technology of Bennequin and of Birman and Menasco. The approach is to consider the knot as a closed braid, and to use the fact that a knot is unknotted if and …

1998-01-28abs ↗pdf ↗

BUDS balances privacy and utility by shuffling data, achieving strong privacy with minimal loss.

problem Balancing privacy and utility in crowd-sourced statistical databases.
method One-hot encoding, iterative shuffling, loss estimation, risk minimization.
result Achieves ε=0.02ε= 0.02 for privacy, maintaining a privacy bound of ε=ln[t/((n11)S)]ε= ln [t/((n_1 - 1)^S)].

The paper tackles ranking experts based on their answers to questions, considering statistical and computational challenges.

problem Ranking experts based on their answers to questions, considering isotonic constraints.
method Investigates the existence of statistically optimal and computationally efficient procedures for ranking experts under isotonic constraints.
result Disproves the existence of computational-statistical gaps for the problem.

We study coordinate-invariance of some asymptotic invariants such as the ADM mass or the Chruściel-Herzlich momentum, given by an integral over a "boundary at infinity". When changing the coordinates at infinity, some terms in the change of integrand do not decay fast enough to have a vanishing integral at infinity; bu…

2010-12-16abs ↗pdf ↗

Gradient descent solves sparse skill estimation in crowdsourcing.

problem Crowd-sourced worker skill estimation with sparse and irregular assignments.
method Rank-one matrix completion and projected gradient descent.
result Skill estimates converge to global optima for specific sampling matrices.

Model predicts political ideology using context vectors to mitigate bias and scarcity.

problem Scarcity and selection bias in political ideology prediction.
method Proposes a statistical model decomposing embeddings into context and position vectors, training an end-to-end model for deployment.
result Model can predict ideological labels even with minimal biased data, outperforming state-of-the-art methods.

Multifractal analysis and extensive statistical tests are performed upon intraday minutely data within individual trading days for four stock market indexes (including HSI, SZSC, S&P500, and NASDAQ) to check whether the indexes (instead of the returns) possess multifractality. We find that the mass exponent τ(q)τ(q) is l…

2007-06-14abs ↗pdf ↗

Predicts lead contamination in Flint's water system based on home attributes.

problem Understanding and predicting lead contamination in Flint's water system.
method Data science approach using a large dataset of water tests and a crowd-sourced prediction challenge.
result Elevated lead risks can be weakly predicted from observable home attributes.

New functions derived from arrow diagrams for spherical curves, invariant under certain deformations.

problem Defining and analyzing integer-valued functions on spherical curves.
method Introducing new functions and relators to study spherical curves and their isotopy classes.
result Functions derived from arrow diagrams are invariant under specific deformations.

This is the first of three articles on the Fibered Isomorphism Conjecture of Farrell and Jones for L-theory. We apply the general techniques developed in [15] and [16] to the L-theory case of the conjecture and prove several results. Here we prove the conjecture, after inverting 2, for poly-free groups. In particular, …

2007-03-29abs ↗pdf ↗

This paper deforms complex tori and their mirrors using gerbes.

problem Deforming complex tori and their mirror partners.
method Using flat gerbes to deform complex tori and their mirrors, constructing holomorphic line bundles over deformed objects.
result Deformed complex tori and their mirrors can be studied using flat gerbes.

Statistical model checking for PCTL on MDPs using reinforcement learning.

problem Model checking PCTL specifications on MDPs with statistical methods.
method Reinforcement learning for policy search, statistical model checking with UCB-based Q-learning.
result Provably guaranteed statistical model checking method for PCTL specifications on MDPs.

Community moderation drifts towards majority, study finds.

problem How to ensure crowd-sourced moderation systems trust and reward accurate evaluations.
method Consensus-based auditing with a two-stage algorithm that weights contributors by the stability of their past residuals.
result Minority contributors' evaluations drift towards the majority, and their participation share falls on controversial topics.