Two novel algorithms improve distributed machine learning in the presence of Byzantine adversaries.
problem Improving distributed machine learning in the presence of Byzantine adversaries.
method Two novel stochastic gradient descent algorithms, ByGARS and ByGARS++, using reputation scores for gradient aggregation.
result Robust to any number of multiplicative noise Byzantine adversaries and converge for strongly convex loss functions.
Well-defined formal definitions for sentiment and opinion are extended to incorporate the necessary elements to provide a formal quantitative definition of reputation. This definition takes the form of a time-based index, in which each element is a function of a collection of opinions mined during a given time period. …
Crowdsourced algorithms identify fake news on Twitter.
problem Identifying fake news on social media platforms.
method Evaluation of reputation algorithms on a large dataset of Twitter news.
result Simple crowdsourcing-based algorithms can identify a significant portion of fake news with low false positive rates.
Model analyzes OTC market making with reputation feedback.
problem Optimizing electronic OTC liquidity provision considering reputation.
method Developed a stochastic-control model with feedback loops.
result Policy alternates between reputation-building and franchise monetization phases.
Model analyzes how reputation feedback affects OTC market making.
problem Understanding and optimizing OTC market making strategies.
method Developed a stochastic-control model with feedback loops.
result Policy alternates between reputation building and franchise monetization.
Study shows social media impacts shareholder returns on ESG risks.
problem Investor sentiment and public opinion on ESG risks.
method Event study design using social media data.
result Statistically significant reduction in abnormal returns after ESG-risk events.
Reducing barriers to entry in large-scale ML markets, study shows multi-objective learning can lower data requirements.
problem Barriers to entry in emerging markets for large-scale machine learning models.
method Defined a multi-objective high-dimensional regression framework to study reputational damage and data requirements.
result The number of data points needed for a new company to enter the market can be significantly smaller than the incumbent company's dataset size.
A new federated learning framework ensures fairness and robustness.
problem Collaborative fairness and adversarial robustness in federated learning.
method RFFL framework with a reputation mechanism to identify and remove non-contributing or malicious participants.
result RFFL achieves high fairness and robustness to different types of adversaries.
Framework scores DeFi users based on liquidity and trading behavior.
problem Distinguishing between liquidity provision and active trading in DeFi.
method Rule-based decomposition, deep residual neural network, pool-level context.
result Deep residual neural network improves user scoring and risk assessment.
We propose a continuum model for the description of buyer and seller dynamics in an Internet market. The relevant variables are the research effort of buyers and the sellers' reputation building process. We show that, if a commercial web-site gives consumers the possibility to rate credibly sellers they bargained with,…
New approach simulates reputation dynamics using information compression.
problem Malicious communication strategies in reputation networks.
method Uses information compression techniques to simulate social phenomena.
result Emergent phenomena like echo chambers and deception are observed.
Model analyzes how firms balance full disclosure with selective disclosure to maintain a good reputation.
problem Managing reputation in financial markets through voluntary disclosure.
method Developed a dynamic model with two disclosure strategies: candid and sparing, using a piecewise-deterministic model.
result Firms are rewarded for full disclosure but may switch to selective disclosure to avoid potential downgrades.
New algorithm reduces regret in strategic prediction problem.
problem Designing an IC algorithm with sublinear regret for strategic experts.
method Developed a new algorithm WSU-UX and proved a worst-case regret bound.
result WSU-UX suffers a Ω(T2/3) lower bound on regret. As the number of contributors to online peer-production systems grows, it becomes increasingly important to predict whether the edits that users make will eventually be beneficial to the project. Existing solutions either rely on a user reputation system or consist of a highly specialized predictor that is tailored to …
The persistence of racial inequality in the U.S. labor market against a general backdrop of formal equality of opportunity is a troubling phenomenon that has significant ramifications on the design of hiring policies. In this paper, we show that current group disparate outcomes may be immovable even when hiring decisio…
CFFL framework improves fairness in FL without sacrificing accuracy.
problem Overfitting and lack of collaborative fairness in Federated Learning.
method CFFL framework uses reputation to ensure participants converge to different models.
result CFFL achieves high fairness, comparable accuracy, and better performance than Standalone and Distributed frameworks.
Machine learning methods have gained a great deal of popularity in recent years among public administration scholars and practitioners. These techniques open the door to the analysis of text, image and other types of data that allow us to test foundational theories of public administration and to develop new theories. …
Android and Facebook provide third-party applications with access to users' private data and the ability to perform potentially sensitive operations (e.g., post to a user's wall or place phone calls). As a security measure, these platforms restrict applications' privileges with permission systems: users must approve th…
The paper proposes a fair and private decentralized deep learning framework.
problem Ensuring fairness and privacy in collaborative deep learning.
method A reputation system and differential privacy are used. FDPDDL framework is built with two stages: initialisation and update.
result FDPDDL achieves high fairness, comparable accuracy to centralised and distributed frameworks, and better accuracy than standalone.
Both generative adversarial networks (GAN) in unsupervised learning and actor-critic methods in reinforcement learning (RL) have gained a reputation for being difficult to optimize. Practitioners in both fields have amassed a large number of strategies to mitigate these instabilities and improve training. Here we show …
Research shows higher damages may encourage more disclosure in corporate disputes.
problem How to resolve disputes over undisclosed material events in a way that encourages voluntary disclosure.
method Dynamic continuous-time model of management's equilibrium disclosure decision.
result Increased damages may lead to an endogenous increase in voluntary disclosure.
System filters inappropriate YouTube content for advertisers.
problem Inadequate detection of inappropriate content on YouTube ads.
method Proposes a system for identifying and filtering inappropriate content.
result Current countermeasures are ineffective in detecting inappropriate content.
Paper uses queue theory to model financial signals with relativistic delay.
problem Relativistic delay in financial trading signals.
method Modified M/M/G queue theory.
result Describes propagation of trading signals with finite velocity.
Nonnegative matrix factorization (NMF) has an established reputation as a useful data analysis technique in numerous applications. However, its usage in practical situations is undergoing challenges in recent years. The fundamental factor to this is the increasingly growing size of the datasets available and needed in …
Optimizes package types for e-commerce to reduce damage and costs.
problem Sub-optimal package types lead to damaged shipments and high costs.
method Multi-stage approach that balances shipment and damage costs using a scalable algorithm.
result Significant cost savings of tens of millions of dollars achieved.
Neural network models have a reputation for being black boxes. We propose to monitor the features at every layer of a model and measure how suitable they are for classification. We use linear classifiers, which we refer to as "probes", trained entirely independently of the model itself. This helps us better understand …
A dealer manages quotes and rejection rules to control slippage risk in FX markets.
problem Managing inventory risk and latency risk in OTC FX market making.
method Dynamic programming and adiabatic-quadratic approximation to optimize quotes and rejection rules.
result Developed a method to optimize quotes and rejection rules for managing slippage risk.
The paper assesses fairness in AI for financial services, using statistical methods.
problem Unintentional bias and insufficient model validation in AI applications.
method Statistical methods for imbalanced data treatment and bias mitigation.
result Fairness evaluation metrics applied to a credit card default payment example.
This paper clarifies Bitcoin's volatility and predictability across daily, weekly, and monthly scales.
problem Clarify Bitcoin's volatility and predictability across different time scales.
method Using daily, weekly, and monthly closing prices and log-returns data, analyze volatility and predictability.
result Bitcoin exhibits high volatility and high predictability, with different behaviors at different time scales.
Clarifies model-based RL's theoretical issues and counterexamples for popular losses.
problem Model-based reinforcement learning's empirical performance vs. theoretical properties and popular loss functions.
method Analyzes empirical and theoretical aspects of model-based RL and constructs counterexamples for losses.
result MuZero loss fails in stochastic and deterministic environments, leading to exponential sample complexity.
The study explores when it's best to remove a real estate broker from the process.
problem Optimal conditions for removing a real estate broker.
method New models analyzing information asymmetry, social capital, and contract types.
result Dis-intermediation can be optimal under certain conditions.
This paper proposes a new hashing-based KNN technique for faster nearest neighbor selection.
problem Slowness of KNN in big datasets due to searching entire dataset.
method Divide data space into subcells, use hashing to map data points, and select nearest neighbors layer by layer.
result The proposed technique offers competitive performance with KNN and KDtree while significantly improving time efficiency.
Delegated votes in Uniswap DAO favor parties with less self-owned votes and a16z-affiliated entities.
problem Incentives for vote delegation in decentralized governance systems.
method Analysis of Uniswap governance DAO using vote delegation data.
result Vote delegation patterns suggest window-dressing around decentralization and merit-based delegation.
Reproducibility of modeling is a problem that exists for any machine learning practitioner, whether in industry or academia. The consequences of an irreproducible model can include significant financial costs, lost time, and even loss of personal reputation (if results prove unable to be replicated). This paper will fi…
A technique to quickly fix mistakes in neural networks.
problem Fixing model errors in neural networks quickly and without affecting other samples.
method Editable Training, a model-agnostic training technique.
result Effectiveness demonstrated on large-scale image classification and machine translation tasks.
Islamophobic hate speech on social media inflicts considerable harm on both targeted individuals and wider society, and also risks reputational damage for the host platforms. Accordingly, there is a pressing need for robust tools to detect and classify Islamophobic hate speech at scale. Previous research has largely ap…
Nostradamus links climate and stock market performance.
problem Understanding the impact of climate on stock prices.
method Analyzing historical data, climate indicators, and natural disasters.
result Significant correlation between climate and stock price fluctuations.
Tests for Esophageal cancer can be expensive, uncomfortable and can have side effects. For many patients, we can predict non-existence of disease with 100% certainty, just using demographics, lifestyle, and medical history information. Our objective is to devise a general methodology for customizing tests using user pr…
Detecting faults and SLA violations in a timely manner is critical for telecom providers, in order to avoid loss in business, revenue and reputation. At the same time predicting SLA violations for user services in telecom environments is difficult, due to time-varying user demands and infrastructure load conditions. In…
The problem of searching for experts in a given academic field is hugely important in both industry and academia. We study exactly this issue with respect to a database of authors and their publications. The idea is to use Latent Semantic Indexing (LSI) and Latent Dirichlet Allocation (LDA) to perform topic modelling i…
Crowdsourcing systems, in which numerous tasks are electronically distributed to numerous "information piece-workers", have emerged as an effective paradigm for human-powered solving of large scale problems in domains such as image classification, data entry, optical character recognition, recommendation, and proofread…
Neural networks predict TED Talk ratings from transcripts, removing bias.
problem Predicting public speaking performance from speech transcripts.
method Causal diagram modeling, word sequence and dependency tree based neural networks.
result Average F-score of 0.77, significantly outperforming baseline methods.
Research creates a taxonomy to bridge AI security and regulatory gaps.
problem Disciplinary disconnect between technical and legal teams in AI risk assessment.
method Developed an AI System Threat Vector Taxonomy with 9 domains and 53 sub-threats.
result Empirically validated and aligned with ISO/IEC 42001 controls and NIST AI RMF functions.
New algorithms improve neural network verification by exploiting piecewise linear structure.
problem Efficiently verify the correctness of neural networks, especially those with high-dimensional inputs.
method Branch-and-Bound (BaB) framework applied to Mixed Integer Linear Programming (MIP) formulation.
result Significant performance improvements and new branching strategies for neural networks.
Machine learning impacts computational math, offering new functions approximations.
problem Machine learning's black box nature hinders further progress in computational math.
method Analyzes machine learning's impact on computational math and vice versa.
result Integrating computational math with machine learning can enhance both fields.
In 1970, E. M. Andreev published a classification of all three-dimensional compact hyperbolic polyhedra having non-obtuse dihedral angles. Given a combinatorial description of a polyhedron, C, Andreev's Theorem provides five classes of linear inequalities, depending on C, for the dihedral angles, which are necessar…
Paper detects review abuse using tensor decomposition.
problem Detecting review abuse by sellers and reviewers.
method Semi-supervised binary multi-target tensor decomposition.
result The model achieves higher precision and recall.
Model shows how advisors can manipulate naive investors.
problem How financial advisors manipulate naive investors.
method Agent-Based Model with Nash equilibria and best response functions.
result Greediness/naivety of investors emerge naturally from the model.