New DP training ensures models behave similarly at training and test time.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Increasingly, discrimination by algorithms is perceived as a societal and legal problem. As a response, a number of criteria for implementing algorithmic fairness in machine learning have been developed in the literature. This paper proposes the Continuous Fairness Algorithm (CFA) which enables a continuous interpol…
This paper is an attempt to explain all the matrix calculus you need in order to understand the training of deep neural networks. We assume no math knowledge beyond what you learned in calculus 1, and provide links to help you refresh the necessary math where needed. Note that you do not need to understand this materia…
New definition reveals encoding explanations that retain predictive power.
The hidden tail of empirical distributions is analyzed using extreme value theory.
New UCB algorithms resist contamination in bandit problems.
With the dawn of the Big Data era, data sets are growing rapidly. Data is streaming from everywhere - from cameras, mobile phones, cars, and other electronic devices. Clustering streaming data is a very challenging problem. Unlike the traditional clustering algorithms where the dataset can be stored and scanned multipl…
What would you do if you were invited to play a game where you were given \$25 and allowed to place bets for 30 minutes on a coin that you were told was biased to come up heads 60% of the time? This is exactly what we did, gathering 61 young, quantitatively trained men and women to play this game. The results, in a nut…
The paper explores the limits of tight PAC-Bayes bounds for cheap models in robust statistics.
With the advancement in argument detection, we suggest to pay more attention to the challenging task of identifying the more convincing arguments. Machines capable of responding and interacting with humans in helpful ways have become ubiquitous. We now expect them to discuss with us the more delicate questions in our w…
Facial attribute editing aims to manipulate single or multiple attributes of a face image, i.e., to generate a new face with desired attributes while preserving other details. Recently, generative adversarial net (GAN) and encoder-decoder architecture are usually incorporated to handle this task with promising results.…
Supervised machine learning models boast remarkable predictive capabilities. But can you trust your model? Will it work in deployment? What else can it tell you about the world? We want models to be not only good, but interpretable. And yet the task of interpretation appears underspecified. Papers provide diverse and s…
What return should you expect when you take on a given amount of risk? How should that return depend upon other people's behavior? What principles can you use to answer these questions? In this paper, we approach these topics by exploring the consequences of two simple hypotheses about risk. The first is a common-sense…
Suppose you have one unit of stock, currently worth 1, which you must sell before time . The Optional Sampling Theorem tells us that whatever stopping time we choose to sell, the expected discounted value we get when we sell will be 1. Suppose however that we are able to see units of time into the future, and ba…
We consider two well known constructions of link invariants. One uses skein theory: you resolve each crossing of the link as a linear combination of things that don't cross, until you eventually get a linear combination of links with no crossings, which you turn into a polynomial. The other uses quantum groups: you con…
Can certain shapes be drawn with a pencil and eraser?
We construct a pair of transverse genuine laminations on an atoroidal 3-manifold admitting transversely orientable uniform 1-cochain. The laminations are induced by the uniform 1-cochain and they are indeed the "straightening" of the coarse laminations defined in [Ca], by using minimal surface techniques. Moreover, whe…
F.: Good morning Hermann, I would like to talk with you about infinitesimals. G.: Tell me Pierre. F.: I'm fed up of all these slanders about my attitude to be non rigorous, so I've started to study nonstandard analysis (NSA) and synthetic differential geometry (SDG). G.: Yes, I've read something ... F.: Ok, no problem …
New approach to counterfactual reasoning avoids demographic interventions.
Have you ever looked at a machine learning classification model and thought, I could have made that? Well, that is what we test in this project, comparing XGBoost trained on human engineered features to training directly on data. The human engineered features do not outperform XGBoost trained di- rectly on the data, bu…
In this report, we talked about a new quantitative strategy for choosing the optimal(s) stock(s) to trade. The basic notions are generally very known by the financial community. The key here is to understand 1) the standard score applied to a sample and 2) the correlation factor applied to different time series in real…
Take a torus with a Riemannian metric. Lift the metric on its universal cover. You get a distance which in turn yields balls. On these balls you can look at the Laplacian. Focus on the spectrum for the Dirichlet or Neumann problem. We describe the asymptotic behaviour of the eigenvalues as the radius of the balls goes …
Study uses chatbot to understand users' needs for ML model explanations.
Despite the widespread usage of machine learning throughout organizations, there are some key principles that are commonly missed. In particular: 1) There are at least four main families for supervised learning: logical modeling methods, linear combination methods, case-based reasoning methods, and iterative summarizat…
MLE works best for covariate shift without modifications.
Solves a problem about deforming symplectic forms on a Klein bottle.
Improves deep learning for Airbnb search ranking.
GP model calibration improves optimization algorithm performance.
Quant firms manipulate stock markets overnight and intraday.
Google Trends can lead to misleading forecasts if not used carefully.
The paper defines and studies the category of Z-graded manifolds, including their intrinsic structure and formal properties.
This paper takes stock of megaproject management, an emerging and hugely costly field of study. First, it answers the question of how large megaprojects are by measuring them in the units mega, giga, and tera, concluding we are presently entering a new "tera era" of trillion-dollar projects. Second, total global megapr…
Silence on suspicious stock market patterns persists despite lack of plausible explanations.
Study shows surfaces sound the same everywhere if they have a transitive isometry group.
Paper proposes a new approach to GDPR compliance using data protection analytics.
Paper proposes a modified uncertainty sampling method to speed up preference learning from noisy humans.
Complex behaviors are often driven by an internal model, which integrates sensory information over time and facilitates long-term planning. Inferring an agent's internal model is a crucial ingredient in social interactions (theory of mind), for imitation learning, and for interpreting neural activities of behaving agen…
Explains agent behavior through intended outcomes in reinforcement learning.
The paper introduces a new insurance pricing model based on driving mileage.
In this small article one compromise monetization strategy is proposed, which hopefully may lead to a more satisfactory coexistence of IP manufacturers and consumers. The motto is "fair exchange": you use our IP-product, we use your product (in form of money); when you do not need our product any more, we change back.
As technology become more advanced, those who design, use and are otherwise affected by it want to know that it will perform correctly, and understand why it does what it does, and how to use it appropriately. In essence they want to be able to trust the systems that are being designed. In this survey we present assura…
A new method to value IPOed companies.
Lecture notes on geodesics in differential geometry.
We reconsider the su(3) link homology theory defined by Khovanov in math.QA/0304375 and generalized by Mackaay and Vaz in math.GT/0603307. With some slight modifications, we describe the theory as a map from the planar algebra of tangles to a planar algebra of (complexes of) `cobordisms with seams' (actually, a `canopo…
Active-memory mechanisms can replace self-attention in Transformers, but optimal results often require both.
We consider a class of auctions (Lowest Unique Bid Auctions) that have achieved a considerable success on the Internet. Bids are made in cents (of euro) and every bidder can bid as many numbers as she wants. The lowest unique bid wins the auction. Every bid has a fixed cost, and once a participant makes a bid, she gets…
In the process of exploring the world, the curiosity constantly drives humans to cognize new things. Supposing you are a zoologist, for a presented animal image, you can recognize it immediately if you know its class. Otherwise, you would more likely attempt to cognize it by exploiting the side-information (e.g., seman…
Local explanation frameworks aim to rationalize particular decisions made by a black-box prediction model. Existing techniques are often restricted to a specific type of predictor or based on input saliency, which may be undesirably sensitive to factors unrelated to the model's decision making process. We instead propo…