In many platforms, user arrivals exhibit a self-reinforcing behavior: future user arrivals are likely to have preferences similar to users who were satisfied in the past. In other words, arrivals exhibit positive externalities. We study multiarmed bandit (MAB) problems with positive externalities. We show that the self…
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
New algorithms reduce costly feature collection in bandits.
New approach to disentangle utility from impulse in recommendation systems.
Many recommendation algorithms rely on user data to generate recommendations. However, these recommendations also affect the data obtained from future users. This work aims to understand the effects of this dynamic interaction. We propose a simple model where users with heterogeneous preferences arrive over time. Based…
In order for an e-commerce platform to maximize its revenue, it must recommend customers items they are most likely to purchase. However, the company often has business constraints on these items, such as the number of each item in stock. In this work, our goal is to recommend items to users as they arrive on a webpage…
Novel mean estimation method under user-level differential privacy reduces noise in continual mean estimates.
Algorithm learns diverse rankings for search engines.
The paper proposes a survival model to optimize mobile notification delivery times.
We study an online multi-task learning setting, in which instances of related tasks arrive sequentially, and are handled by task-specific online learners. We consider an algorithmic framework to model the relationship of these tasks via a set of convex constraints. To exploit this relationship, we design a novel algori…
Real-time personalization for HAR models learns from new users without prior data.
Recommender systems take inputs from user history, use an internal ranking algorithm to generate results and possibly optimize this ranking based on feedback. However, often the recommender system is unaware of the actual intent of the user and simply provides recommendations dynamically without properly understanding …
New algorithms improve decision-making with limited offline data.
Traffic prediction plays a vital role in efficient planning and usage of network resources in wireless networks. While traffic prediction in wired networks is an established field, there is a lack of research on the analysis of traffic in cellular networks, especially in a content-blind manner at the user level. Here, …
Machine Learning (ML) models trained on data from multiple demographic groups can inherit representation disparity (Hashimoto et al., 2018) that may exist in the data: the model may be less favorable to groups contributing less to the training process; this in turn can degrade population retention in these groups over …
In this paper, we consider decentralized sequential decision making in distributed online recommender systems, where items are recommended to users based on their search query as well as their specific background including history of bought items, gender and age, all of which comprise the context information of the use…
Interpretable machine learning tackles the important problem that humans cannot understand the behaviors of complex machine learning models and how these models arrive at a particular decision. Although many approaches have been proposed, a comprehensive understanding of the achievements and challenges is still lacking…
Mobile edge computing (MEC) emerges recently as a promising solution to relieve resource-limited mobile devices from computation-intensive tasks, which enables devices to offload workloads to nearby MEC servers and improve the quality of computation experience. Nevertheless, by considering a MEC system consisting of mu…
We tackle the problem of building explainable recommendation systems that are based on a per-user decision tree, with decision rules that are based on single attribute values. We build the trees by applying learned regression functions to obtain the decision rules as well as the values at the leaf nodes. The regression…
Inferring intent from observed behavior has been studied extensively within the frameworks of Bayesian inverse planning and inverse reinforcement learning. These methods infer a goal or reward function that best explains the actions of the observed agent, typically a human demonstrator. Another agent can use this infer…
Study ruin probabilities in risk processes on stochastic networks.
Paper shows how to quantify uncertainty in medical ML models.
New model optimizes assortment and pricing with dynamic customer arrivals.
Automatic estimation of relative difficulty of a pair of questions is an important and challenging problem in community question answering (CQA) services. There are limited studies which addressed this problem. Past studies mostly leveraged expertise of users answering the questions and barely considered other properti…
We tackle the problem of collaborative filtering (CF) with side information, through the lens of Gaussian Process (GP) regression. Driven by the idea of using the kernel to explicitly model user-item similarities, we formulate the GP in a way that allows the incorporation of low-rank matrix factorisation, arriving at o…
By leveraging the concept of mobile edge computing (MEC), massive amount of data generated by a large number of Internet of Things (IoT) devices could be offloaded to MEC server at the edge of wireless network for further computational intensive processing. However, due to the resource constraint of IoT devices and wir…
Study off-policy evaluation and learning in dynamic pricing with context.
Algorithm solves job acceptance problem with random arrivals and values.
Sequential screening and dynamic regret in multi-armed bandits with arriving arms
Proves Arnold-Thom conjecture for surfaces' arrival times.
This paper improves indoor positioning accuracy by deploying reference nodes to ensure Line-of-Sight.
Estimate arrival times in random recursive trees using iterated Jordan centralities.
ADER addresses continual learning in session-based recommendation by periodically replaying exemplars with adaptive distillation.
We consider a classical risk process with arrival of claims following a non-stationary Hawkes process. We study the asymptotic regime when the premium rate and the baseline intensity of the claims arrival process are large, and claim size is small. The main goal of the article is to establish a diffusion approximation …
This paper tackles inventory control with general arrival dynamics and post-processing, improving profitability.
In this paper, we introduce Ballooning Multi-Armed Bandits (BL-MAB), a novel extension of the classical stochastic MAB model. In the BL-MAB model, the set of available arms grows (or balloons) over time. In contrast to the classical MAB setting where the regret is computed with respect to the best arm overall, the regr…
In this paper, we obtain the finite-horizon and infinite-horizon ruin probability asymptotics for risk processes with claims of subexponential tails for non-stationary arrival processes that satisfy a large deviation principle. As a result, the arrival process can be dependent, non-stationary and non-renewal. We give t…
Study on queues with Hawkes arrivals, proving steady-state behavior and developing an efficient algorithm.
Optimal fund deployment strategy under uncertain deal arrivals.
We studied non-dynamical stochastic resonance for the number of trades in the stock market. The trade arrival rate presents a deterministic pattern that can be modeled by a cosine function perturbed by noise. Due to the nonlinear relationship between the rate and the observed number of trades, the noise can either enha…
In this paper, we address the Online Unsupervised Domain Adaptation (OUDA) problem, where the target data are unlabelled and arriving sequentially. The traditional methods on the OUDA problem mainly focus on transforming each arriving target data to the source domain, and they do not sufficiently consider the temporal …
In a previous analysis the problem of "zero-inflated" time data (caused by high frequency trading in the electronic order book) was handled by left-truncating the inter-arrival times. We demonstrated, using rigorous statistical methods, that the Weibull distribution describes the corresponding stochastic dynamics for a…
For a monotonically advancing front, the arrival time is the time when the front reaches a given point. We show that it is twice differentiable everywhere with uniformly bounded second derivative. It is smooth away from the critical points where the equation is degenerate. We also show that the critical set has finite …
A new framework enhances binaural audio for moving talkers.
Order book dynamics play an important role in both execution time and price formation of orders in an exchange market. In this study, we aim to model the limit order arrival rates in the vicinity of the best bid and the best ask price levels. We use limit order book data for Garanti Bank, which is one of the most trade…
Study models market volatility with persistent and temporary impacts.
A learning-based algorithm optimizes admission control in a queuing system.
Proposes a new simulator for complex arrival processes.
We introduce a multivariate Hawkes process that accounts for the dynamics of market prices through the impact of market order arrivals at microstructural level. Our model is a point process mainly characterized by 4 kernels associated with respectively the trade arrival self-excitation, the price changes mean reversion…