Selective planning with imperfect models reduces harmful effects of model inadequacy.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Information planning enables faster learning with fewer training examples. It is particularly applicable when training examples are costly to obtain. This work examines the advantages of information planning for text data by focusing on three supervised models: Naive Bayes, supervised LDA and deep neural networks. We s…
Statistical depth metrics help identify risky power grid scenarios.
Unified view on selective credit assignment for reinforcement learning.
We introduce Dynamic Planning Networks (DPN), a novel architecture for deep reinforcement learning, that combines model-based and model-free aspects for online planning. Our architecture learns to dynamically construct plans using a learned state-transition model by selecting and traversing between simulated states and…
New framework guides resource usage to achieve sublinear regret in adversarial settings.
Automated planning is one of the foundational areas of AI. Since no single planner can work well for all tasks and domains, portfolio-based techniques have become increasingly popular in recent years. In particular, deep learning emerges as a promising methodology for online planner selection. Owing to the recent devel…
The paper explores how planning with models improves credit assignment in reinforcement learning.
PS framework selects best policy from library for CSO problems.
Study experiment planning with function approximation in contextual bandit problems.
New method selects critical DER scenarios for distribution grid investment planning.
It is well established that humans decision making and instrumental control uses multiple systems, some which use habitual action selection and some which require deliberate planning. Deliberate planning systems use predictions of action-outcomes using an internal model of the agent's environment, while habitual action…
New method for selecting clusters in residential electricity data.
DDPD separates generation into planning and denoising for improved efficiency.
In the context of tree-search stochastic planning algorithms where a generative model is available, we consider on-line planning algorithms building trees in order to recommend an action. We investigate the question of avoiding re-planning in subsequent decision steps by directly using sub-trees as action recommender. …
An accurate load forecasting has always been one of the main indispensable parts in the operation and planning of power systems. Among different time horizons of forecasting, while short-term load forecasting (STLF) and long-term load forecasting (LTLF) have respectively got benefits of accurate predictors and probabil…
Proposes FROT for high-dimensional data, avoiding curse of dimensionality.
Constructing agents with planning capabilities has long been one of the main challenges in the pursuit of artificial intelligence. Tree-based planning methods have enjoyed huge success in challenging domains, such as chess and Go, where a perfect simulator is available. However, in real-world problems the dynamics gove…
This paper optimizes DC pension plan investments using O-U process and loan.
A microscopic approach to macroeconomic features is intended. A model for macroeconomic behavior under heterogeneous spatial economic conditions is reviewed. A birth-death lattice gas model taking into account the influence of an economic environment on the fitness and concentration evolution of economic entities is nu…
Study shows how optimal transport behaves in higher dimensions.
In this paper, we solve the arms exponential exploding issue in multivariate Multi-Armed Bandit (Multivariate-MAB) problem when the arm dimension hierarchy is considered. We propose a framework called path planning (TS-PP) which utilizes decision graph/trees to model arm reward success rate with m-way dimension interac…
PLANS synthesizes programs from noisy inputs using neural specs and filtering.
The automatic digitizing of paper maps is a significant and challenging task for both academia and industry. As an important procedure of map digitizing, the semantic segmentation section mainly relies on manual visual interpretation with low efficiency. In this study, we select urban planning maps as a representative …
Director learns hierarchical behaviors from pixels, outperforming exploration methods.
Computing optimal transport (OT) between measures in high dimensions is doomed by the curse of dimensionality. A popular approach to avoid this curse is to project input measures on lower-dimensional subspaces (1D lines in the case of sliced Wasserstein distances), solve the OT problem between these reduced measures, a…
An investor with constant relative risk aversion and an infinite planning horizon trades a risky and a safe asset with constant investment opportunities, in the presence of small transaction costs and a binding exogenous portfolio constraint. We explicitly derive the optimal trading policy, its welfare, and implied tra…
Active localization is the problem of generating robot actions that allow it to maximally disambiguate its pose within a reference map. Traditional approaches to this use an information-theoretic criterion for action selection and hand-crafted perceptual models. In this work we propose an end-to-end differentiable meth…
The paper evaluates criteria for selecting cryptocurrencies based on historical data.
Proposes new models to predict student grades more accurately.
A new method for efficient inference and model selection in SBMs using OT.
Random investment strategies outperform sensible ones, even with forecasts.
Study on investment strategy for agents with periodic preferences and discounting.
GraSP-RL uses graph neural networks to improve job shop scheduling.
In an online contract selection problem there is a seller which offers a set of contracts to sequentially arriving buyers whose types are drawn from an unknown distribution. If there exists a profitable contract for the buyer in the offered set, i.e., a contract with payoff higher than the payoff of not accepting any c…
New method designs fairer transport plans with uncertainty.
Improves RL planning by proposing sub-goals hierarchically.
In this short note we consider a dynamic assortment planning problem under the capacitated multinomial logit (MNL) bandit model. We prove a tight lower bound on the accumulated regret that matches existing regret upper bounds for all parameters (time horizon , number of items and maximum assortment capacity )…
Power load forecast with Machine Learning is a fairly mature application of artificial intelligence and it is indispensable in operation, control and planning. Data selection techniqies have been hardly used in this application. However, the use of such techniques could be beneficial provided the assumption that the da…
Complex networks are often either too large for full exploration, partially accessible, or partially observed. Downstream learning tasks on these incomplete networks can produce low quality results. In addition, reducing the incompleteness of the network can be costly and nontrivial. As a result, network discovery algo…
Robust forecast framework reduces distribution error by 63%.
We aim to reduce the burden of programming and deploying autonomous systems to work in concert with people in time-critical domains, such as military field operations and disaster response. Deployment plans for these operations are frequently negotiated on-the-fly by teams of human planners. A human operator then trans…
Study integrates reliability constraints into generation planning models.
New approach improves black-box planning efficiency by discovering focused macros.
Study motion planning for points avoiding obstacles in a plane.
Model-based reinforcement learning is an appealing framework for creating agents that learn, plan, and act in sequential environments. Model-based algorithms typically involve learning a transition model that takes a state and an action and outputs the next state---a one-step model. This model can be composed with itse…
CoMPNetX uses neural networks to efficiently solve constrained motion planning problems.
This article asks how planning scholarship may effectively gain impact in planning practice through media exposure. In liberal democracies the public sphere is dominated by mass media. Therefore, working with such media is a prerequisite for effective public impact of planning research. Using the example of megaproject…