Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

169,341 papers · 148 categories

Trend · papers per month

61122183244 · Jun 202019922001200920182026
48 results for transport policies

A new framework for generative modeling using value-driven transport.

problem Developing efficient methods for generative modeling.
method A discrete-time stochastic control formulation of measure transport, formulated as a linear program with dual variables corresponding to the optimal value function.
result Well-trained VDT policies lead to straight transport paths that can be simulated quickly and robustly.

New algorithm reduces variance in Monte Carlo simulations using deep neural networks and policy gradients.

problem Reducing variance in Monte Carlo simulations for estimating function values.
method Optimal correlation search using deep neural networks and policy gradients.
result Optimal correlation function reduces variance by approximating and calibrating policy.

Optimizes taxi carpool policies using RL and spatio-temporal data.

problem Reduce car usage and minimize traffic congestion through efficient carpooling.
method Developed a deep neural network (ST-NN) for trip time prediction and a carpooling RL simulation environment.
result RL learned policy outperformed a fixed policy in maximizing efficiency and minimizing congestion.

Developed a cost and revenue model for HEMS to estimate breakeven transport volumes under different reimbursement and labor cost assumptions.

problem Estimating breakeven transport volumes for HEMS under varying reimbursement and labor cost assumptions.
method Developed a two-part model: cost framework and actuarial revenue model using healthcare encounter data and payer reimbursement rates.
result Estimated breakeven transport volumes under different reimbursement and labor cost assumptions.

Neural Index Policy for multi-action bandits with heterogeneous budgets.

problem Real-world settings often involve multiple interventions with heterogeneous costs and constraints, breaking classical assumptions.
method Introduces a Neural Index Policy (NIP) that learns to assign budget-aware indices to arm-action pairs using a neural network and differentiable knapsack layer.
result Empirically achieves near-optimal performance while strictly enforcing heterogeneous budgets and scaling to hundreds of arms.

New graph distances derived from optimal transport framework using path flows.

problem Develop new graph distances for clustering and classification.
method Bag-of-paths framework with Gibbs-Boltzmann distribution and optimal transport relaxation.
result Interpolates between shortest-path and resistance distances, improving performance.

Paper introduces OTR for efficient offline RL in surgical robotics.

problem Lack of annotated datasets for offline RL in surgical robotics.
method OTR algorithm using Optimal Transport to assign rewards to unlabeled trajectories.
result OTR enables efficient policy learning from large datasets without handcrafted rewards.

This paper improves transportation efficiency by teaching automated vehicles to cooperate.

problem Improving efficiency and safety of transportation systems with automated vehicles.
method Multi-agent graph reinforcement learning with attention mechanism.
result Automated vehicles can achieve better performance when learning to cooperate with each other.

The paper provides theoretical guarantees for behavior cloning using generative models.

problem Behavior cloning of complex expert demonstrations using generative models.
method The paper proposes a theoretical framework invoking low-level controllers to stabilize imitation around expert demonstrations. It shows that with suitable low-level stability guarantees and powerful generative models, pure supervised behavior cloning can match expert trajectories.
result The paper proves that with a suitable low-level stability guarantee and a powerful enough generative model, pure supervised behavior cloning can generate trajectories matching the per-time step distribution of essentially arbitrary expert trajectories in an optimal transport cost.

Paper forecasts dynamic transportation networks using probabilistic models.

problem Forecasting temporal evolution of transportation networks.
method Probabilistic latent network model with Bayesian inference.
result Models accurately predict future network states and community structures.

A new method uses optimal transport for imitation learning, improving efficiency and performance.

problem Recovering expert policies from demonstrations with efficient and general reward functions.
method Proposes Wasserstein Adversarial Imitation Learning, using Kantorovich potentials as reward functions and regularized optimal transport.
result Significantly improves sample-efficiency and average cumulative rewards in robotic experiments.

This research optimizes feature selection for predicting transportation modes in smart cities.

problem Finding the best subset of features for predicting transportation modes.
method Wrapper and information retrieval methods were used to find the best feature subset.
result The proposed framework achieved better performance compared to related studies.

Proposes a method to correct for covariate shift in meta-analysis of randomized trials.

problem Invalidation of standard IPD meta-analysis due to covariate shift across studies.
method Placebo-anchored transport framework that treats source-trial outcomes as proxy signals and target-trial placebo outcomes as gold labels.
result Yields target-identified effect estimates in connected targets and a principled screen--then--transport procedure in disconnected targets.

Study shows ethanol blends and incentives can significantly reduce transportation carbon emissions.

problem Rapid growth in electric vehicles requires complementary strategies to decarbonize transportation.
method Analysis of ethanol blending, regulatory incentives, and economic assessments.
result Ethanol blending, especially E15 and E85, can substantially reduce carbon emissions and provide economic benefits.

Study of repeated games with unobserved agent rewards using MAB framework.

problem Designing policies for principals in repeated principal-agent games with unobservable agent rewards.
method Developed a policy achieving low regret (square-root regret up to a log factor) for perfect-knowledge agents.
result Constructed an estimator for agent's expected reward and designed a policy achieving low regret.

CJE calibrates cheap LLM judges against an oracle, achieving high accuracy at a fraction of the cost.

problem Inexpensive LLM judges can produce biased rankings, leading to unreliable outcomes.
method CJE uses a small oracle to calibrate cheap scores, then evaluates at scale with valid uncertainty.
result CJE achieves 99% pairwise ranking accuracy at 14x lower cost compared to a 16x oracle/judge cost ratio.

Deep learning improves trip prediction accuracy in transportation planning.

problem Traditional models fail to accurately predict person and vehicle trips due to complexity and dynamics.
method Developed and trained a deep learning model using NHTS data.
result Deep learning model achieved 98% accuracy for person trip prediction and 96% for vehicle trip estimation.

This article presents results from the first statistically significant study of cost escalation in transportation infrastructure projects. Based on a sample of 258 transportation infrastructure projects worth US$90 billion and representing different project types, geographical regions, and historical periods, it is fou…

2013-03-06abs ↗pdf ↗

Active learning selects optimal measurement times for inferring continuous paths from sparse data.

problem Inferring continuous probability paths from sparse snapshots in high-fidelity domains like single-cell biology.
method Extends active experimentation to the space of measures using Linearized Optimal Transport (LOT) for probabilistic surrogate modeling.
result Empirical results show that the proposed strategy outperforms uncertainty-agnostic baselines.

SDPA is shown to be an optimal transport problem in deep learning.

problem The mathematical foundation and optimization perspective of SDPA.
method SDPA is shown to be the exact solution to a degenerate, one-sided Entropic Optimal Transport (EOT) problem.
result The SDPA mechanism is a principled mechanism where the forward pass performs optimal inference and the backward pass implements a rational, manifold-aware learning update.

This paper tackles online strategic decision making with asymmetry and knowledge transportability.

problem Strategic decision making with information asymmetry and knowledge transportability challenges.
method Developed a sample-efficient algorithm for online learning under these conditions.
result Proved sample complexity of O(1/ε2)O(1/ε^2) for learning an εε-optimal policy.

A new framework uses multi-agent reinforcement learning for evaluating policies in two-sided markets.

problem Evaluating the effects of different policies in two-sided markets with spatial and temporal interference.
method Introduces a multi-agent reinforcement learning (MARL) framework to address policy evaluation challenges in large-scale fleet management.
result Proposes novel estimators for mean outcomes under different products that are consistent despite high-dimensionality.

Improves data efficiency in multi-agent control tasks using model-based reinforcement learning.

problem Limited data efficiency in reinforcement learning for multi-agent tasks.
method Decentralized model-based policy optimization (DMPO) framework.
result DMPO achieves superior data efficiency and matches model-free methods using true models.

Study examines active travel in Chicago communities, revealing mixed perceptions.

problem Transport disadvantage and lack of active mobility in underserved communities.
method Focus groups, qualitative discourse analysis, quantitative text-mining (topic modeling, sentiment analysis).
result Residents view active travel as both necessity and symbol of privilege, influenced by local culture.

A new algorithm for optimizing probability distributions converges linearly.

problem Optimizing functionals over families of probability distributions.
method Variational transport: particle-based algorithm approximating Wasserstein gradient descent.
result Variational transport converges linearly to the global minimum of the objective functional.

Study \ell_\infty bounds for MRP value function estimation.

problem Estimate MRP value function from samples.
method Analyze standard and robust plug-in approaches, establish bounds.
result Non-asymptotic and data-dependent \ell_\infty-norm bounds.

New approach uses deep reinforcement learning for vehicle dispatching, reducing waiting times.

problem Dynamic vehicle dispatching problem in various contexts.
method Event-based semi-Markov decision process with deep q-learning.
result Deep reinforcement learning policies outperform heuristic methods in New York City data.

Deep RL for dynamic pricing of express lanes considers multiple origins, destinations, and access locations.

problem Dynamic pricing of express lanes with multiple access points and traveler heterogeneity.
method Formulated as a POMDP, uses policy gradient methods and neural networks to determine stochastic tolls.
result Deep RL outperforms traditional methods in maximizing revenue and minimizing travel time.

This paper tackles efficient cooperative control for large-scale traffic signals using tensor-based deep learning.

problem Efficient training and control for large-scale multi-intersection traffic signals.
method Tensor representation, multi-task learning, imitation learning, proximal policy optimization.
result The proposed model achieves better performance compared to existing methods.

This article presents results from the first statistically significant study of causes of cost escalation in transport infrastructure projects. The study is based on a sample of 258 rail, bridge, tunnel and road projects worth US$90 billion. The focus is on the dependence of cost escalation on (1) length of project imp…

2013-04-16abs ↗pdf ↗

New methods estimate transport-growth pairs in unbalanced optimal transport.

problem Statistical guarantees for Monge-type estimation in unbalanced optimal transport remain limited.
method Developed two estimators for transport-growth pairs under different setups.
result Achieved minimax optimal rate for estimation of transport-growth pairs.

Novel approach learns optimal transport using convex neural networks.

problem Learning optimal transport between distributions from samples.
method Solving a minimax optimization to learn two convex functions, representing the optimal transport map.
result The approach finds optimal transport mappings that are independent of initialization and can handle discontinuous distributions.

Study shows how optimal transport behaves in higher dimensions.

problem Characterizing optimal transport in higher dimensions with Euclidean distance.
method Investigates the small regularization limit of entropic optimal transport.
result The limiting transport plan is supported on transport rays and uniquely minimizes a relative entropy functional.

Optimizes e-hailing drivers' passenger seeking to reduce congestion and pollution.

problem Reduces congestion and pollution by optimizing e-hailing drivers' passenger seeking.
method Uses Markov Decision Process (MDP) and imitation learning to model and optimize drivers' decisions.
result Achieves a 17.5% improvement in passenger return rate over a heuristic strategy.

Deep learning models optimize urban transportation scheduling.

problem Optimizing transportation systems with complex dynamics and large data sets.
method Developed deep learning metamodels for simulators and reinforcement learning algorithms.
result Improved optimal scheduling of travelers on transportation networks.

New algorithm solves unbalanced optimal transport on trees in quasi-linear time.

problem Efficiently solving unbalanced optimal transport problems on trees.
method Proposed an algorithm that solves a more general unbalanced optimal transport problem exactly in quasi-linear time on a tree metric.
result Solves unbalanced optimal transport on trees in quasi-linear time (less than one second for a tree with one million nodes).