Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,742 papers · 148 categories

Trend · papers per month

2.3%4.6%7.0%9.3% · Nov 202419922001200920172026
48 results for tactical decisions

A reinforcement learning framework combining value function and tree search planner for strategic and tactical decisions.

problem Strategic and tactical decision-making in discrete environments.
method Combines value function and tree search planner, using uncertainty modeling and risk measurement.
result Improves performance and learning speed on hard exploration environments.

TacticAI helps football coaches improve tactics by analyzing corner kicks.

problem Developing effective responses to rival teams' tactics algorithmically.
method TacticAI combines predictive and generative components to suggest player setups and position adjustments.
result TacticAI's model suggestions are favored over existing tactics 90% of the time by football domain experts.

This article provides a novel framework to evaluate limit order tactics that highlights expected fill price, adverse price selection cost, and opportunity cost. We formulate the problem of optimal execution of market orders with nonlinear market impact, power law decay kernel, and stochastic and deterministic liquidity…

2014-09-04abs ↗pdf ↗

Paper improves asset allocation using machine learning for regime detection.

problem Improving asset allocation strategies in uncertain economic conditions.
method Machine learning for regime detection, modified k-means algorithm, portfolio optimization.
result Significant portfolio performance improvements over traditional benchmarks.

SoccerCPD detects tactical changes in soccer matches using spatiotemporal tracking data.

problem Detecting consistent team formations in fluid sports like soccer.
method Two-step change-point detection: formation and role changes.
result Accurately detects tactical changes and estimates formation and role assignments.

We consider an agent who needs to buy (or sell) a relatively small amount of asset over some fixed short time interval. We work at the highest frequency meaning that we wish to find the optimal tactic to execute our quantity using limit orders, market orders and cancellations. To solve the agent's control problem, we b…

2018-03-15abs ↗pdf ↗

Humans prove theorems by relying on substantial high-level reasoning and problem-specific insights. Proof assistants offer a formalism that resembles human mathematical reasoning, representing theorems in higher-order logic and proofs as high-level tactics. However, human experts have to construct proofs manually by en…

2019-05-21abs ↗pdf ↗

Technology offers new ways to measure the locations of the players and of the ball in sports. This translates to the trajectories the ball takes on the field as a result of the tactics the team applies. The challenge professionals in soccer are facing is to take the reverse path: given the trajectories of the ball is i…

2015-08-10abs ↗pdf ↗

In this paper, we introduce a system called GamePad that can be used to explore the application of machine learning methods to theorem proving in the Coq proof assistant. Interactive theorem provers such as Coq enable users to construct machine-checkable proofs in a step-by-step manner. Hence, they provide an opportuni…

2018-06-02abs ↗pdf ↗

We introduce two tactics to attack agents trained by deep reinforcement learning algorithms using adversarial examples, namely the strategically-timed attack and the enchanting attack. In the strategically-timed attack, the adversary aims at minimizing the agent's reward by only attacking the agent at a small subset of…

2017-03-08abs ↗pdf ↗

Machine learning research for developing countries can demonstrate clear sustainable impact by delivering actionable and timely information to in-country government organisations (GOs) and NGOs in response to their critical information requirements. We co-create products with UK and in-country commercial, GO and NGO pa…

2018-11-29abs ↗pdf ↗

This paper compares forecasting techniques for sales data, focusing on profit-driven models.

problem Choosing the best forecasting technique for sales data is challenging.
method Compares various forecasting methods including ML, statistics, and econometrics.
result Simple seasonal models consistently outperform other methodologies.

Simultaneously estimates travel times and route choice model parameters.

problem Interdependent estimation of arc travel times and route choice model parameters.
method Maximum likelihood estimation for any differentiable route choice model.
result Strong performance in real-world data, even compared to arc travel time estimation methods.

LLM forecasting benchmarks suffer from information leakage, which confounds model performance.

problem LLM forecasting benchmarks suffer from information leakage.
method A retrieval-augmented LLM forecaster observes only decision-time information.
result The full pipeline obtains a median monthly Spearman rank IC of +0.154.

Paper uses deep imitation learning to predict aircraft trajectories accurately.

problem Inefficient and costly Air Traffic Management system limits predictability.
method Generative Adversarial Imitation Learning framework with trajectory clustering and classification.
result Accurate predictions for entire trajectory stages, pre- and tactical.

Geometric Brownian motion (GBM) is a model for systems as varied as financial instruments and populations. The statistical properties of GBM are complicated by non-ergodicity, which can lead to ensemble averages exhibiting exponential growth while any individual trajectory collapses according to its time-average. A com…

2012-09-20abs ↗pdf ↗

Paper presents a new method for better financial market forecasting.

problem Traditional investment strategies fail to capture market nuances and risks.
method Combines deep learning, factor integration, and correlated stock analysis.
result Enhanced diversification and performance capture in financial markets.

Minimalistic attacks reveal deep RL policies' vulnerabilities with little perturbation.

problem Tackling the vulnerability of deep reinforcement learning policies to minimal perturbations.
method Three key settings: black-box policy access, fractional-state adversary, and tactically-chanced attack. Formulated adversarial attacks on six Atari games.
result Deep RL policies can be significantly fooled by minimal perturbations, even in 0.01% of the input state.

MAP-Elites generates diverse trading strategies for improved execution performance.

problem Optimizing trading execution schedules in volatile market conditions.
method Quality-diversity algorithm (MAP-Elites) generating a portfolio of specialized strategies.
result Diverse strategies achieve 8-10% performance improvements, validating quality-diversity methods.

Obtaining more accurate equity value estimates is the starting point for stock selection, value-based indexing in a noisy market, and beating benchmark indices through tactical style rotation. Unfortunately, discounted cash flow, method of comparables, and fundamental analysis typically yield discrepant valuation estim…

2007-07-24abs ↗pdf ↗

This paper demonstrates the use of genetic algorithms for evolving: 1) a grandmaster-level evaluation function, and 2) a search mechanism for a chess program, the parameter values of which are initialized randomly. The evaluation function of the program is evolved by learning from databases of (human) grandmaster games…

2017-11-21abs ↗pdf ↗

This study uses HMM and RL to dynamically allocate equities, Treasuries, and gold based on market regimes.

problem Developing a dynamic portfolio allocation strategy for different market conditions.
method Characterizes market regimes using Markov switching models and HMM, then applies RL for allocation decisions.
result RL-based allocation outperforms passive strategies, providing lower drawdowns and higher Sharpe ratios.

QTNet uses deep reinforcement learning to automate trading strategies.

problem Handling noisy and high-frequency financial data, balancing exploration and exploitation.
method QTNet employs deep reinforcement learning (DRL) with imitative learning to autonomously formulate trading strategies.
result QTNet demonstrates proficiency in extracting robust market features and adaptability to diverse conditions.

Many machine learning models have important structural tuning parameters that cannot be directly estimated from the data. The common tactic for setting these parameters is to use resampling methods, such as cross--validation or the bootstrap, to evaluate a candidate set of values and choose the best based on some pre--…

2014-05-27abs ↗pdf ↗

Optimizes trade execution with reinforcement learning for limit orders.

problem Maximizing revenue in a limit order book with market and limit orders.
method Formulated as a dynamic allocation task, uses multivariate logistic-normal distributions for efficient training.
result Outperforms traditional strategies in simulated environments.

AI-driven framework improves enterprise financial audits and risk identification.

problem Manual auditing is inefficient and limited by data complexity and evolving fraud tactics.
method Machine learning algorithms (SVM, RF, KNN) applied to a dataset of audit project counts, violations, and fraud instances.
result Random Forest achieves best performance with F1-score of 0.9012, identifying fraud and compliance anomalies.

Paper uses NMT to predict solutions to stochastic optimization problems quickly.

problem Predicting solutions to stochastic discrete optimization problems under uncertainty.
method Applied a state-of-the-art NMT algorithm with minimal adaptations and hyperparameter tuning.
result NMT can produce accurate solutions in milliseconds with less variability.

Framework analyzes physical metrics in soccer to link performance with value.

problem Understanding and quantifying the value of physical performance in soccer.
method Estimating physical indicators from tracking data, contextualizing runs, linking with possession-value model.
result Physical performance correlates with value generation in soccer.

New research shows existing information-theoretic methods can't establish minimax rates for gradient descent in stochastic convex optimization.

problem Establishing minimax rates for gradient descent in stochastic convex optimization using information-theoretic methods.
method Examined several information-theoretic frameworks including input-output mutual information bounds, conditional mutual information bounds, PAC-Bayes bounds, and their variants.
result Proved that none of the examined information-theoretic frameworks can establish minimax rates for gradient descent in stochastic convex optimization.

Communication networks have evolved from specialized, research and tactical transmission systems to large-scale and highly complex interconnections of intelligent devices, increasingly becoming more commercial, consumer-oriented, and heterogeneous. Propelled by emergent social networking services and high-definition st…

2012-11-29abs ↗pdf ↗