Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,932 papers · 148 categories

Trend · papers per month

0.3%0.5%0.8%0.2% · Jun 201219922001200920172026
7 results for pre-clustering

Develops algorithms to exploit historical and pre-clustered arm information in bandit problems.

problem Optimizing decision-making in multi-armed bandit and contextual bandit problems with historical observations and pre-clustered arms.
method META algorithm that combines historical observations and pre-clustering information, deriving regret bounds for various scenarios.
result META algorithm effectively balances between using historical observations and clustering, outperforming the other in different scenarios.

A new method estimates treatment effects in mixed groups, improving accuracy.

problem Estimating treatment effects in mixed groups with heterogeneous responses.
method PCM (pre-cluster and merge) approach for nonparametric estimation.
result Asymptotic consistency and significant improvement in accuracy over existing methods.

This paper improves keyword recommendation for sponsored search using deep reinforcement learning.

problem Selecting keywords from given candidates considering internal and external competitions.
method Solves the combinatorial optimization problem of keyword recommendations with a modified pointer network structure trained in a deep reinforcement learning framework.
result Remarkable improvements in performance observed both offline and online.

New algorithms for clustering and synthetic data generation of heterogeneous tabular datasets.

problem Clustering and generating synthetic data from heterogeneous tabular datasets with hidden cluster structure.
method Developed MMM and MMMsynth algorithms for clustering and synthetic data generation.
result MMMsynth algorithm outperforms other literature tabular-data generators and approaches real data performance.