DG improves policy gradient efficiency by selectively backpropagating only valuable samples.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
DE is a new exploration method that limits resource usage based on expected improvement and surprise.
DG improves policy gradients by weighting actions with a sigmoid of advantage and surprisal.
DG separates successes and failures by gating updates with advantage and surprisal.
The Weil conjecture is a delightful theorem for algebraic varieties on finite fields and an important model for dynamical zeta functions. In this paper, we prove a functional equation of Lefschetz zeta functions for infinite cyclic coverings which is analogous to the Weil conjecture. Applying this functional equation t…
HedgeAgents boosts financial trading with balanced strategies.
For a company looking to provide delightful user experiences, it is of paramount importance to take care of any customer issues. This paper proposes COTA, a system to improve speed and reliability of customer support for end users through automated ticket classification and answers selection for support representatives…
RAMANMETRIX simplifies Raman spectroscopy data analysis.